y0news
AnalyticsDigestsSourcesTopicsRSSAICrypto

#edge-computing News & Analysis

214 articles tagged with #edge-computing. AI-curated summaries with sentiment analysis and key takeaways from 50+ sources.

214 articles
AINeutralarXiv – CS AI · Jun 96/10
🧠

Collaborative Edge-to-Server Inference for Vision-Language Models

Researchers propose a collaborative edge-to-server inference framework for vision-language models that reduces communication costs by selectively transmitting only high-entropy regions of interest rather than full-resolution images. The two-stage approach maintains inference accuracy while substantially decreasing bandwidth requirements across visual question-answering tasks.

AINeutralarXiv – CS AI · Jun 96/10
🧠

Efficient Skill Grounding via Code Refactoring with Small Language Models

Researchers introduce RECENT, a framework that enables small language models to effectively ground robot skills through code refactoring rather than full regeneration. By decoupling skill semantics from embodiment-specific details, the approach matches LLM-based performance while remaining practical for resource-constrained embodied agents.

AINeutralarXiv – CS AI · Jun 86/10
🧠

Towards Efficient and Exact Forgetting Services in Pre-Trained-Model-based Continual Learning

Researchers propose Analytic Continual Unlearning (ACU), a gradient-free method enabling efficient removal of specific knowledge from pre-trained models during continuous learning phases while preserving privacy. The approach uses closed-form solutions to handle sequential forgetting requests, addressing gaps in existing unlearning techniques that struggle with privacy violations and adversarial request patterns.

AINeutralThe Verge – AI · Jun 56/10
🧠

This is your laptop… on AI

Major tech companies including Nvidia, Microsoft, and Google are pushing AI-integrated laptops as the next computing paradigm, with Nvidia's Jensen Huang unveiling new hardware designed specifically for on-device AI workloads. However, the article raises a critical question about market demand: whether consumers and enterprises actually want these AI-centric devices or if vendors are creating solutions in search of problems.

This is your laptop… on AI
🏢 Nvidia🧠 Gemini
AIBullishBlockonomi · Jun 56/10
🧠

Hitachi (6501.T) Stock Surges on Intel Collaboration in Industrial AI Sector

Hitachi's stock rose 2.26% to ¥5,300 following the announcement of a strategic partnership with Intel centered on industrial AI and advanced computing infrastructure. The collaboration signals growing corporate investment in AI-driven manufacturing and enterprise solutions.

AINeutralarXiv – CS AI · Jun 55/10
🧠

Dimensionality Reduction for Cyberattack Classification: A Comparative Evaluation of PCA and Linear Predictive Coding

Researchers compare Principal Component Analysis (PCA) and Linear Predictive Coding (LPC) for reducing feature dimensionality in cyberattack detection systems. The study demonstrates that aggressive compression of high-dimensional data maintains classification accuracy while significantly reducing computational overhead, enabling deployment in resource-constrained environments.

AINeutralarXiv – CS AI · Jun 56/10
🧠

Cognitive Threat Intelligence and Explainable Federated Security Analytics for distributed Infrastructure Systems

Researchers propose a Cognitive Threat Intelligence framework combining Federated Learning and Explainable AI to detect cyber threats across distributed infrastructure systems while preserving data privacy. The approach eliminates the need to transmit sensitive network traffic to centralized servers, instead training models locally and sharing only encrypted parameters.

AINeutralarXiv – CS AI · Jun 56/10
🧠

DisasterBench: A Multimodal Benchmark for UAV-Based Disaster Response in Complex Environments

Researchers introduced DisasterBench, a multimodal AI benchmark designed to improve UAV-based disaster response by testing reasoning across 14 disaster types and 9 response-critical tasks. They also developed DisasterVL, a lightweight 2B-parameter model that achieves GPT-4o-level reasoning accuracy while operating efficiently on edge devices with limited computational resources.

🧠 GPT-4
AINeutralarXiv – CS AI · Jun 56/10
🧠

Surrogate Neural Architecture Codesign Package (SNAC-Pack)

SNAC-Pack is an open-source AutoML framework that automates neural architecture design for FPGA deployment by combining hardware-aware search with quantization and pruning. The tool reduces design cycles from months to hours while matching or exceeding baseline performance on tasks like jet classification and quantum computing applications.

AINeutralarXiv – CS AI · Jun 45/10
🧠

Gravity-Aware Hierarchical Routing for Lightweight SensorLLM on Human Activity Recognition

Researchers propose a gravity-aware hierarchical routing method to improve human activity recognition in compressed language models used with wearable sensors. The lightweight adaptation addresses a specific failure mode where static activities like standing and sitting are poorly recognized when using compact models like TinyLlama, while maintaining strong performance on dynamic activities.

AIBullisharXiv – CS AI · Jun 46/10
🧠

LLM Compression with Jointly Optimizing Architectural and Quantization choices

Researchers introduce a differentiable Neural Architecture Search framework that jointly optimizes LLM architecture and mixed-precision quantization, achieving 1.4x faster inference speeds or 6% higher accuracy compared to sequential optimization approaches. This compression technique addresses the critical challenge of deploying large language models on edge devices without requiring extensive GPU training.

AIBullisharXiv – CS AI · Jun 46/10
🧠

TITAN-FedAnil+: Trust-Based Adaptive Blockchain Federated Learning for Resource-Constrained Intelligent Enterprises

TITAN-FedAnil+ presents a blockchain-based federated learning framework designed to address data privacy and security challenges in resource-constrained enterprise environments. The system uses adaptive clustering and GPU acceleration to filter malicious updates while reducing memory overhead by up to 81%, making secure distributed learning more practical for edge devices.

AINeutralarXiv – CS AI · Jun 46/10
🧠

Multi-SPIN: Multi-Access Speculative Inference for Cooperative Token Generation at the Edge

Researchers propose Multi-SPIN, a distributed speculative inference architecture that enables edge servers and resource-constrained devices to collaboratively generate language model tokens. The system optimizes draft-length control and bandwidth allocation to maximize throughput, achieving up to 88% goodput improvement over baseline methods in real-world testing.

🧠 Llama
AIBullisharXiv – CS AI · Jun 46/10
🧠

Toward Autonomous O-RAN: A Multi-Scale Agentic AI Framework for Real-Time Network Control and Management

Researchers propose a multi-scale agentic AI framework for Open Radio Access Networks (O-RAN) that uses hierarchical AI agents—from Large Language Models to wireless foundation models—to autonomously manage 6G network control across different timescales. The framework addresses operational complexity in disaggregated networks by enabling coordinated AI decision-making across standardized interfaces, demonstrated through proof-of-concept scenarios.

AIBullishCrypto Briefing · Jun 36/10
🧠

Nvidia unveils RTX Spark, advancing AI integration in Windows PCs

Nvidia has unveiled RTX Spark, a technology designed to enhance local AI capabilities on Windows PCs. The innovation promises to strengthen security through on-device processing while creating new commercial opportunities for technology companies.

Nvidia unveils RTX Spark, advancing AI integration in Windows PCs
🏢 Nvidia
AINeutralDecrypt – AI · Jun 36/10
🧠

Perplexity Wants Your Laptop to Do Part of the AI Work—So It Doesn't Have To

Perplexity has introduced a hybrid inference system that distributes AI computational tasks between user devices and cloud servers automatically. The approach aims to reduce server costs, improve privacy, and lower latency by leveraging local processing power where feasible.

Perplexity Wants Your Laptop to Do Part of the AI Work—So It Doesn't Have To
🏢 Perplexity
AIBullishOpenAI News · Jun 36/10
🧠

How Wasmer used Codex to build a Node.js runtime for the edge

Wasmer leveraged OpenAI's Codex (GPT-5.5) to accelerate development of a Node.js runtime for edge computing, reducing typical development timelines from months to weeks while achieving a 10x-20x productivity multiplier. This demonstrates how AI-assisted coding tools can substantially compress software engineering cycles for complex infrastructure projects.

🧠 GPT-5
AINeutralarXiv – CS AI · Jun 36/10
🧠

Toward a Modular Architecture for Embedded AI Agent Systems at the Edge

Researchers propose a modular reference architecture for deploying AI agents on resource-constrained embedded devices, combining on-device compressed neural networks with cloud-based small language models. The framework introduces a governance layer for safety and observability across distributed autonomous systems, addressing the gap between real-time control and agentic reasoning in edge computing environments.

AIBullishHugging Face Blog · Jun 26/10
🧠

Holo3.1: Fast & Local Computer Use Agents

Holo3.1 represents an advancement in local, fast computer-use AI agents that operate without requiring constant cloud connectivity. This development enables more efficient, privacy-preserving autonomous agents for developers and enterprises seeking decentralized AI infrastructure.

AINeutralarXiv – CS AI · Jun 26/10
🧠

CoMIC: Collaborative Memory and Insights Circulation for Long-Horizon LLM Agents in Cloud-Edge Systems

Researchers introduce CoMIC, a cloud-edge framework that enables lightweight LLM agents on edge servers to handle long-horizon tasks by combining local execution with centralized cloud-based reflection and experience aggregation. The parameter-update-free approach improves performance across symbolic planning and text interaction tasks without requiring model fine-tuning.

AINeutralarXiv – CS AI · Jun 26/10
🧠

LP5X-PIM Sim: A High-Fidelity HW/SW Integrated Simulator for LPDDR5X-PIM

Samsung Electronics has developed LP5X-PIM Sim, a high-fidelity hardware-software integrated simulator for LPDDR5X-PIM technology that models both data paths and control layers. The simulator enables precise evaluation of system performance and energy efficiency while optimizing processing-in-memory resource utilization, representing an advancement in memory architecture simulation for emerging computing paradigms.

AIBullisharXiv – CS AI · Jun 26/10
🧠

A Lightweight Context-Driven Training-Free Network for Scene Text Segmentation and Recognition

Researchers propose a training-free, lightweight framework for scene text recognition that leverages pre-trained models and context-driven understanding to achieve state-of-the-art performance with significantly reduced computational requirements. The approach uses attention-based segmentation and semantic evaluation to enable faster inference suitable for real-time deployment scenarios.

AIBullisharXiv – CS AI · Jun 26/10
🧠

A Communication-Centric 6G-LLM Architecture for Scalable Tactical Autonomous Defense Vehicle Networks

Researchers propose a 6G-LLM architecture for coordinating autonomous defense vehicle networks that combines edge-based large language models with semantic communication. Simulations show the system achieves 75% latency reduction and 83% mission success rates at 30-vehicle scale compared to 5G baselines, suggesting significant operational advantages for military autonomous systems.

← PrevPage 5 of 9Next →