Skip to content

Large AI Models Empowered Edge Intelligence for Next-Gen Consumer Electronics

Sep 2026 · IEEE Consumer Electronics Magazine · Vol 15, pp. 10-16 · 0 citations · 20 references

Abstract

This article addresses challenges in the Internet of Consumer Electronics (ICE), such as random task arrivals, limited resources, and system stability, by proposing a collaborative computing framework that integrates edge intelligence with Lyapunov-based deep reinforcement learning (DRL). The framework adopts a three-tier architecture. 1) The application layer generates multiple types of tasks; 2) the intelligent decision-making layer incorporates large artificial intelligence (AI) models to extract global features and employs Lyapunov optimization to transform long-term stochastic problems into deterministic optimization while utilizing an actor-critic DRL architecture for resource allocation; and 3) the resource layer integrates distributed edge nodes to form a unified resource pool. Experiments demonstrate that the framework achieves efficient, stable, and scalable intelligent services on the edge.

View source

Similar papers

Review Open access Jul 2026

Towards Intelligent 6G Networks: A Comprehensive Review of AI-Driven Control and Optimization

The findings indicate that while AI techniques substantially improve network adaptability, resource management, and autonomous operation, significant challenges remain regarding scalability, computational complexity, data dependency, interoperability, explainability, and deployment in real-world environments.

Ali Ahmed Mirza, T. Mahmood, E. Dhulkefl · 0 citations
Preprint Jul 2026

Self-organizing Architecture of Receptron Units: a Hardware-Aware Framework for Edge Intelligence

The growing demand for intelligent processing at the edge of IoT networks is constrained by the severe computational and memory limitations of microcontroller units, which render impractical conventional deep learning approaches. We propose a neuromorphicinspired classifier based on the Receptron model, a single-unit architecture capable of implementing non-linearly separable decision boundaries, without resorting to multi-layer networks. The model is designed for direct deployment on mid-range MCUs, while supporting continuous on-device adaptation. Experimental evaluation on basic dataset benchmarks yields cross-validated accuracies compatible with standard machine learning method baselines. These results position the Receptron as a viable and interpretable alternative for resource-constrained neuromorphic edge systems operating in dynamic, non-stationary environments.

Stefano Radice, Ludovico Casaccia, Riccaro Emanuele Beccalli et al. · 0 citations
Review Aug 2026

Deep Reinforcement Learning for 6G AI-RAN: A Comprehensive Survey

This article presents the first dedicated and comprehensive survey of DRL for Open AI-RAN, and reviews the foundations of model-free, model-based, offline, safe, multi-agent, federated, and transfer learning, and provides an O-RAN-aware framework for formulating RAN control problems through states, observations, actions, rewards, constraints, and temporal structure.

Jie Lu, Peihao Yan, Qijun Wang et al. · 0 citations
Open access Jul 2026

A unified machine learning framework for intelligent resource allocation toward 6G wireless communications.

A Dual-Stage Multi-Time-Scale Temporal Attention-Based LSTM network (D-MTSTA-LSTM) has been architected, which effectively learns short- and long-term relationships in network trends, thereby precisely predicting optimal communication routes and associated power and spectrum allocation.

Nishu Gupta, Rupali Bhartiya, S. Rathod et al. · 0 citations
Conference Jul 2026

A Generative Artificial Intelligence–based ANFIS Approach for Low-Latency Video Communication in Cloud Environments

Generative AI (GAI) refers to advanced models that can generate new, realistic, context-aware data or solutions by learning from existing datasets, making them extremely valuable in adaptive intelligent systems. Traditional AI technologies often face limitations, including a lack of interpretability, sensitivity to noisy data, and an inability to generalize to dynamic environments. These limitations can be effectively addressed by GAI-driven adaptive neuro-fuzzy inference systems (ANFIS). To facilitate the ease of scalable processing and real-time deployment, the proposed framework is developed in a cloud computing system, whereby the aggregation of large-scale QoE data, the training of generative models and the optimization of the fuzzy rules are handled using the cloud computing resources, and latencysensitive inference is done effectively at the edge. To address these challenges, we propose a Generative Reinforcement Learning (GRL) method that enhances decision-making skills by generating artificial experiences, utilizing the video transmission system as an environment, and learns policies to optimize latency. The Generative Diffusion Model (GDM) for feature denoising offers a powerful mechanism for removing noise in high- dimensional data, thereby enhancing the accuracy and stability of prediction tasks. Physics-Guided Generative Modeling (PG2M) combines domain-specific physics laws with AI learning to ensure scientifically consistent and interpretable outputs. Finally, Generative Adversarial Networks (GANs) are employed for generative rule evolution, combining evolutionary computation with adversarial learning to dynamically generate and improve classification rules. ANFIS utilizes fuzzy rules to model delay shapes, while GRL optimizes video transmission strategies based on the output of ANFIS. Reduce latency dynamically by learning adaptive frame scheduling; it improves network responsiveness to fluctuations. The presented method achieved 94% accuracy and improved the latency of the video communication service.

Ashis Kumar Mohapatra · 0 citations

Related blog posts

Microsoft Research Blog Aug 31, 2026

GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models

What if pathology foundation models could do more with less? GigaPath-Flash and GigaTIME-Flash cut computational demands while maintaining strong performance, opening the door to larger studies and broader exploration. The post GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models appeared first on Microsoft Research.

MIT News · Artificial Intelligence Aug 27, 2026

Looking beyond natural sequences

A new machine-learning framework aims to improve the success rate of computational protein design while moving away from results that reproduce sequences found in nature.