A-MADiff: Attention-Guided Multi-Agent DRL with Diffusion Policies for Memory-Aware Task Orchestration in Mobile AIGC Networks
Chongzhi Wu, Zhengtao Li, Jiawen Kang, Jinbo Wen, Xiaohuan Li, Maomao Zhang, Ekram Hossain
Abstract
Artificial Intelligence-Generated Content (AIGC) services employ Generative AI (GenAI) models to automatically generate diverse content. Mobile AIGC networks host GenAI models on edge-located AIGC Service Providers (ASPs) to deliver low-latency and personalized AIGC services for mobile users. However, AIGC inference tasks typically occupy GPU memory until task completion, causing GPU memory exhaustion at serving ASPs and triggering out-of-memory failures rather than merely increasing service latency. Existing studies on AIGC task orchestration have largely overlooked GPU memory feasibility constraints. To address this issue, we develop a cooperative multi-agent orchestration framework, in which each edge node is equipped with a scheduling agent to route tasks to local ASPs or neighboring edge nodes. Since scheduling agents make decisions based only on local observations, while peer offloading couples their resource states and long-term utilities, we formulate the orchestration process as a cooperative Decentralized Partially Observable Markov Decision Process (Dec-POMDP). To solve the Dec-POMDP, we propose an Attention-guided Multi-Agent deep reinforcement learning algorithm with Diffusion policies (A-MADiff) under the centralized training with a decentralized execution paradigm. A-MADiff employs diffusion-based decentralized actors to generate multi-modal preferences over feasible orchestration actions, and an attention-guided centralized critic to estimate per-agent values from cross-agent states under GPU memory heterogeneity. Numerical results demonstrate that A-MADiff significantly improves the cumulative reward over the state-of-the-art baseline.
Create a lesson
Related papers
WiSDoM: Wireless Sparse Decision Transformer with Mixture-of-Experts for Multi-Task Mobile Network Optimization
Fatih Temiz, Shavbo Salehi, Melike Erol-Kantarci
Ray Tracing-Based LoRaWAN Gateway Placement for Reliable Connectivity in Amazonian Regions
Cláudio Modesto, Lucas Mozart, Cleverson Nahum et al.
Predictive Traffic Shaping as a UE Network Control Loop in Wireless Systems
Shriram Vasudevan, Subramanian Vasudevan
Matched-View Cross-Domain Evaluation of WireGuard VPN Traffic Classification Using Early-Flow Fingerprints
Yasameen Sajid Razooqi, Adrian Pekar
RadioSight: Predictive mmWave XR Network Optimization from Dynamic Neural Radio Fields
Lihao Zhang, Paul Kudyba, Zhenlin An et al.
RL-based Network Slice Embedding over Space Division Multiplexed Elastic Optical Networks
Divya Khanure, Riti Gour†, Congzhou Li et al.