RL-based Network Slice Embedding over Space Division Multiplexed Elastic Optical Networks
Divya Khanure, Riti Gour†, Congzhou Li, Jason P. Jue
Abstract
Network slicing over space-division-multiplexed elastic optical networks (SDM-EONs) requires jointly managing spectrum, spatial cores, and compute resources, a coupling that many existing studies ignore by treating compute placement independently from routing and spectrum decisions. This disconnect can cause the spectrum to be allocated along a path, only for the request to fail due to insufficient compute resources along the path, or may result in compute resources being allocated without consideration for spectrum resource availability on the path between compute nodes. We propose a path-constrained reinforcement learning framework that addresses compute node selection and RMCSA, being aware of both resources, restricting the RL agent's action space to nodes along k-shortest paths between request endpoints. Training incorporates reward shaping to improve robustness under high load. We propose PPO-Full (Proximal Policy Optimization-Full), which jointly selects compute nodes and routing paths via a multi-dimensional action space, against distance-based heuristics, a greedy baseline, and a decoupled VONE-DRL baseline on a 24-node USNET topology under hotspot traffic conditions. Results demonstrate consistent improvements in acceptance rate over all baselines at high load, with gains becoming more pronounced as traffic intensity increases.
Create a lesson
Related papers
Predictive Traffic Shaping as a UE Network Control Loop in Wireless Systems
Shriram Vasudevan, Subramanian Vasudevan
Matched-View Cross-Domain Evaluation of WireGuard VPN Traffic Classification Using Early-Flow Fingerprints
Yasameen Sajid Razooqi, Adrian Pekar
RadioSight: Predictive mmWave XR Network Optimization from Dynamic Neural Radio Fields
Lihao Zhang, Paul Kudyba, Zhenlin An et al.
Who Resolves Your DNS? Measuring Resolver Opacity and Closing the Visibility Gap
Kedar Thiagarajan, Fabian E. Bustamante
Sense Once, Serve Many: Common-Trace Factorized Constrained PPO for Online Sensing-Session Consolidation in Multi-Tenant ISAC Networks
Dang-Dung Vu
A-MADiff: Attention-Guided Multi-Agent DRL with Diffusion Policies for Memory-Aware Task Orchestration in Mobile AIGC Networks
Chongzhi Wu, Zhengtao Li, Jiawen Kang et al.