Adaptive Beam Hopping and Power Control for Dual-Layer Over-the-Air Online Federated Learning in LEO Satellite Networks

📅 2026-09-02
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
本文研究了低轨卫星网络中通过空中计算实现的在线联邦学习,采用双层空中聚合架构,并利用深度强化学习框架优化自适应波束跳变和功率控制以最大化数据利用率。
📝 Abstract
This paper investigates over-the-air (OTA) computation enabled online federated learning (FL) in low-Earth orbit (LEO) satellite networks. Specifically, we consider a dual-layer OTA aggregation architecture, where ground devices upload analog model updates to serving satellites via uplink OTA aggregation, and satellites forward the aggregated signals to a data processing center through the second round OTA aggregation. Then, we formulate a long-term data-utilization maximization problem in which devices continuously collect new data and untrained samples gradually lose freshness. The problem is subject to the satellite beam budget, transmit-power limit, and global mean squared error (MSE) constraint that governs end-to-end aggregation distortion. This yields a coupled mixed-integer nonlinear programming (MINLP) problem, involving tightly coupled discrete beam-hopping decisions and continuous power control. Due to the combinatorial action space and nonconvex constraints, the problem is NP-hard and computationally intractable. Furthermore, the time-varying satellite topology and dynamic data generation render it a sequential decision-making problem, necessitating adaptive online scheduling. To address these issues, we cast the problem as a Markov decision process and develop a proximal policy optimization (PPO)-based deep reinforcement learning framework that jointly optimizes adaptive beam hopping and power control, using an MSE-aware reward to balance data utilization and aggregation accuracy. Numerical simulation results verify that the proposed algorithm consistently outperforms other benchmark schemes, achieving superior long-term data utilization and faster FL convergence while satisfying the MSE requirement.
Problem

Research questions and friction points this paper is trying to address.

over-the-air computation
online federated learning
LEO satellite networks
data-utilization maximization
global mean squared error
Innovation

Methods, ideas, or system contributions that make the work stand out.

Adaptive Beam Hopping
Power Control
Dual-Layer OTA Aggregation
PPO-based Deep Reinforcement Learning
MSE-aware Reward
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
Zhendong Li
Zhendong Li
Beijing Normal University
Theoretical and Computational Chemistry
S
Shaojie Wang
School of Information and Communication Engineering, Xi’an Jiaotong University, Xi’an 710049, China
Zhou Su
Zhou Su
Xi'an Jiaotong University
Zihao Zhang
Zihao Zhang
天津大学
计算机视觉
Haixia Peng
Haixia Peng
Professor, School of Information and Communications Engineering, Xi'an Jiaotong University
Resource managementReinforcement learningMulti-access edge computingSDNVehicular networks
Nan Cheng
Nan Cheng
University of Michigan
condensed matter physics
Y
Ying Wang
State Key Laboratory of Networking and Switching Technology, Beijing University of Posts and Telecommunications, Beijing 100876, China
W
Wen Chen
Department of Electronic Engineering, Shanghai Jiao Tong University, Shanghai 200240, China