Abstract:This paper introduces a full solution for decentralized routing in Low Earth Orbit satellite constellations based on continual Deep Reinforcement Learning (DRL). This requires addressing multiple challenges, including the partial knowledge at the satellites and their continuous movement, and the time-varying sources of uncertainty in the system, such as traffic, communication links, or communication buffers. We follow a multi-agent approach, where each satellite acts as an independent decision-making agent, while acquiring a limited knowledge of the environment based on the feedback received from the nearby agents. The solution is divided into two phases. First, an offline learning phase relies on decentralized decisions and a global Deep Neural Network (DNN) trained with global experiences. Then, the online phase with local, on-board, and pre-trained DNNs requires continual learning to evolve with the environment, which can be done in two different ways: (1) Model anticipation, where the predictable conditions of the constellation are exploited by each satellite sharing local model with the next satellite; and (2) Federated Learning (FL), where each agent's model is merged first at the cluster level and then aggregated in a global Parameter Server. The results show that, without high congestion, the proposed Multi-Agent DRL framework achieves the same E2E performance as a shortest-path solution, but the latter assumes intensive communication overhead for real-time network-wise knowledge of the system at a centralized node, whereas ours only requires limited feedback exchange among first neighbour satellites. Importantly, our solution adapts well to congestion conditions and exploits less loaded paths. Moreover, the divergence of models over time is easily tackled by the synergy between anticipation, applied in short-term alignment, and FL, utilized for long-term alignment.

Shaping Rewards, Shaping Routes: On Multi-Agent Deep Q-Networks for Routing in Satellite Constellation Networks

Multi-Agent Deep Reinforcement Learning for Distributed Satellite Routing

Stigmergy and Hierarchical Learning for Routing Optimization in Multi-domain Collaborative Satellite Networks

Reinforcement learning based dynamic distributed routing scheme for mega LEO satellite networks

Heterogeneous Satellite Network Routing Algorithm Based on Reinforcement Learning and Mobile Agent.

Continual Deep Reinforcement Learning for Decentralized Satellite Routing

Q-learning for distributed routing in LEO satellite constellations

Fully-Distributed Dynamic Packet Routing for LEO Satellite Networks: A GNN-Enhanced Multi-Agent Reinforcement Learning Approach

Distributed Routing Algorithm for LEO Satellite Network Based on Deep Reinforcement Learning

Enabling High-Throughput Routing for LEO Satellite Broadband Networks: A Flow-Centric Deep Reinforcement Learning Approach

An Intelligent Routing Algorithm for LEO Satellites Based on Deep Reinforcement Learning

Traffic Optimization in Satellites Communications: A Multi-agent Reinforcement Learning Approach

Satellite Network Routing Planning Method Based on Reinforcement Learning

Multi-Commodity Flow Routing for Large-Scale LEO Satellite Networks Using Deep Reinforcement Learning

A robust routing strategy based on deep reinforcement learning for mega satellite constellations

A Topology Design Method for Satellite Networks Based on Deep Reinforcement Learning

Dynamic Routing for Integrated Satellite-Terrestrial Networks: A Constrained Multi-Agent Reinforcement Learning Approach

Dynamic Routing Planning Method for Large-Scale Low-Orbit Satellite Networks Based on Location Guided and Multi-Agent DQN Network

Research on Routing Optimization in Satellite Internet Based on Deep Reinforcement Learning

Graph Neural Network and Reinforcement Learning Based Routing for Mega LEO Satellite Constellations

Trustworthy and Load-Balancing Routing Scheme for Satellite Services with Multi-Agent DRL.