Article
Multi-Agent Reinforcement Learning with Two Layer Control Plane for Traffic Engineering
2025-08-12
Abstract excerpt
The article presents a new method for multi-agent traffic flow balancing. It is based on the MAROH multi-agent optimization method. However, unlike MAROH, the agent's control plane is built on principles of human decision-making and consists of two layers. The first layer ensures autonomous decision-making by the agent based on accumulated experience—representatives of states the agent has encountered and kno...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- e8a68e56-776c-57c7-9641-ee95c8768991
- DOI
- 10.20944/preprints202508.0818.v1
