Back to search

Article

Multi-Agent Reinforcement Learning with Two Layer Control Plane for Traffic Engineering

2025-08-12

Abstract excerpt

The article presents a new method for multi-agent traffic flow balancing. It is based on the MAROH multi-agent optimization method. However, unlike MAROH, the agent's control plane is built on principles of human decision-making and consists of two layers. The first layer ensures autonomous decision-making by the agent based on accumulated experience—representatives of states the agent has encountered and kno...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
e8a68e56-776c-57c7-9641-ee95c8768991
DOI
10.20944/preprints202508.0818.v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Multi-Agent Reinforcement Learning with Two Layer Control Plane for Traffic EngineeringDOI 10.20944/preprints202508.0818.v1
Select a neighboring publication to make it the new centre.