This is the pre-proceedings for the RLC 2026. You may expect minor changes.

An Agent-Centric Dynamical Systems Perspective on Multi-Agent Reinforcement Learning

By James Rudd-Jones, Maria Perez-Ortiz, and Mirco Musolesi

Reinforcement Learning Journal, vol. 7, 2026, pp. TBD.

Will be presented at the Reinforcement Learning Conference (RLC), MontrĂ©al, Quebec, Canada, August 15–17, 2026.


Download:

Keywords: Multi-Agent Reinforcement Learning, Dynamical Systems, Agent-Centric Behaviour,

Abstract:

Analysing learning in Multi-Agent Reinforcement Learning (MARL) environments is challenging, in particular with respect to individual decision-making. Practitioners frequently struggle to compare training runs due to the inherent stochasticity in algorithms arising from random dithering exploration, environment transition noise, and stochastic gradient updates to name a few. Traditional analytical approaches, such as replicator dynamics, oft rely on mean-field approximations to remove stochastic effects, but this simplification, whilst able to provide general overall trends, can lead to dissonance between analytical predictions and actual agent realisations. We propose modelling MARL training as a coupled stochastic dynamical systems, capturing both agent interactions and environmental characteristics. Leveraging tools from dynamical systems theory, we pragmatically analyse the stability and sensitivity of agent behaviour, which are key dimensions for their practical deployments, for example, in presence of strict safety requirements. This framework allows us to rigorously study the inherent stochasticity of MARL, providing a deeper understanding of system behaviour.


Citation Information:

James Rudd-Jones, Maria Perez-Ortiz, and Mirco Musolesi. "An Agent-Centric Dynamical Systems Perspective on Multi-Agent Reinforcement Learning." Reinforcement Learning Journal, vol. 7, 2026, pp. TBD.

BibTeX:
@article{ruddjones2026an,
    title={An Agent-Centric Dynamical Systems Perspective on Multi-Agent Reinforcement Learning},
    author={James Rudd-Jones and Maria Perez-Ortiz and Mirco Musolesi},
    journal={Reinforcement Learning Journal},
    volume={7},
    pages={},
    year={2026}
}