MARATTO

article · Zenodo (CERN European Organization for Nuclear Research)

A New Learning Technique for Planning in Cooperative Multi-Agent Systems

Abstract

This presentation introduces a new learning technique for planning in cooperative multi-agent systems (MAS), proposing a taxonomy for MAS based on rationality and optimality, and formally defining cooperative, competitive, and mixed matrix games (MGs). It presents the Cooperative Multi-agent Markov Decision Process (CMMDP) as a mathematical framework and introduces the Extended-Q algorithm, which integrates reinforcement learning with game-theoretic equilibrium concepts like Nash equilibrium to solve coordination problems. The algorithm is extended to handle weakly competitive scenarios and is enhanced with neural network-based generalization (Neuro-Extended-Q) for large state spaces. Experimental validation using grid games demonstrates its effectiveness, while future work includes convergence proofs, extensions to competitive MAS, partial observability, and improved exploration techniques.

Research topics

  • Reinforcement Learning in Robotics
  • Adaptive Dynamic Programming Control
  • Game Theory and Applications

Sustainable Development Goals

Read the original research

This page summarises published work. The authoritative version sits with the publisher.

DOI: 10.5281/zenodo.18202514

Is something wrong with this record? Report it or request removal.

Discussion

Discuss this research

Have you built on this work, tried to replicate it, or seen it applied in practice? Share what you know. Verified researchers and MARATTO™ domain experts can open a discussion, and any member can reply. Contributions are reviewed before they appear.

No discussion yet. Open the first thread.