Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
102 tokens/sec
GPT-4o
59 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
50 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Recursive Reasoning Graph for Multi-Agent Reinforcement Learning (2203.02844v1)

Published 6 Mar 2022 in cs.LG, cs.AI, and cs.MA

Abstract: Multi-agent reinforcement learning (MARL) provides an efficient way for simultaneously learning policies for multiple agents interacting with each other. However, in scenarios requiring complex interactions, existing algorithms can suffer from an inability to accurately anticipate the influence of self-actions on other agents. Incorporating an ability to reason about other agents' potential responses can allow an agent to formulate more effective strategies. This paper adopts a recursive reasoning model in a centralized-training-decentralized-execution framework to help learning agents better cooperate with or compete against others. The proposed algorithm, referred to as the Recursive Reasoning Graph (R2G), shows state-of-the-art performance on multiple multi-agent particle and robotics games.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (5)
  1. Xiaobai Ma (6 papers)
  2. David Isele (38 papers)
  3. Jayesh K. Gupta (25 papers)
  4. Kikuo Fujimura (22 papers)
  5. Mykel J. Kochenderfer (215 papers)
Citations (3)