Pith. sign in

REVIEW 7 cited by

Multi-agent Reinforcement Learning: A Comprehensive Survey

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2312.10256 v2 pith:LM5TSEEX submitted 2023-12-15 cs.MA cs.AIcs.LG

classification cs.MAcs.AIcs.LG
keywords marlapplicationschallengeslearningmulti-agentsurveyagentscomprehensive
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Multi-agent systems (MAS) are widely prevalent and crucially important in numerous real-world applications, where multiple agents must make decisions to achieve their objectives in a shared environment. Despite their ubiquity, the development of intelligent decision-making agents in MAS poses several open challenges to their effective implementation. This survey examines these challenges, placing an emphasis on studying seminal concepts from game theory (GT) and machine learning (ML) and connecting them to recent advancements in multi-agent reinforcement learning (MARL), i.e. the research of data-driven decision-making within MAS. Therefore, the objective of this survey is to provide a comprehensive perspective along the various dimensions of MARL, shedding light on the unique opportunities that are presented in MARL applications while highlighting the inherent challenges that accompany this potential. Therefore, we hope that our work will not only contribute to the field by analyzing the current landscape of MARL but also motivate future directions with insights for deeper integration of concepts from related domains of GT and ML. With this in mind, this work delves into a detailed exploration of recent and past efforts of MARL and its related fields and describes prior solutions that were proposed and their limitations, as well as their applications.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 7 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 9 citations worldwide. Full citation record

  1. Subjective-Graph LLM Agents for Simulating Uncertainty in Classroom Social Perception

    cs.AI 2026-03 conditional novelty 6.0 of 10

    Subjective-graph LLM agents on 12 real classrooms accumulate collective ranking error from 0.066 to 0.124 over six exams despite repeated score anchors.

  2. Hierarchical Message-Passing Policies for Multi-Agent Reinforcement Learning

    cs.LG 2025-07 conditional novelty 6.0 of 10

    A feudal hierarchical MARL method where lower-level policies are rewarded with the upper level's advantage function, with theoretical alignment guarantees and strong benchmark results.

  3. LLM-Powered Decentralized Generative Agents with Adaptive Hierarchical Knowledge Graph for Cooperative Planning

    cs.AI 2025-02 conditional novelty 6.0 of 10

    DAMCS combines an adaptive knowledge-graph memory with structured communication to let LLM agents cooperate in an open-world game, cutting the steps needed to reach a diamond goal.

  4. TrustChain-Review: A Risk-Adaptive Blockchain and Game-Theoretic Framework for Trustworthy AI-Assisted Code Review

    cs.SE 2026-07 conditional novelty 5.0 of 10

    TrustChain-Review's simulation shows full-evidence code-review governance maximizes trust and detection but is costliest; risk-adaptive switching cuts cost ~38% and boosts cost-efficiency ~71% with lower detection.

  5. AgentGroupChat-V2: Divide-and-Conquer Is What LLM-Based Multi-Agent System Need

    cs.CL 2025-06 reject novelty 4.0 of 10

    A divide-and-conquer multi-agent framework with task forests and specialized roles improves math and code benchmarks but not commonsense or domain QA, and the adaptive heterogeneous-LLM engine is never tested.

  6. Multi-Agent Reinforcement Learning in Cybersecurity: From Fundamentals to Applications

    cs.MA 2025-05 conditional novelty 3.0 of 10

    A narrative survey of multi-agent reinforcement learning for cyber defense, reviewing game-theoretic models, cyber gyms, and applications, concluding MARL is promising but faces scalability and simulation-to-real tran...

  7. Multi-Agent Reinforcement Learning in Wireless Distributed Networks for 6G

    cs.IT 2025-02 conditional novelty 1.0 of 10

    A comprehensive survey of multi-agent reinforcement learning for wireless distributed networks in 6G, covering structures, algorithms, enhanced techniques, and applications.

Pith tools