Pith. sign in

REVIEW 1 cited by

Scalable Multi-Agent Reinforcement Learning through Intelligent Information Aggregation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2211.02127 v3 pith:Y45UKSPP submitted 2022-11-03 cs.MA cs.AIcs.RO

classification cs.MAcs.AIcs.RO
keywords agentsinformarlinformationlocalmulti-agentagentenvironmentsgoals
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We consider the problem of multi-agent navigation and collision avoidance when observations are limited to the local neighborhood of each agent. We propose InforMARL, a novel architecture for multi-agent reinforcement learning (MARL) which uses local information intelligently to compute paths for all the agents in a decentralized manner. Specifically, InforMARL aggregates information about the local neighborhood of agents for both the actor and the critic using a graph neural network and can be used in conjunction with any standard MARL algorithm. We show that (1) in training, InforMARL has better sample efficiency and performance than baseline approaches, despite using less information, and (2) in testing, it scales well to environments with arbitrary numbers of agents and obstacles. We illustrate these results using four task environments, including one with predetermined goals for each agent, and one in which the agents collectively try to cover all goals. Code available at https://github.com/nsidn98/InforMARL.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SRMT: Shared Memory for Multi-agent Lifelong Pathfinding

    cs.LG 2025-01 conditional novelty 6.0 of 10

    Shared recurrent memory with global broadcast improves coordination in decentralized multi-agent pathfinding and generalizes to longer corridors better than private-memory baselines.

Pith tools