Pith. sign in

REVIEW

CoDreamer: Communication-Based Decentralised World Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2406.13600 v1 pith:GUQZ244C submitted 2024-06-19 cs.AI

classification cs.AI
keywords codreamerapplicationcommunicationdreamerenvironmentslearnedmodelsmulti-agent
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Sample efficiency is a critical challenge in reinforcement learning. Model-based RL has emerged as a solution, but its application has largely been confined to single-agent scenarios. In this work, we introduce CoDreamer, an extension of the Dreamer algorithm for multi-agent environments. CoDreamer leverages Graph Neural Networks for a two-level communication system to tackle challenges such as partial observability and inter-agent cooperation. Communication is separately utilised within the learned world models and within the learned policies of each agent to enhance modelling and task-solving. We show that CoDreamer offers greater expressive power than a naive application of Dreamer, and we demonstrate its superiority over baseline methods across various multi-agent environments.

Discussion (0). Continue with ORCID to comment.

Pith tools