REVIEW 3 cited by
Multi-Agent Reinforcement Learning for Autonomous Driving: A Survey
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Reinforcement Learning (RL) is a potent tool for sequential decision-making and has achieved performance surpassing human capabilities across many challenging real-world tasks. As the extension of RL in the multi-agent system domain, multi-agent RL (MARL) not only need to learn the control policy but also requires consideration regarding interactions with all other agents in the environment, mutual influences among different system components, and the distribution of computational resources. This augments the complexity of algorithmic design and poses higher requirements on computational resources. Simultaneously, simulators are crucial to obtain realistic data, which is the fundamentals of RL. In this paper, we first propose a series of metrics of simulators and summarize the features of existing benchmarks. Second, to ease comprehension, we recall the foundational knowledge and then synthesize the recently advanced studies of MARL-related autonomous driving and intelligent transportation systems. Specifically, we examine their environmental modeling, state representation, perception units, and algorithm design. Conclusively, we discuss open challenges as well as prospects and opportunities. We hope this paper can help the researchers integrate MARL technologies and trigger more insightful ideas toward the intelligent and autonomous driving.
Forward citations
Cited by 3 Pith papers
-
MUTE: Return-Preserving Communication Unlearning for Efficient Multi-Agent Coordination
Value-guided unlearning of low Counterfactual Message Value channels from an unrestricted MARL policy yields 80–90% bandwidth cuts with bounded return loss.
-
Automated Vehicles Should be Connected with Natural Language
A vision paper recommending natural language as the universal communication medium for connected and automated vehicles.
-
Multi-Agent Reinforcement Learning in Wireless Distributed Networks for 6G
A comprehensive survey of multi-agent reinforcement learning for wireless distributed networks in 6G, covering structures, algorithms, enhanced techniques, and applications.
Discussion (0). Continue with ORCID to comment.