Pith. sign in

REVIEW 2 cited by

Multi-Agent Target Assignment and Path Finding for Intelligent Warehouse: A Cooperative Multi-Agent Deep Reinforcement Learning Perspective

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.13750 v3 pith:X2LC7GQE submitted 2024-08-25 cs.AI cs.MA

classification cs.AIcs.MA
keywords multi-agentassignmentdeeppathtargetcooperativeintelligentmethod
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Multi-agent target assignment and path planning (TAPF) are two key problems in intelligent warehouse. However, most literature only addresses one of these two problems separately. In this study, we propose a method to simultaneously solve target assignment and path planning from a perspective of cooperative multi-agent deep reinforcement learning (RL). To the best of our knowledge, this is the first work to model the TAPF problem for intelligent warehouse to cooperative multi-agent deep RL, and the first to simultaneously address TAPF based on multi-agent deep RL. Furthermore, previous literature rarely considers the physical dynamics of agents. In this study, the physical dynamics of the agents is considered. Experimental results show that our method performs well in various task settings, which means that the target assignment is solved reasonably well and the planned path is almost shortest. Moreover, our method is more time-efficient than baselines.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion

    cs.RO 2025-08 conditional novelty 4.0 of 10

    Limb-as-agent MAPPO with a shared global critic speeds up humanoid walking policy training and improves gait smoothness over single-agent PPO in simulation and on hardware.

  2. Where Paths Collide: A Comprehensive Survey of Classic and Learning-Based Multi-Agent Pathfinding

    cs.AI 2025-05 conditional novelty 4.0 of 10

    A broad survey of MAPF methods that documents inconsistent evaluation practices and proposes a unified taxonomy.

Pith tools