Pith. sign in

REVIEW 2 cited by

Learning Safe Multi-Agent Control with Decentralized Neural Barrier Certificates

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2101.05436 v4 pith:IO34ZFDY submitted 2021-01-14 cs.MA cs.AIcs.SYeess.SY

classification cs.MAcs.AIcs.SYeess.SY
keywords controlagentsmulti-agentdecentralizedframeworkotherpolicybarrier
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We study the multi-agent safe control problem where agents should avoid collisions to static obstacles and collisions with each other while reaching their goals. Our core idea is to learn the multi-agent control policy jointly with learning the control barrier functions as safety certificates. We propose a novel joint-learning framework that can be implemented in a decentralized fashion, with generalization guarantees for certain function classes. Such a decentralized framework can adapt to an arbitrarily large number of agents. Building upon this framework, we further improve the scalability by incorporating neural network architectures that are invariant to the quantity and permutation of neighboring agents. In addition, we propose a new spontaneous policy refinement method to further enforce the certificate condition during testing. We provide extensive experiments to demonstrate that our method significantly outperforms other leading multi-agent control approaches in terms of maintaining safety and completing original tasks. Our approach also shows exceptional generalization capability in that the control policy can be trained with 8 agents in one scenario, while being used on other scenarios with up to 1024 agents in complex multi-agent environments and dynamics.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Stochastic Neural Control Barrier Functions

    eess.SY 2025-06 reject novelty 6.0 of 10

    A framework for synthesizing and verifying neural control barrier functions for stochastic systems, including new Tanaka-formula-based safety conditions for ReLU networks.

  2. Tackling Uncertainties in Multi-Agent Reinforcement Learning through Integration of Agent Termination Dynamics

    cs.LG 2025-01 conditional novelty 4.0 of 10

    Adding a penalty for ally deaths to distributional multi-agent Q-learning improves win rates on StarCraft II and driving benchmarks compared with six baseline algorithms.

Pith tools