Pith. sign in

Paper Citation Record · LEDGER

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL

As of 19 August 2026, this Paper Citation Record lists 99 of 99 outbound references and 1 inbound Pith citation observation for arXiv:2504.15425.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.15425 v1

Coverage vector

measured 99 of 99 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:35:49.596258Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-13T22:56:33.073457Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T22:58:23.930563Z

Reference resolution

99 of 99 outbound references displayed

  • verified exact0
  • verified fuzzy56
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 021ed894-e2a8-4fad-af50-233bc9b42bc9 · outbound

This paper cites Constrained policy optimization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Constrained policy optimization

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.100505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.100505Z digest=sha256:0eb1986c7516af686bf78e5878ed7e6e46a8c0fc82ffb04bc55883bb908f6ff9

Observation 3f762109-487a-40e7-a875-1da5a0add65d · outbound

This paper cites Learning transferable cooperative be- havior in multi-agent team.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Learning transferable cooperative be- havior in multi-agent team

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.105943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.105943Z digest=sha256:a5aa7a9eca41e00a950c5f8ad75b3ccf0c150902a54ab978f8371a9312f9dfc0

Observation f011d933-34b0-4ec8-956f-40751c62e234 · outbound

This paper cites Constrained Markov decision processes.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Constrained Markov decision processes

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.111335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.111335Z digest=sha256:4f41fff8d418536e561469fba54bf066aaf47cc26a9a312090a2ee28d7c7c38f

Observation bddfa60f-aab9-4940-a474-151512c7062a · outbound

This paper cites Casadi: a software framework for nonlinear optimization and optimal con- trol.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Casadi: a software framework for nonlinear optimization and optimal con- trol

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.116234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.116234Z digest=sha256:5499248baaf3a29682cc8d73d7905d458a25165b6e0a16fd3dd7c2f7917628a7

Observation ba6db136-ebf3-4e72-8c1b-16766ba2e162 · outbound

This paper cites Hamilton-jacobi reachability: A brief overview and recent advances.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Hamilton-jacobi reachability: A brief overview and recent advances

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.121126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.121126Z digest=sha256:0dfa547f92146d2e75bca427313f65dd68be943e47acdf5838b88a8177da075f

Observation 66e5c626-3936-4ea5-98fb-f527b5901967 · outbound

This paper cites Dynamic programming and optimal control: Volume I, volume 4.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Dynamic programming and optimal control: Volume I, volume 4

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.125942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.125942Z digest=sha256:da71e5d73147f00cc29c76346b051f69a4c5a1c55d3c58296c1cd2dcb39c4daa

Observation 647f9e09-641c-4aa1-b5cb-6078b0a34493 · outbound

This paper cites Synthesis of minimum-cost shields for multi-agent systems.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Synthesis of minimum-cost shields for multi-agent systems

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.131262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.131262Z digest=sha256:32710d074ea2ed6efd701a558977640c6d1a476fb44b25b06e2b56931b8b5870

Observation 01d8c8e2-5ed1-4d53-b4da-b702c9846195 · outbound

This paper cites An actor-critic algorithm for constrained markov decision processes.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL An actor-critic algorithm for constrained markov decision processes

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.136145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.136145Z digest=sha256:7101e0c2f47b67bb66ebbc92365b4dc8d7c7b96f5f52ea116ec9004d31c6a319

Observation 387cca71-61aa-4619-9945-16a5dd173f73 · outbound

This paper cites Stochastic Approximation: A Dynamical Systems Viewpoint, volume 48.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Stochastic Approximation: A Dynamical Systems Viewpoint, volume 48

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.141464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.141464Z digest=sha256:ee85ea555feae8450de0e4317834471c59dfc71e5ba4d4c88a1e4adcc445674f

Observation f2eede24-1a56-4d49-8870-c9075d9289b3 · outbound

This paper cites Convex optimization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Convex optimization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.146263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.146263Z digest=sha256:e575f126ab1150df6093380b17177ff6e5d68eb544e2c9fb3cd8e15f93f7692e

Observation b3dae811-fdde-4419-b16c-879328c11f8f · outbound

This paper cites Safe Multi-Agent Reinforcement Learning through Decentralized Multiple Control Barrier Functions.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe Multi-Agent Reinforcement Learning through Decentralized Multiple Control Barrier Functions

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.151496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.151496Z digest=sha256:5084064969b3fd7d1b89c8e27d3e21d47bc6a39121f722b163f4f10e2b8ff169

Observation 47be9f9e-05d5-44dd-8c33-ac89b9305b93 · outbound

This paper cites A new hybrid quadratic/bisection algorithm for finding the zero of a nonlinear function without using derivatives.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A new hybrid quadratic/bisection algorithm for finding the zero of a nonlinear function without using derivatives

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.156876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.156876Z digest=sha256:19bdbdde76158148409cf554a6137b2f38737bc01dfc681b1eb74d617ff0543e

Observation 1182b6e9-eee5-4d80-9d2c-bf4c05b5cc3c · outbound

This paper cites Socially aware motion planning with deep re- inforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Socially aware motion planning with deep re- inforcement learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.161771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.161771Z digest=sha256:954af2383d4253d0b22fdc4f3394cb5aa1cbc464a422e66aab17ecf478a2a430

Observation 4924cc63-5136-465e-8dee-b868114ab8ec · outbound

This paper cites Decentralized non-communicating multiagent col- lision avoidance with deep reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Decentralized non-communicating multiagent col- lision avoidance with deep reinforcement learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.166199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.166199Z digest=sha256:5aa6764d03c928e2c83fe352e1bb7c4468416b6801062604093c8d0844c916c8

Observation 2986275c-4cda-40fe-b0ac-f6f90ac66d74 · outbound

This paper cites On the duality gap of constrained cooperative multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL On the duality gap of constrained cooperative multi-agent reinforcement learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.170705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.170705Z digest=sha256:b6063d8d2e488f263dcfdd0340cab5f252c5641874c0e397cd64dc76985076b8

Observation dfd586fa-ea56-4f86-b855-92705b94571f · outbound

This paper cites Computational aspects of distributed optimization in model predictive control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Computational aspects of distributed optimization in model predictive control

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.175559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.175559Z digest=sha256:9accafaa456646f9b1369874ceeca987bc061f21d280a350a156ed69892933c7

Observation e1c51361-7511-40ce-b559-ece92db3f9bd · outbound

This paper cites De- tecting, localizing, and tracking an unknown number of moving targets using a team of mobile robots.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL De- tecting, localizing, and tracking an unknown number of moving targets using a team of mobile robots

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.180383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.180383Z digest=sha256:58a0dda6fdfe8ced43db1a03b0b3a6405df3abea163aa17f6dac97649a472a9e

Observation 357a9e3c-ec56-4522-860a-ca97b7967865 · outbound

This paper cites Provably efficient gener- alized lagrangian policy optimization for safe multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Provably efficient gener- alized lagrangian policy optimization for safe multi-agent reinforcement learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.185205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.185205Z digest=sha256:5d5403141a2f3014aeeacfde6dbc6ed240827adbc49678b49be10362b9efb750

Observation 81d691fe-8670-41e6-b37e-f155dddd9993 · outbound

This paper cites Safe Multi-Agent Reinforcement Learning via Shielding.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe Multi-Agent Reinforcement Learning via Shielding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.190136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.190136Z digest=sha256:04109157eb2bb84d61498faa85faed5b9ac620c75bf4f412ff2d79b56ad0ce49

Observation e0024367-a868-409e-8988-6a3c1dbe99a2 · outbound

This paper cites Safe multi- agent reinforcement learning via shielding.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe multi- agent reinforcement learning via shielding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.194900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.194900Z digest=sha256:c125a0a28d0b1bdf7251d4593a45b1ac5f47bc289c5b362c323d15453a6bb4f0

Observation 5101e1bb-d11f-46c4-b075-44e424cad64b · outbound

This paper cites Mo- tion planning among dynamic, decision-making agents with deep reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Mo- tion planning among dynamic, decision-making agents with deep reinforcement learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.199462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.199462Z digest=sha256:f2409635957eb004bd80d4ab0137812476fea9e1f001cfd56eb2c49552075343

Observation a2625dc9-4098-48a1-a000-2456fb48fdf1 · outbound

This paper cites A distributed model predictive control strategy for constrained multi- agent systems: The uncertain target capturing scenario.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A distributed model predictive control strategy for constrained multi- agent systems: The uncertain target capturing scenario

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.960217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.203864Z digest=sha256:4293c75a6447ab0a60a582f7e41c911b2e45726ebaad28ea69698ec5542df1f3

Observation a37f3fec-48f4-496b-9e91-42bd4d4dca69 · outbound

This paper cites Counterfactual multi-agent policy gradients.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Counterfactual multi-agent policy gradients

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.944198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.208585Z digest=sha256:a9a5937346764f4d8ca43ed34b7602378ec14d6d8a01887151062fe7b3dfa42a

Observation a36e666d-7a33-4222-aba9-4073d813675b · outbound

This paper cites Iterative reachability estimation for safe reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Iterative reachability estimation for safe reinforcement learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.213377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.213377Z digest=sha256:2673ae9d76cb8b74854f23c5df67230a1b8ad868019b585fbc43c21490c82079

Observation 3ca00963-31c1-42f3-99d1-b440454e11c5 · outbound

This paper cites Learning safe control for multi- robot systems: Methods, verification, and open chal- lenges.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Learning safe control for multi- robot systems: Methods, verification, and open chal- lenges

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.916330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.218177Z digest=sha256:08acdc459e15818cb96f9025a53fac33ad6d30f45a72536d3f82b4c94531ba30

Observation 2e19ab2a-784e-41e9-9806-908398d7a7fc · outbound

This paper cites A reinforce- ment learning framework for vehicular network routing under peak and average constraints.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A reinforce- ment learning framework for vehicular network routing under peak and average constraints

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.899242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.222823Z digest=sha256:c09c1d9856acbf85ce3171cded384e4f1b54698226572c7cc152cce6d1a21da4

Observation 7e679174-932b-4bb9-9733-7274dc2c81d5 · outbound

This paper cites Crazyflie 2.0 quadrotor as a platform for research and education in robotics and control engineering.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Crazyflie 2.0 quadrotor as a platform for research and education in robotics and control engineering

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.882587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.228778Z digest=sha256:f00658ba60aa2cd3cf485a88044a7786d37ae43b562b666a4c79e3a5af7d2cb9

Observation 6f6e2a00-a966-4689-a670-3e1230b068cd · outbound

This paper cites Snopt: An sqp algorithm for large-scale constrained optimization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Snopt: An sqp algorithm for large-scale constrained optimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.234123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.234123Z digest=sha256:0cbff42c743f89028ac2b626cdb83cfd7b094b917c1ef3d0a1e6d048a2e61edd

Observation c7c38f09-8592-41a2-b2e7-97f7db9891a4 · outbound

This paper cites Nonlinear model predictive control: theory and algorithms.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Nonlinear model predictive control: theory and algorithms

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.853203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.239025Z digest=sha256:c262d4c6a7e1e4c3487a7f003bd0fa3fd2de751c009cc38738423a91de98eb2b

Observation cd1c5b7b-32eb-4073-a64f-0fea3ec9c907 · outbound

This paper cites Multi-Agent Constrained Policy Optimisation.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-Agent Constrained Policy Optimisation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.243705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.243705Z digest=sha256:c45fdbc89d91779a73a5ca7ed9a0fb5cfc8a11bea6612bb00a9771f639efc58c

Observation 75bcca73-1d38-4d39-b5ae-6474eddf8218 · outbound

This paper cites A Review of Safe Reinforcement Learning: Methods, Theory and Applications.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A Review of Safe Reinforcement Learning: Methods, Theory and Applications

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.248824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.248824Z digest=sha256:06d54a67aa3ff984d18094d2d1d9e716e4ec5b8970b40ba901893f43779e5081

Observation df65dcad-5e97-4d33-84b7-8389674db06d · outbound

This paper cites Safe multi-agent reinforcement learning for multi-robot control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe multi-agent reinforcement learning for multi-robot control

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.834594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.253790Z digest=sha256:100d3aa6544830586ce197ff510a1389e880d8d7cfebef3cace1a376996e7056

Observation 7fe7dc21-ca19-41c5-a1c5-f1443941b0a2 · outbound

This paper cites Coordinated reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Coordinated reinforcement learning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.816795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.258380Z digest=sha256:f8ce7db1ac5d4f26a024d4a16cef5e39cb02e055b248cc3a539ca49621f436db

Observation bce72a09-d609-4b10-9671-4cce7b9d10ed · outbound

This paper cites Deep recurrent q-learning for partially observable mdps.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Deep recurrent q-learning for partially observable mdps

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.800028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.263244Z digest=sha256:85951227a11c02dde014c69f937d9d19ab7ebbbb92f775bb23cb5f4f5115a8f4

Observation efc781f4-8330-4f9e-8ab0-b0fa59a789d6 · outbound

This paper cites Autocost: Evolving intrinsic cost for zero-violation reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Autocost: Evolving intrinsic cost for zero-violation reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.782922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.269034Z digest=sha256:78040c973ca7c6a3248a028e744ea37cc4b74be24158208997be672ea23f8b0d

Observation 60b7bea2-7cf5-4183-ab78-6e180f5141a5 · outbound

This paper cites Safedreamer: Safe reinforcement learning with world models.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safedreamer: Safe reinforcement learning with world models

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.765390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.273968Z digest=sha256:3dd790d8f8de3fd0de81dc8d72b6c4504ff03213e72d2e880db772032b956f12

Observation bcef9d51-07e5-449c-8798-89fdffd4bdb4 · outbound

This paper cites Distributed optimization in multi-agent robotics for industry 4.0 warehouses.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Distributed optimization in multi-agent robotics for industry 4.0 warehouses

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.747003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.280957Z digest=sha256:649de528964604068ba516bcc8144d3ec3186087721347ac2664b16b289cdab2

Observation ea097793-3e3b-4905-be52-a3a255c5c929 · outbound

This paper cites Cmix: Deep multi- agent reinforcement learning with peak and average con- straints.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Cmix: Deep multi- agent reinforcement learning with peak and average con- straints

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.728160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.286188Z digest=sha256:e7a1113a48af1ef2d1934406243cde2af50c721ad919983720711939fa03f9fb

Observation 1406ba9c-9992-4145-9e9c-3710efa7d9fc · outbound

This paper cites Towards optimally decentralized multi-robot collision avoidance via deep reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Towards optimally decentralized multi-robot collision avoidance via deep reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.705889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.290715Z digest=sha256:010731af571b4e01e4b59be4f7224f0bfd40e8ad79c7f67524361c78fad1b16e

Observation 08f7fcea-19d9-480d-a403-5c4ed1d81e61 · outbound

This paper cites Multi-agent actor- critic for mixed cooperative-competitive environments.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-agent actor- critic for mixed cooperative-competitive environments

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.687627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.295450Z digest=sha256:dc8f0e574269d9a7ff56a85d96f58863cb2f850916d3aa25e8d24eab9e085087

Observation 1ec2c0d6-124d-4e49-8d1c-7e4796f40a9c · outbound

This paper cites Decentralized policy gradient descent ascent for safe multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Decentralized policy gradient descent ascent for safe multi-agent reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.670639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.300993Z digest=sha256:a6bb258c590da90d7d0344cfb69aab1203b8a66ed9c346212ade855463ceadfa

Observation 37f2e7e6-78ba-46cb-8ac7-34d294efe1ca · outbound

This paper cites Trajectory generation for multiagent point-to-point transitions via distributed model predictive control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Trajectory generation for multiagent point-to-point transitions via distributed model predictive control

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.653222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.306462Z digest=sha256:abe8e3f7983566e2aa2f068c65a4c86f25228466fc7ad8f320f7069e510bedad

Observation 94b03caa-0cb1-4e78-8c51-3b1b3280dc16 · outbound

This paper cites Online trajectory generation with distributed model predictive control for multi-robot motion planning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Online trajectory generation with distributed model predictive control for multi-robot motion planning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.634565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.311335Z digest=sha256:dcfc83cb43778435b12ef93971d0e72cf4a6f76e4d0aedd80b08db41c32c950a

Observation 3d62944b-59f5-43d1-a914-1388f859acdf · outbound

This paper cites On reachability and minimum cost optimal control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL On reachability and minimum cost optimal control

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.316420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.316420Z digest=sha256:35ea002d10bf0fac8e2e4829e3560f771101dadd4c760563cd4d3e30618ae9c0

Observation a67092fb-f45b-4168-8ce2-1c802594f8d9 · outbound

This paper cites Lifelong Multi-Agent Path Finding for Online Pickup and Delivery Tasks.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Lifelong Multi-Agent Path Finding for Online Pickup and Delivery Tasks

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.321247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.321247Z digest=sha256:202f1071fe1006c66c04226038320fd3589a3833a445bac73ea99e40d63d254f

Observation 93311ed5-1199-4c60-9bf5-029f404aa922 · outbound

This paper cites Hamilton–jacobi formulation for reach–avoid differential games.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Hamilton–jacobi formulation for reach–avoid differential games

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.605556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.326900Z digest=sha256:5484bd67dd0e4e0da023982d1b41a52bffc52b485fade683a3e7bf6327a8d827

Observation bfd02cfa-a22b-4953-be02-47eba0a8793a · outbound

This paper cites Safe value functions.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe value functions

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.587653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.331833Z digest=sha256:c1b78f951403f8087d0b972ec577f5e9e9b44f33628f0994080d2ea047ee2c83

Observation bc854208-b398-4db6-beec-579021b599d8 · outbound

This paper cites Shield decentralization for safe multi-agent reinforce- ment learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Shield decentralization for safe multi-agent reinforce- ment learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.570150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.337215Z digest=sha256:293ab7c11ff28c01a0c83d045b19a6f4deee8f0816c45dbc173380bba83b0409

Observation 140a3280-0ba2-4594-8868-3c5926e56dfd · outbound

This paper cites A time-dependent hamilton-jacobi formulation of reachable sets for continuous dynamic games.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A time-dependent hamilton-jacobi formulation of reachable sets for continuous dynamic games

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.342067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.342067Z digest=sha256:b0e7803617bed3901f4e3756ea459928529b23980dccabfb28bf67c9849240b9

Observation e0fec4a7-ff8b-4d82-a8bc-93d754141da9 · outbound

This paper cites Distributed model predictive safety certification for learning-based control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Distributed model predictive safety certification for learning-based control

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.540428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.346763Z digest=sha256:16cbe60dcb5278ca2f4d52028ad8f6c8b78ed698fb877e3d6f538d4751abcde9

Observation 2f02698a-e874-44db-974c-e81feffb43b8 · outbound

This paper cites Scalable multi-agent reinforcement learning through intelligent information aggregation.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Scalable multi-agent reinforcement learning through intelligent information aggregation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.523861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.351927Z digest=sha256:ff748abfc854a945318c19faee5cc9a6cd2728ab49e4474d9b185c2c89b9e268

Observation fd4865da-e1d0-4b21-a8d3-347e590770d3 · outbound

This paper cites Distributed optimization for control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Distributed optimization for control

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.506329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.356555Z digest=sha256:324c916ed3b9e11b68a3eeeb02872e0c3a833700846f89bd0a405b734c81488a

Observation a9691ad1-2670-49cf-b45e-2c22cdba0e9c · outbound

This paper cites Numerical opti- mization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Numerical opti- mization

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.361981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.361981Z digest=sha256:20d7b2827456e6c99d8d9c116d7aeb02a129f935b573b22d816fdc0e28183521

Observation 13192686-87fe-4a3b-9a5c-6f1a74d3a50f · outbound

This paper cites Facmac: Factored multi- agent centralised policy gradients.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Facmac: Factored multi- agent centralised policy gradients

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.479369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.366973Z digest=sha256:0499aabfe87347bb0f4c15829200ea3971e82278c3b29a8448777b0d1c44fa58

Observation 21340052-cdf4-4616-b3b5-61c0855e9881 · outbound

This paper cites Decentralized Safe Multi-agent Stochastic Optimal Control using Deep FBSDEs and ADMM.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Decentralized Safe Multi-agent Stochastic Optimal Control using Deep FBSDEs and ADMM

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.372051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.372051Z digest=sha256:05d231e7a062a1ec40378cbce3d916e9935dfaedf89f9698f534a9a00c2081f3

Observation aabf5350-5c2b-4d98-8168-1bd78439b3fb · outbound

This paper cites Learning safe multi-agent control with decentralized neural barrier certificates.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Learning safe multi-agent control with decentralized neural barrier certificates

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.462917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.377313Z digest=sha256:4b2aa12e07cd5a75f6450275f02ab1c8647666adf1b7654dde0944a67f5a0102

Observation f2d235d6-6442-4c49-b465-007add8b0b91 · outbound

This paper cites Weighted qmix: Expanding monotonic value function factorisation for deep multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Weighted qmix: Expanding monotonic value function factorisation for deep multi-agent reinforcement learning

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.446467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.382001Z digest=sha256:f94722ca943e036c5770b360f8e87704e13d104534fef3dabb72983c919b78d4

Observation 9146e6a4-38c0-47b9-b853-b50494f6c11c · outbound

This paper cites Monotonic value function factorisation for deep multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Monotonic value function factorisation for deep multi-agent reinforcement learning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.430474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.386675Z digest=sha256:e664b968188510dee926f113e27ef1f1a70b76f5c21b38b5b24b4272833fd073

Observation bd3fb1fe-ecc5-4994-8921-57747eb556e1 · outbound

This paper cites A stochastic approx- imation method.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A stochastic approx- imation method

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.414429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.391565Z digest=sha256:e6f41b9d9eefa778e33a1f12cc994e3289a19c5b37e6a2ffc2473970779d0a3b

Observation 3c9c3079-afe6-4d49-845b-79ebf0a70db0 · outbound

This paper cites Con- strained markov decision processes via backward value functions.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Con- strained markov decision processes via backward value functions

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.398173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.396466Z digest=sha256:465bab1c6c124ac108da9e5ab003064dfe7e440a75dfcfc6f7c9681eacbf93d0

Observation d6075347-e324-4dba-8847-a101ca763e47 · outbound

This paper cites Trust region policy optimiza- tion.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Trust region policy optimiza- tion

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.381182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.401269Z digest=sha256:fa3b1da14fef0e1d98fdccc4cf2ef7d5616aa706543b8c10f875da9a1ea6d51b

Observation 3b10c4ad-aecd-42e6-93e6-5b509e001dc4 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.405908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.405908Z digest=sha256:c54ccca44246e427d764a1c434a5aa391d3fc4790db872f0f52183f18120e7b5

Observation fee63235-7e9a-4da2-a553-945bc6e6ac5b · outbound

This paper cites Proximal Policy Optimization Algorithms.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Proximal Policy Optimization Algorithms

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.411092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.411092Z digest=sha256:9537889989a08ceedb9605f761c98734cae9ad6568e50d01c20ff0ebb0c58493

Observation 3adb61f4-599e-4493-8fef-ffab520f8d3c · outbound

This paper cites Multi-agent motion planning for dense and dynamic environments via deep reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-agent motion planning for dense and dynamic environments via deep reinforcement learning

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.363450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.415932Z digest=sha256:d7de88c6999189c7c9c612b4267b860d2b539c3c43d9471f286475f73fb28e9a

Observation 41fad7e6-d630-4b0c-ba21-c3394c99f9dd · outbound

This paper cites Masked Label Prediction: Unified Message Passing Model for Semi-Supervised Classification.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Masked Label Prediction: Unified Message Passing Model for Semi-Supervised Classification

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.420858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.420858Z digest=sha256:3db29b2897d69ce915b0b75b850e40e13e93fe6d8231e3641568ac875a5e1d10

Observation 1309f659-a663-49a7-8ca9-119b5083d7bb · outbound

This paper cites Solving stabilize-avoid optimal control via epigraph form and deep reinforce- ment learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Solving stabilize-avoid optimal control via epigraph form and deep reinforce- ment learning

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.347386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.426325Z digest=sha256:d03ced8cfb8ca26ba14a943f817ee2932d9bf3050cede7c1fbd41fc5a51e7c30

Observation 29aebdd6-f22a-4bed-ab12-78370cd16582 · outbound

This paper cites Solving minimum-cost reach avoid using reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Solving minimum-cost reach avoid using reinforcement learning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.330772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.431478Z digest=sha256:20a9c039d8eb18224c5a761b706f16b48cd0db3ce185bfd3714d9a669702bf5d

Observation bf9028ad-636b-40e7-baee-5aa78a81ff94 · outbound

This paper cites Predictive control of aerial swarms in cluttered environ- ments.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Predictive control of aerial swarms in cluttered environ- ments

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.311938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.436382Z digest=sha256:deff7ef78b11c77d8ea7315c27f2f54fe52fece7fdf9cd2a1ceeacd032c7d054

Observation 1926dfd9-129d-44d3-9ebb-4915cc2c9790 · outbound

This paper cites Value-Decomposition Networks For Cooperative Multi-Agent Learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.441863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.441863Z digest=sha256:ab07664d192120b423d1b7dbd00c9055168e69627e43bda5a20b2a618b5b341f

Observation 54313299-44bc-4c9e-ba05-01afdf70c610 · outbound

This paper cites Mankowitz, and Shie Mannor.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Mankowitz, and Shie Mannor

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.293658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.447051Z digest=sha256:780d51124374285d8234bc7ba628bdca6bc2deb0a5d2905e1206dbee78d039e6

Observation 1522d750-8cc9-46c0-92a8-36bb1816632f · outbound

This paper cites A game theoretic approach to controller design for hybrid systems.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A game theoretic approach to controller design for hybrid systems

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.276454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.452061Z digest=sha256:ea5597bd42176820f56e81449f8c2dd9759258fb553a5238722915e60a1a3db8

Observation fcfdc6fa-dcc9-44db-82af-7ef417761499 · outbound

This paper cites Decentralized multi-agent planning using model predictive control and time-aware safe corridors.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Decentralized multi-agent planning using model predictive control and time-aware safe corridors

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.260275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.456929Z digest=sha256:b585af373311dd8d0691153ab0d1fffd1e5ac091f940d73475990d3fa1b86b55

Observation f2d69c81-cba1-4274-9b0a-19aa5399a1d1 · outbound

This paper cites Initial guess generation for aircraft landing trajec- tory optimization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Initial guess generation for aircraft landing trajec- tory optimization

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.241930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.461859Z digest=sha256:e0635626c6eab416543ed302aa9a8e4b72ffa1dde4ee01942bfee139a83a1844

Observation 85c5ed1a-92ab-48d1-ad9b-02b126e0355f · outbound

This paper cites QPLEX: Duplex Dueling Multi-Agent Q-Learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL QPLEX: Duplex Dueling Multi-Agent Q-Learning

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.466812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.466812Z digest=sha256:83ec99517c7816753e7836d47ceb2c22d6c350e08d2bfaba0fd23fa7f9f2a626

Observation 587b983f-2f82-460d-b4f8-21478fb37347 · outbound

This paper cites A synthesis approach of distributed model predictive control for homogeneous multi-agent system with collision avoidance.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A synthesis approach of distributed model predictive control for homogeneous multi-agent system with collision avoidance

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.225310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.472565Z digest=sha256:4d7769202564965d0d34537df261e75bf6f6eee4997ef0356106c3ac1a826da7

Observation 04916c6e-b7d6-41bc-be07-e2a7093bf891 · outbound

This paper cites Multi-agent deep reinforcement learning for urban traffic light control in vehicular networks.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-agent deep reinforcement learning for urban traffic light control in vehicular networks

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.478723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.478723Z digest=sha256:bfdb4ec96613e303cc3a38176e3ff836d2c3855888d7a5f2b4c24fced9657872

Observation abafd294-d515-4529-a413-1c0fd57fab81 · outbound

This paper cites Model-based Dynamic Shielding for Safe and Efficient Multi-Agent Reinforcement Learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Model-based Dynamic Shielding for Safe and Efficient Multi-Agent Reinforcement Learning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.483573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.483573Z digest=sha256:0ba8e4d9c4ceed77e28e6417de569bbd8210aa1921a25ad059f558130b39e2ad

Observation 1f6914fa-a1f2-4d92-8206-66da2e5e14e4 · outbound

This paper cites Crpo: A new approach for safe reinforcement learning with convergence guarantee.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Crpo: A new approach for safe reinforcement learning with convergence guarantee

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.196643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.488687Z digest=sha256:5772df0a3e63208907e4dea5674477f1cf1a5928e51a36bf64271bf4c70fb417

Observation 71321466-c55e-4511-a348-002b28d81941 · outbound

This paper cites Qatten: A General Framework for Cooperative Multiagent Reinforcement Learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Qatten: A General Framework for Cooperative Multiagent Reinforcement Learning

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.493384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.493384Z digest=sha256:70931756fb7cbdc9ff594227a686ad9ac31a0f1d0bd7bec102be573f5beedfa4

Observation 69237a14-51b7-4de9-a903-c724df5cc2e1 · outbound

This paper cites The surprising effectiveness of ppo in cooperative multi-agent games.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL The surprising effectiveness of ppo in cooperative multi-agent games

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.178088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.499001Z digest=sha256:aeadac74d247c414278a72217cf9d7819ccd61b4d2843fcd848d7151d7c60d18

Observation 9095540a-b7ea-4b7b-9411-59b023a5f2e6 · outbound

This paper cites Reachability constrained reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Reachability constrained reinforcement learning

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.161635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.504513Z digest=sha256:49ca9a47f5560147fae7e129de82fe4c3fe0f2200c07fa191f52304af38fb7b8

Observation b2c0f567-268e-4b73-9e38-2d1f76ccf2b8 · outbound

This paper cites Safe reinforcement learning using robust mpc.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe reinforcement learning using robust mpc

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.143568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.509472Z digest=sha256:4f8809bc3683539504050ea09aa63da5b2fbc416fb174abe9784f8049923a1ad

Observation 73060d35-0d38-480d-bfe0-fee0857b7def · outbound

This paper cites Fully decentralized multi-agent re- inforcement learning with networked agents.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Fully decentralized multi-agent re- inforcement learning with networked agents

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.124360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.514827Z digest=sha256:7c73e7f743b7e75b81b5bae55460d854eea87b0550dfa264eed8a12bf485c941

Observation ffe68379-54e4-4d0b-a739-3bd8773eda9e · outbound

This paper cites Multi- agent reinforcement learning: A selective overview of theories and algorithms.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi- agent reinforcement learning: A selective overview of theories and algorithms

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.105282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.520678Z digest=sha256:53850009188bdeac0acbb1c7a9ebcf45349dc18a0ac63606e13cceb6b23073cd

Observation 6eb47d82-ce9b-450e-a34e-ae2cb5161338 · outbound

This paper cites Neu- ral graph control barrier functions guided distributed collision-avoidance multi-agent control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Neu- ral graph control barrier functions guided distributed collision-avoidance multi-agent control

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.088952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.525301Z digest=sha256:b5361dc08b7fa7dd701e24bfb26393883d72f3f20346742ce4966d4e68e5673f

Observation cc55ae64-209c-4442-a094-ae8e0610b2d1 · outbound

This paper cites Discrete GCBF proximal policy optimization for multi-agent safe optimal control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Discrete GCBF proximal policy optimization for multi-agent safe optimal control

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.072861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.530170Z digest=sha256:71870d08e4e2b423959d9785f6aac3543e0d9bb3828e6b4251fbc46bf421aa4b

Observation 56eb93fb-3c0f-401d-b7cd-674dd7a7498e · outbound

This paper cites GCBF+: A neural graph control barrier function framework for distributed safe multiagent control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL GCBF+: A neural graph control barrier function framework for distributed safe multiagent control

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.055969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.534940Z digest=sha256:b561eb6e82646ea6bc3d8a4753d59788c74544f832d05e3382ee1ccdb4db10e6

Observation 89e57848-5be3-4969-8f1a-b94fca74a625 · outbound

This paper cites MAMPS: Safe Multi-Agent Reinforcement Learning via Model Predictive Shielding.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL MAMPS: Safe Multi-Agent Reinforcement Learning via Model Predictive Shielding

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.539360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.539360Z digest=sha256:d6c0e293ec41cdfc491a73fb18b1a47d544e6f9f26def92ca64db9bd079a5036

Observation cc1d6b07-680b-41d2-9377-4f86753bcc21 · outbound

This paper cites Model-free safe control for zero-violation reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Model-free safe control for zero-violation reinforcement learning

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.040156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.544295Z digest=sha256:1a47fc5ac9a0618dbca519460a796d440264c3b508bb94aaf7e0fd1ad6fc2c45

Observation 02e63a0a-c108-48df-95a8-786b32f80c81 · outbound

This paper cites Multi-agent first order con- strained optimization in policy space.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-agent first order con- strained optimization in policy space

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.023762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.549368Z digest=sha256:6e5868956c022c77160187920aedd859b68772bf538331b322d54224203c1fb7

Observation c23fb004-13ae-4b79-bebd-59a7e5ab4f6f · outbound

This paper cites Fast, on-line collision avoidance for dynamic vehicles using buffered voronoi cells.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Fast, on-line collision avoidance for dynamic vehicles using buffered voronoi cells

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.006851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.555561Z digest=sha256:08591fbec6f6f9678daa0dff275a3443df5bb8dfe389859aec298810c16a149e

Observation d8b0142c-0fc5-41ba-b1f1-3b1efdd930f5 · outbound

This paper cites Trajectory optimization for nonlinear multi-agent systems using decentralized learning model predictive control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Trajectory optimization for nonlinear multi-agent systems using decentralized learning model predictive control

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:49.990352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.560635Z digest=sha256:7463d8a887f87e0d52fb2c075881ffe8d3cecc2a29a5c662209571bc4d93c36e

Observation 26b1b56b-96f7-4307-bc1d-b75711f7a03f · outbound

This paper cites In other words, for a given z0, the value at the kth timestep is only a function of zk and xk instead of the z0 and the entire trajectory up to the kth timestep.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL In other words, for a given z0, the value at the kth timestep is only a function of zk and xk instead of the z0 and the entire trajectory up to the kth timestep

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:49.972993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.565966Z digest=sha256:31b6868ace9dd62f1253f853ef6e54339a3d45e29a01c62d6184196c964dc38a

Observation f0cdaa1e-0e66-4000-aae9-24dc084b845f · outbound

This paper cites an unresolved cited work.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Unresolved cited work

Reference 94

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:35:49.955745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.570731Z digest=sha256:0487b48f29c529a5f8b41bdcf2baf4bc15274ee728828b3cb8a212d60904f6c0

Observation 6cffaff4-cbbf-4eac-8f7f-8c27fe1a09b2 · outbound

This paper cites an unresolved cited work.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:35:49.938382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.575826Z digest=sha256:54cc7a661757042b4b0391e06d03a72f88cf8fc6daba2069f015a6454b764142

Observation 53d0bdd6-665c-4045-a16d-ea6f6b03ab4d · outbound

This paper cites an unresolved cited work.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:35:49.922510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.580801Z digest=sha256:27b99f4d2c2a04ea71d858aafa883244124f21eeb82adda81135c5e61ef88cec

Observation be1d4854-de22-48a7-876d-7e1571c4931d · outbound

This paper cites APPENDIX D ALGORITHM PSEUDOCODE We describe the centralized training process of Def-MARL in Algorithm 1 and the distributed execution process in Algorithm 2.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL APPENDIX D ALGORITHM PSEUDOCODE We describe the centralized training process of Def-MARL in Algorithm 1 and the distributed execution process in Algorithm 2

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:49.907043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.585636Z digest=sha256:01431da64d4e79b7dceab57eebe0fefa0aeff79390e153d47acee2aff20fc521

Observation 99cdc52b-244d-4645-aa18-6dae73dd4be0 · outbound

This paper cites E⊆{ (i,j )|i∈V a,j ∈V} is the set of edges, denoting the information flow from a sender node j to a receiver agent i.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL E⊆{ (i,j )|i∈V a,j ∈V} is the set of edges, denoting the information flow from a sender node j to a receiver agent i

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:49.889638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.590785Z digest=sha256:150204133847eb70df5b765d1a213eef3fafd0193e7db830fe77381a8cc733c6

Observation 2884d2e1-7170-4d80-beb3-d044cd74f8c9 · outbound

This paper cites an unresolved cited work.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:35:49.870781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T11:35:49.596258Z digest=sha256:ec54b45d3daff5c5ac817153f8fae2dce4ea6159adcf8c71c29d5fd75632fa45

Pith citing papers

Observation e8d1989e-45f5-44b0-bf8c-d9b029676de5 · inbound

Distributed Safety-Critical Control of Multi-Agent Systems with Time-Varying Communication Topologies cites this paper.

Distributed Safety-Critical Control of Multi-Agent Systems with Time-Varying Communication Topologies Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:58:23.933705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T22:56:33.073457Z digest=sha256:6cf09839abdcfbedf39d49784063470c61733e0d5c32eee0a399212998b8d4bc