Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-25T21:17:28.832301Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2606.25526.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-25T21:17:28.832301Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
42 of 42 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6b45067e-eff9-4101-a366-75840814ead9 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: International Conference on Machine Learning, pp
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 081a4368-3940-4d0d-b334-364f9e32731a · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d110a2c1-6cf4-4362-b53b-90fa480d5a73 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: International Conference on Learning Representations (2022)
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 202a086b-79e2-43eb-a267-286f5187da70 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning The Journal of Machine Learning Research21(1), 7234–7284 (2020)
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c70f578c-41db-43cd-a757-e387554a43c7 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Advances in neural information processing systems30(2017)
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef5355c4-7fd3-44a2-91fb-08055c8b5910 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 57ae7e59-881e-41da-a366-7b67e4fae686 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Advances in Neural Information Processing Systems35, 24611–24624 (2022)
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0477c5c-04a3-475f-8ca9-6d16aa313ea7 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Advances in neural information processing systems33, 5527–5540 (2020)
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cca2e325-ab9a-49b3-b088-3e086cbbddb8 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: International Con- ference on Learning Representations (2022)
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af68f09f-19b0-4cd0-8ddc-64cf6cba99ae · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: International Conference on Machine Learning, pp
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 344996c8-efe8-4f74-b1e2-985b9130ecb2 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: Proceedings of the AAAI Conference on Artificial Intelligence, vol
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c690da5-698d-4fbc-a0af-0a9dcdc76f00 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: International Conference on Learning Representations (2019)
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cb1eea1-b308-4ea9-b4fe-07034c582e38 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning A Survey of Learning in Multiagent Environments: Dealing with Non-Stationarity
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e39db308-939c-4c0e-b727-21b694d1c6a4 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Advances in Neural Information Processing Systems35, 16509–16521 (2022)
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b2eeeb1-963f-4da9-9de6-166efad42d45 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: Proceedings of the 40th International Conference on Machine Learning, vol
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a78b47f-b8f3-4dc1-ae06-122fe7d14509 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: International Conference on Learning Representations (2022)
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81d129e0-2fd9-49dd-bb0a-7f64b5207bde · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Advances in Neural Information Processing Systems34, 12208–12221 (2021)
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f533fb55-fe2b-49c2-bd15-cf3793ef68b1 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems, pp
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5b3ae74-7ac5-4cc4-84ec-1b071287dffb · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Advances in Neural Information Processing Systems34, 26437–26448 (2021)
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b9f5360-f391-4140-9275-3cdc57d47d9d · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Advances in neural information processing systems32(2019)
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb52bd12-3557-4508-8120-10976269fba7 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Neural PPO-Clip Attains Global Optimality: A Hinge Loss Perspective
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cfa13852-eb11-4d7d-b145-5950e759a4f2 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: Proceedings of the Nineteenth International Conference on Machine Learning, pp
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cc91ded-77f3-47a8-ab08-45cda041c5ac · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: Proceedings of the 34 AAAI Conference on Artificial Intelligence, vol
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e311621e-b17b-4d73-ae61-dd336cf0c76e · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0b956c29-dace-4739-a5c3-f9f2956407ac · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Transactions on Machine Learning Research (2023)
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0b0c554-dee5-4b5d-9566-b958c3a26a3b · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning FACMAC: Factored Multi-Agent Centralised Policy Gradients
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9395ea3d-2d17-4909-b2b8-ddebff0f6c08 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning The StarCraft Multi-Agent Challenge
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0d02a4e9-d44e-4bc8-8949-982c2760b064 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Journal of Machine Learning Research25(32), 1–67 (2024)
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8807a59-f697-48f9-9f45-c7ff58acf8bd · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Proceedings of the national academy of sciences114(13), 3521–3526 (2017)
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 959bbcdd-7c7a-4448-ae56-0d1c8bcd3177 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning IEEE Transactions on Neural Networks and Learning Systems (2023)
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbd9fe14-791d-41a0-a134-a41219cda2c3 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Multi-Agent Constrained Policy Optimisation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b68e7cfd-eafb-4dda-b757-7aa3507cba34 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: International Conference on Machine Learning, pp
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f49671a-167e-493a-87a9-1ee90f45b5ec · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: International Conference on Machine Learning, pp
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8ca9ee7-b34a-48c8-b977-662d778cd1d8 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Advances in neural information processing systems 29(2016)
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15b98025-7d42-4040-b4c7-e9744dc19410 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Advances in Neural 35 Information Processing Systems34, 13458–13470 (2021)
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5584e8d-19dd-4c61-8afd-9e0ae4990794 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: The Eleventh International Confer- ence on Learning Representations (2023)
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54fc2b56-a0ac-4865-85e6-66d9825432f9 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning In: Proceed- ings of the 17th International Conference on Autonomous Agents and MultiAgent Systems
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2109b79-44ed-46e2-b8c9-20eb78192180 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning A unified view of entropy-regularized Markov decision processes
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 25b37762-775e-409e-89b2-322c9752f19e · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Mathematical programming 198(1), 1059–1106 (2023)
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48fc81a1-f672-4791-82d0-21e6ab3f41fd · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning SIAM-Society for Industrial and Applied Mathematics, Philadelphia, PA, USA (2017)
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1276ca49-ae62-47a4-86a5-8145d237f1a2 · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning The Journal of Machine Learning Research22(1), 4431–4506 (2021)
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d22bef29-31a7-4ef7-a9c1-3be7f32bb2cc · outbound
Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning Equivalence Between Policy Gradients and Soft Q-Learning
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.