Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:42:57.122356Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 100 of 168 outbound references and 0 inbound Pith citation observations for arXiv:2507.10142.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:42:57.122356Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
100 of 168 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 79b5e472-bed4-41ba-b5c8-887dec4dcccf · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review A comprehensive survey of multia- gent reinforcement learning.IEEE Transactions on Systems, Man, and Cybernetics, Part C (Applications and Reviews), 38(2):156–172, 2008
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5201f86-d96c-4afc-850e-6c3a1c0d5475 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review A Survey of Progress on Cooperative Multi-agent Reinforcement Learning in Open Environment
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0f2f134-c785-428b-86c0-de3151eb2c18 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review A survey of multi-agent deep reinforcement learning with communication.Autonomous Agents and Multi-Agent Systems, 38(1):4, 2024
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd336ca3-e500-4b0d-9ea8-de6c3136fc11 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Single and multi-agent deep reinforcement learning for ai-enabled wireless networks: A tutorial
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70018174-e6d4-4f24-90cb-f8cdbb5dba69 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Multi-agent deep reinforcement learning for multi-robot applica- tions: A survey.Sensors, 23(7):3625, 2023
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2723a35-7894-4bb9-b300-e293b1d02da8 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Multi-agent deep reinforcement learning: a survey.Arti- ficial Intelligence Review, 55(2):895–943, 2022
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 696121d7-a3fe-4624-aa0e-322df50c2068 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review A Survey on Large-Population Systems and Scalable Multi-Agent Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8f5e70e-db62-4177-b0db-980ca59950ff · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Multi-agent reinforcement learning: A selective overview of theories and algorithms.Handbook of reinforcement learning and control, pages 321–384, 2021
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 235e5c4d-84d6-44bf-9238-820e84b031fd · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review MIT press Cambridge, 1998
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e8b7457-8ad0-4659-a19c-885484f81958 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Human-level control through deep reinforcement learning.nature, 518(7540):529–533, 2015
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 426e0eee-2e97-445b-afbb-3eea80e1e4e6 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Deep reinforcement learning: A brief survey.IEEE Signal Processing Magazine, 34(6):26–38, 2017
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8db56f8-81f9-48c6-b81b-24186bc7a58a · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Multi-agent reinforcement learning: Independent vs
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f29492cd-8904-4fd9-94f5-291d256fe61d · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Mean field multi-agent reinforcement learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98c18353-78df-4abc-bf33-bd2bd85f988a · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9297e378-1fce-4a9b-8f71-7bc6345edb63 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Partially observable mean field reinforcement learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16673b40-a680-40dd-ae84-54f199d44cbd · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Rode: Learning roles to decompose multi-agent tasks
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0843ae6c-cfe7-4429-b545-52d132d0353e · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Heterogeneous-Agent Mirror Learning: A Continuum of Solutions to Cooperative MARL
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b755f29-f056-402b-85a1-531e3aeb4c24 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review The surprising effectiveness of PPO in cooperative multi-agent games
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c5869a5-f014-4cda-abcf-89e6ca35181a · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Trust region policy optimisation in multi-agent reinforcement learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba5a190e-3a9b-41c6-854b-2280b374f3ea · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Learning mean-field games.Advances in neural information processing systems, 32, 2019
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05af5343-9291-49f3-851d-cd21fcb346ca · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Believewhatyousee: Implicitconstraintapproachforofflinemulti-agent reinforcement learning.Advances in Neural Information Processing Systems, 34:10299–10312, 2021
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1f5f8ad-56e1-4a5e-aa6b-8f1770f30055 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Plan better amid conservatism: Offline multi-agent reinforcement learning with actor rectification
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f9f8dc2-c385-4cdd-bd30-bda11056eafa · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Off-the-grid marl: Datasets and baselines for offline multi-agent reinforcement learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1abd791-7b37-4b6b-90c7-8d45c50357bd · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Networked multi-agent reinforcement learn- ing in continuous spaces
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8369d25-4724-42ef-b220-93dba4f2d522 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Fully decentralized multi-agent reinforcement learning with networked agents
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12ae6d12-1e4c-4320-b61e-8df57f7649a4 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Graph Convolutional Reinforcement Learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5bfb0a5-5d4a-4406-9cbf-dbf0c4d4cb57 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Mambpo: Sample-efficient multi-robot reinforcement learning using learned world models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1628151-c620-47ff-b7bc-232a4737f9f5 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Mingling foresight with imagination: Model-based cooperative multi-agent reinforcement learning.Advances in Neural Information Processing Systems, 35:11327–11340, 2022
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd410ad3-650c-4f5f-8427-59dd1723b374 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Scalable Multi-Agent Model-Based Reinforcement Learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 730365c9-7faa-4738-bd25-624fb7dcc895 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review MAMPS: Safe Multi-Agent Reinforcement Learning via Model Predictive Shielding
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 42b8179a-8e2a-4e0b-8455-65cafe8fa10f · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Cmix: Deep multi-agent reinforcement learning with peak and average constraints
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c0295ad-d3f8-4abe-8c72-6a21f3c50382 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Safe multi-agent reinforcement learning for multi-robot control
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33de91ae-11d5-4864-8ac3-a2c87d441b0d · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Multi- agent actor-critic for mixed cooperative-competitive environments
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 968df221-589d-45a4-9dd7-3c199bdb0cb0 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Counterfactual multi-agent policy gradients
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2b0ea3d-2fdb-4560-b58e-a0181e4df166 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Value-decomposition networks for cooperative multi-agent learning based on team reward
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ac68691-e695-462b-ade7-861768e4b6ed · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review {QPLEX}: Duplex dueling multi-agent q-learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba858a96-3d44-4313-9dc9-fc19e00628a0 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Maximum entropy heterogeneous-agent reinforcement learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d6d0e12-5e14-4f16-9bf0-0e082fc43a13 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Heterogeneous-agentreinforcementlearning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9799ffa7-ab9b-4e4f-8663-bd88c8c11cef · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Metadrive: Composing diverse driving scenarios for generalizable reinforcement learning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68e57559-9b2f-4ad9-862f-7fce322b378f · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Smarts: Scalable multi-agent reinforcement learning training school for autonomous driving, 11 2020
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afb1cddd-2469-47a0-ab2d-5a77b2d1e01d · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Multi-agent reinforcement learning aided intelligent uav swarm for target tracking
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8394315-7eb7-46f0-b0b2-8ef85c7983fe · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Mean-field theory for scale-free random networks
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5db80373-f978-48a4-8a4d-761293c3d6aa · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Mean field games and mean field type control theory, volume 101
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3606999c-b135-4006-a4d4-0ccdf8d31dcb · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Optimal control of partially observable markovian systems.Journal of The Franklin Institute, 280(5):367–386, 1965
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a30baf4-9bf6-402d-a8c3-928413191b87 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Unresolved cited work
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b50b0b9-de87-4b6d-acb5-d477c186496d · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Offline Reinforcement Learning with Implicit Q-Learning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 066bd643-83c4-4ed3-bbb9-4ac20b2eaca4 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Off-policy deep reinforcement learning with- out exploration
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd90c377-709e-4c4c-8c86-e90e86655d27 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Conservative q-learning for offline reinforcement learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b23a61ee-781c-413f-9251-76f4793effce · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review A review of cooperative multi-agent deep rein- forcement learning.Applied Intelligence, 53(11):13677–13722, 2023
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d5ebb4b-2f56-4c4c-a0b4-c539cb611fa1 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review A Review of Safe Reinforcement Learning: Methods, Theory and Applications
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e8f0393-2fad-4f5b-9c8a-e71f9b3cd339 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e71f0e7-e243-44ff-9ab6-3c5b43416426 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review A Survey of Learning in Multiagent Environments: Dealing with Non-Stationarity
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7087da8f-895d-42b2-b821-96b7724e497b · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Bench- marking multi-agent deep reinforcement learning algorithms in cooperative tasks
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b8cd852-ba1b-4e12-820d-f6579cc41dcf · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Qtran: Learning to factorize with transformation for cooperative multi-agent reinforcement learning
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17c844a7-f063-4d85-b255-7d2199ea6984 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Weighted qmix: Ex- panding monotonic value function factorisation for deep multi-agent reinforcement learning
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4afce7d-2c9f-4335-9bda-d28de958d940 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review On the approximation of cooperative heterogeneous multi-agent reinforcement learning (marl) using mean field control (mfc).Journal of Machine Learning Research, 23(129):1–46, 2022
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b40a73a-daa7-4048-9d87-2312d28f4230 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Multi type mean field reinforcement learning
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1af8edd8-ce22-4ac3-b536-a821161c6de6 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Efficient model-based multi-agent mean- field reinforcement learning.Transactions on Machine Learning Research, 2021
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 615c251b-16ed-45fd-ba04-a7212c1dff1c · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Centralized Model and Exploration Policy for Multi-Agent RL
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6e99958c-96ee-4e23-8233-8f5ed1c2173f · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Model-based opponent modeling
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05fb65e1-6447-4cf5-9c30-fc7f732f9cd2 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Decentralized policy gradient descent ascent for safe multi-agent reinforcement learning
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b0c29b1-8e7a-4a0d-9a5a-ee28ed5fbf3a · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Shield decentralization for safe multi-agent reinforcement learning
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e2bbb27-78af-4e0b-8495-1fdfb8b77f7b · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Scalable primal- dual actor-critic method for safe multi-agent rl with general utilities
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88bc0658-dabb-4d83-aede-6a309bf8511e · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Multiagent planning with factored MDPs
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a87f53b-e471-4d6c-b1b9-9df15f0e0cb0 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Efficient solution algorithms for factored mdps.Journal of Artificial Intelligence Research, 19:399–468, 2003
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 392b09ff-5c2e-4593-a266-943da601bd9e · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Coordinated reinforcement learning
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 011e0078-26ba-4f09-819d-3602cd67d4b6 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Sparse cooperative q-learning
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b0fb32c-aa30-4af7-a32f-531c467cac40 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Kok and Nikos Vlassis
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4a7f225-f1ef-4cac-8873-6794cef8c31f · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Networked distributed pomdps: A synthesis of distributed constraint optimization and pomdps
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 703ab691-d170-41c1-9932-9fd21da1706b · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Approximate solutions for factored dec-pomdps with many agents
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8627cbeb-7a74-4c0c-8750-8ea9e243ae45 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Scalable reinforcement learning of localized policies for multi-agent networked systems
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f704428-c377-4535-b5a6-1b81b6069ab9 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Multi-agent reinforcement learning in stochastic networked systems.Advances in neural information processing systems, 34:7825–7837, 2021
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f873834b-6371-43e7-bb08-66b5eee707fa · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Cooperative Multi-Agent Reinforcement Learning with Hypergraph Convolution
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 84bdabf7-7355-455e-a5c9-76ee6cd09aef · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Efficient Policy Generation in Multi-Agent Systems via Hypergraph Neural Network
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ac6d326e-eccc-40b4-a831-ad126e5cc38c · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Magent: Amany-agentreinforcementlearningplatformforartificialcollectiveintelligence
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef7c0188-9e16-498d-b2b6-f31da72bf65d · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Partially observable mean field reinforcement learning
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf61d131-fffc-4ce6-9570-c4ede11667d5 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Model-Free Mean-Field Reinforcement Learning: Mean-Field MDP and Mean-Field Q-Learning
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad84a013-25ce-4147-8a77-c11f6c8a8d2b · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Mean-Field Multi-Agent Reinforcement Learning: A Decentralized Network Approach
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b2e4bd79-8882-4cce-a2a0-159e1d9a4ecc · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Swarm robotics: a review from the swarm engineering perspective.Swarm Intelligence, 7:1–41, 2013
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec4f636b-5848-4156-87fa-410f48e704c3 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Neural mmo 2.0: A massively multi-task addition to massively multi-agent learning.Advances in Neural Information Processing Systems, 36:50094–50104, 2023
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a38645d7-db43-4e74-8ec1-58c0e8138c6c · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Bsk-rl: Modular, high-fidelity reinforcement learning environments for spacecraft tasking
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6859b8f1-f144-43ab-83d2-96f18b3de7ae · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Multi-agent reinforcement learning for active voltage control on power distribution networks.Advances in Neural Information Processing Systems, 34:3271–3284, 2021
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b26e763f-8eb5-476b-8f9a-56764aff8f5f · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Updet: Universal multi-agent rl via policy decoupling with transformers
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c043b144-fdd1-454a-ac53-87b11891cea0 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Randomized entity-wise factorization for multi-agent reinforcement learning
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00c49b66-a4cd-440c-9783-ec533dddb001 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Boosting multiagent reinforcement learning via permutation invariant and permutation equivariant networks
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9bd22fd-2bb8-446d-990b-08be171f2651 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review The StarCraft Multi-Agent Challenge
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36df1842-a95a-4bdb-93c3-dc215743678e · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Cityflow: Amulti-agentreinforcementlearning environment for large scale city traffic scenario
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52a69144-0bea-4f2f-b124-1db10b013dc5 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Google research football: A novel reinforcement learning environment
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37e2b72f-7a80-44e7-91f6-0be1028f01e4 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Shaq: Incorporating shap- ley value theory into multi-agent q-learning
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 968eed5f-7fca-41b0-b009-f27a79e367e0 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Learning correlated communication topology in multi-agent reinforcement learning
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a882716-132c-496b-86b1-6eed3294d532 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Facmac: Factored multi-agent centralised policy gradients
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 187d299c-0158-4063-ad9d-51dab532fe3a · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Shapley q-value: A local reward approach to solve global reward games
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97825d7d-1f68-4e67-b6d2-acb5dbdb88aa · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Nucleolus Credit Assignment for Effective Coalitions in Multi-agent Reinforcement Learning
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 84bf7819-75ae-4946-a4bc-0d9d3a9ee72b · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Model-free mean-field reinforcement learning: mean-field mdp and mean-field q-learning
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2313382a-c76e-4846-a8df-201d1c4cfa00 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Learning to communicate with deep multi-agent reinforcement learning.Advances in neural information processing systems, 29, 2016
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a98fdd1a-6609-4746-ab49-8ab009cdc1d8 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Learning multiagent communication with backprop- agation
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 597d1900-586b-4566-a113-b6fec07fe698 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Learning structured communication for multi-agent reinforcement learning
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b479907-8072-4f78-a28e-4c5e8b2aae38 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Model-based Multi-agent Policy Optimization with Adaptive Opponent-wise Rollouts
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9801fe57-121e-478f-aa61-f1a24951ddf2 · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Offline pre-trained multi-agent decision transformer
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7606781d-9617-4c41-a680-c24168f5da4b · outbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review Hgap: boosting permutation invariant and permutation equivariant in multi-agent reinforcement learning via graph attention network
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.