Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T05:54:06.573779Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2502.03125.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T05:54:06.573779Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 48f5fa52-df87-417b-8141-0b7224d97926 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Exploration by Random Network Distillation
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 552fff4d-4135-4836-968b-14c081cbb5d4 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning The learning rate for the neural networks is uniformly set to 5 × 10−4
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1404e1d8-62a7-4700-9fcd-17ae18856a99 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Multi-agent reinforcement learning as a rehearsal for decentralized planning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e687fb5f-3da2-484f-baf3-f2e3aa1980ba · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning A review of cooperative multi-agent deep reinforcement learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5b26b8c-dd7d-4921-ac1f-8fa888b00e75 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning The StarCraft Multi-Agent Challenge
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2757e937-711f-4bd7-ad69-c80edacf0647 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Leibo, Karl Tuyls, and Thore Grae- pel
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a5b02059-4af0-4f9a-b961-bd9129a3dd41 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning [Tan and Motani, 2023] Chong Min John Tan and Mehul Motani
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 50471655-d193-4b81-a699-2f054d02c5c1 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Knowledge distillation and student-teacher learning for visual intelligence: A review and new outlooks
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a6ed0743-853c-437f-859a-e9b410c48b11 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning QPLEX: Duplex Dueling Multi-Agent Q-Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b511698-04de-4d4e-9c19-e25362176728 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Regularization-adapted anderson acceleration for multi- agent reinforcement learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 02f6a21c-6f92-4bea-ace6-b8c716b07a55 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning En- hancing collaboration in multi-agent reinforcement learn- ing with correlated trajectories
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efd6b89f-c0be-4762-8fae-508fcdbc5254 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Deep multiagent reinforce- ment learning: Challenges and directions
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8ddcc6d2-dcce-43ea-8014-c070bc7a4b80 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning A Survey on Knowledge Distillation of Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2586a3b6-04be-4ae0-a96f-be108bf7cfaa · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning A comprehensive survey on multi-agent rein- forcement learning for connected and automated vehicles
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6b275445-775f-4696-8925-8a90c405438f · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Qatten: A General Framework for Cooperative Multiagent Reinforcement Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 637f0a19-6f55-459a-be36-0952e487841a · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning The surprising effectiveness of ppo in cooperative multi-agent games
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61ddb500-7886-4255-badd-21fde4199a89 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Unmanned aerial vehicle swarm cooperative decision-making for sead mis- sion: A hierarchical multiagent reinforcement learning ap- proach
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a27919d6-9bef-4f0a-b71d-5001f75d1df3 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Ctds: Centralized teacher with decentralized student for mul- tiagent reinforcement learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 05ad1047-7ed1-43f9-a2db-1217bd85ded3 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Qdap: Downsizing adaptive policy for cooperative multi-agent reinforcement learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66fafe73-7433-4bff-830e-edc9a625f629 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Multirobot collaborative task dynamic scheduling based on multiagent reinforcement learning with heuristic graph convolution considering robot service performance
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 78f83113-7418-44f7-ad00-7baf8001470c · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Rethinking individual global max in cooperative multi- agent reinforcement learning
Reference 2015
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0f8ea0de-e950-4414-8c5b-33cbe4911f08 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning A concise introduction to decentralized POMDPs, volume
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 565be0e4-968d-4fb5-8edf-5854024a28dc · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Ptde: Personalized training with dis- tilled execution for multi-agent reinforcement learning
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 14790022-c54d-44db-892d-caad3236684c · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Qtran: Learn- ing to factorize with transformation for cooperative multi- agent reinforcement learning
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 235ba8c1-eaa0-4f1a-b396-d1547713155e · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Policy Distillation
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 702019eb-c527-456e-9f4e-35afaff8472a · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Distilling the knowledge in a neural network,
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5ada8b0b-73c4-4fd4-9799-87597cad877e · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning Rethinking the implementation tricks and monotonicity constraint in cooperative multi-agent re- inforcement learning
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 79c5c1e2-f014-41c4-b1eb-e8be8436a3b9 · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning [Huang et al., 2024] Anqi Huang, Yongli Wang, Xiaoliang Zhou, Haochen Zou, Xu Dong, and Xun Che
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 792082a7-e14b-46e9-b124-1d85147ce2ea · outbound
Double Distillation Network for Multi-Agent Reinforcement Learning [Gou et al., 2021] Jianping Gou, Baosheng Yu, Stephen J Maybank, and Dacheng Tao
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.