Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T22:59:50.876652Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2505.05968.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T22:59:50.876652Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
63 of 63 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2557e51e-a8d3-413e-9784-23a30e8223bc · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Principal component analysis
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7653a9b4-f54d-4df9-9924-d38b20b46270 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition A Model-Based Solution to the Offline Multi-Agent Reinforcement Learning Coordination Problem
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a2f15c44-b325-4109-936d-b3c5440dbf8d · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Dota 2 with Large Scale Deep Reinforcement Learning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea42da4e-c88c-42a7-b134-358f45640970 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition The complexity of decentralized control of markov decision processes
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a5a7f5d3-7d74-45f9-b049-6328a4a2c595 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition ComaDICE: Offline Cooperative Multi-Agent Reinforcement Learning with Stationary Distribution Shift Regularization
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 00393317-47cd-4435-957c-500b7422bc3b · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Powernet: Multi-agent deep reinforcement learning for scalable powergrid control
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c03b083a-0ceb-4ed8-abef-b93a7077cd91 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Score regularized policy optimization through diffusion behavior
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2a856e86-3d4e-458e-bb3b-16110f8b7eac · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb4cca48-bf70-41fa-8346-88bc0883e05e · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Consistency Models as a Rich and Efficient Policy Class for Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14dfefd4-2b6d-41fb-b68c-debe33a3db79 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Putting Data at the Centre of Offline Multi-Agent Reinforcement Learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f86591a4-e4da-4b67-a624-57c5b97389f2 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Off-the-grid marl: Datasets and baselines for offline multi-agent reinforcement learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 573b543a-2e51-4a1a-8a2c-b31233752b23 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Dis- pelling the mirage of progress in offline marl through standardised baselines and evaluation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cb9f4b0b-1b31-471f-a05b-5bc1c5015d0f · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Henriques
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a520a7d7-d5da-49c3-ac04-faba70847c67 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Rein- forcement learning-based consensus reaching in large-scale social networks
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ac1f2933-4fef-426b-aa4c-07f5f4a472fe · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Dynamic programming for partially observable stochastic games
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b5693bad-8ae4-4d43-abe8-f2837e5e62f2 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3ac3c8b-7ac8-473a-88c7-fa6d9eaf7cfb · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Denoising diffusion probabilistic models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c025c414-851f-41b3-b217-08c388c2aa64 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Actor-attention-critic for multi-agent reinforcement learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96e22250-89ad-46ca-9d6d-3c89db93c8c0 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Offline Decentralized Multi-Agent Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50f6e36f-f3ae-4b6c-9d0f-4af051ac523a · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Efficient diffusion policies for offline reinforcement learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 357029c6-3312-4851-a72d-ffb045cd2a71 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Offline Reinforcement Learning with Implicit Q-Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bcbc435-866e-4fa3-9c06-fa1f94c1ebe3 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Trust region policy optimisation in multi-agent reinforcement learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6765777d-6c39-478e-99ad-79bab16f3ea3 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da8636cf-aef0-465d-9924-07ff3de1a78e · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Dof: A diffusion factorization framework for offline multi-agent reinforcement learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5dc0bb25-9d7a-4249-9024-544558c00531 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Improving Generalization and Data Efficiency with Diffusion in Offline Multi-agent RL
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e6c36f2-5d08-452d-a05b-c98dddedf5d5 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition A kernelized stein discrepancy for goodness-of-fit tests
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b9ada247-d453-46bb-a675-4764fe349322 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Offline multi-agent reinforcement learning via in-sample sequential policy optimization
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37ba1b02-ed87-49c2-a356-318bc6324c64 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Multi-agent actor-critic for mixed cooperative-competitive environments
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0615b0d0-235a-48a3-8ca1-0c61ebc4179b · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Efficient and scalable reinforcement learning for large-scale network control
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c03c630a-e96d-48ba-bd9a-5332322fe7d3 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Learning to coordinate from offline datasets with uncoordinated behavior policies
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c388c804-e5f4-456b-aaf6-2bd70cb16212 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Dynamic economic emissions dispatch optimisation using multi-agent reinforcement learning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1da9306d-1e1b-4b3f-bee0-44382af1c89d · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Diffusion-DICE: In-Sample Diffusion Guidance for Offline Reinforcement Learning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71f12ca2-e6d5-4a04-8ff2-318d5312c190 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition AlberDICE: Addressing Out-Of-Distribution Joint Actions in Offline Multi-Agent RL via Alternating Stationary Distribution Correction Estimation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d39dea81-78f9-4b71-bdc1-efafbffcc13c · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Offline Pre-trained Multi-Agent Decision Transformer: One Big Sequence Model Tackles All SMAC Tasks
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d194515-c6e7-4869-8a5d-7791e3210820 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 260eb825-7af6-4704-ac99-d38c03e6ac8c · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Plan better amid conservatism: Offline multi-agent reinforcement learning with actor rectification
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a85312c7-bba3-429e-9846-996f0f272feb · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Facmac: Factored multi-agent centralised policy gradients
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c6bdc2e9-f4ee-4db5-b32a-9f7db8e1aede · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition A survey on offline reinforcement learning: Taxonomy, review, and open problems
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5978b74e-ac05-4102-a484-f603438d68fe · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Monotonic value function factorisation for deep multi-agent reinforcement learning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9180358f-4391-4bbb-a6ce-28c01ca85939 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Counterfactual Conservative Q Learning for Offline Multi-agent Reinforcement Learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 440ee5e2-2047-4642-a67b-565297168f08 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Deep unsuper- vised learning using nonequilibrium thermodynamics
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2bca209-df4e-45d0-9489-0acdc0c0eeeb · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Generative modeling by estimating gradients of the data distribution.yang-song.net, May 2021
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 424fe638-8309-4658-b87b-c02f3a47c5b9 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Score-Based Generative Modeling through Stochastic Differential Equations
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa912955-c13d-494d-bb8a-bfa64493113a · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Arena: A general evaluation platform and building toolkit for multi-agent intelligence
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2b10dfcf-8119-432a-bcc2-09d1048aa270 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Corl: Research-oriented deep offline reinforcement learning library
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8994c3c3-e1a8-43f0-b9d4-115e116cd291 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Learning from good trajectories in offline multi-agent reinforcement learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3e5bd0d7-6e9f-4856-98ad-9edccc9b9a6b · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Coordination Failure in Cooperative Offline MARL
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5333fee4-92a2-4ebd-a56e-7aee2cdbc5c1 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Offline multi- agent reinforcement learning with knowledge distillation
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2cbc3ce4-6ade-432a-ad7b-68f8c16a4c44 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Visualizing data using t-sne
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3af2ea25-f153-45ff-976c-f7df1b6591f2 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Offline multi-agent reinforce- ment learning with implicit global-to-local value regularization
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0cf2474d-2b95-47c3-8f6f-317b9b3cd0e1 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Order Matters: Agent-by-agent Policy Optimization
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 190a48fe-a9cf-42f3-b51f-360084e11a24 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Diffusion policies as an expressive policy class for offline reinforcement learning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c48582e8-3d0b-4cba-8152-6d636294c178 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition MACCA: Offline Multi-agent Reinforcement Learning with Causal Credit Assignment
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa076a28-958a-43c8-93a7-518dbd173854 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Multi-agent reinforcement learning is a sequence modeling problem
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a95ce8f-2087-49cd-8691-632001322ab2 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Behavior Regularized Offline Reinforcement Learning
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96f34187-5050-49fd-8303-b587aa7aec38 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Game-Theoretic Multiagent Reinforcement Learning
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12eadc7e-a6bc-4888-af22-58cfc68e34ca · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Believe what you see: Implicit constraint approach for offline multi-agent reinforcement learning
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2d2f3544-5939-4d31-a95a-8cc75e2cd10d · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Multi-agent reinforcement learning: A selective overview of theories and algorithms
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e331b35d-80f7-4c13-a4a5-76714090dad9 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Fop: Factorizing optimal joint policy of maximum-entropy multi-agent reinforcement learning
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e18fa33e-eba4-4f3c-b0e6-c78371fa46ec · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition The AI Economist: Improving Equality and Productivity with AI-Driven Tax Policies
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de72b18a-441d-4e2c-bf6b-5805072b165b · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition guide-then-select
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 561a1827-18f8-4ce6-a718-760e318fd1d7 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e1c3942d-e424-488b-b4fc-a8ba2f9b7fb8 · outbound
Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition The key hyperparameters for OMSD are summarized in Table 3
Reference 512
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
No inbound Pith citation observations are available.