Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T14:42:28.270299Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:1908.02805.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T14:42:28.270299Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
58 of 58 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e15827cd-0928-4329-b202-777b519b224e · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28d695ab-85b1-4a7f-9e67-87ff82746cfb · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Learning to predict by the methods of temporal differences,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b85d0643-9bca-4e5c-aaf1-4cba95ea484a · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f097c121-60d1-48e4-9657-2d4f9196d5ed · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Residual algorithms: reinforcement learning with function approximation,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 012a80d8-3e06-4301-8e8d-96b9f264a07b · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Human-level control through deep reinforcement learning,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7fa4b313-3b5a-4079-8380-02ec11b83324 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Mastering the game of Go with deep neural networks and tree search,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 92d3b55c-5b72-454e-ac5f-94f028458fe6 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Residential energy management in smart grid: A Markov decision process-based approach,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e95ae937-a098-4317-aebf-beff5d4a8211 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multiagent reinforcement learning for urban traffic control using coordination graphs,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation fbb5bb53-4125-4ee2-b78b-e8f348687ee4 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization A distributed actor-critic algorithm and applications to mobile sensor network coordination problems,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 49766629-fc3d-454e-9cc1-9b5052ce5a04 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Reinforcement learning in robotics: A survey,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9b700b1d-9cf8-455e-b181-e9d553b1c303 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Distributed policy evaluation under multiple behavior strategies,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b2b6d37e-4d05-4f0a-bb5e-c999ed00be54 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Distributed reinforcement learning via gossip,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a3990ef9-fa7d-433a-a8e8-0f316331d07f · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Primal-dual algorithm for distributed reinforcement learning: distributed GTD,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 517665fd-0af4-4b0c-9e51-34bec6416c86 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multi-agent reinforcement learning via double averaging primal-dual optimization,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f08ca55e-3645-4e75-8322-1e38b754e677 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multi-Agent Fully Decentralized Value Function Learning with Linear Convergence Rates
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5a347f68-9944-4d32-b06b-89d9b2ca8741 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-time analysis of distributed TD(0) with linear function approximation on multi-agent reinforcement learning,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d057ba14-535b-4eea-a1e0-a34e572e6da5 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-Time Performance of Distributed Temporal Difference Learning with Linear Function Approximation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2f048e6b-ac78-4cbd-8448-70eadbe156a0 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-Sample Analysis of Decentralized Temporal-Difference Learning with Linear Function Approximation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 59102187-fb98-4eef-81bb-b01ed09342aa · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Fully Asynchronous Policy Evaluation in Distributed Reinforcement Learning over Networks
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f0da8c8e-5c83-45ba-a38c-8d6038b48645 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization An analysis of temporal-difference learning with function approximation,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4e0211f5-e45b-4b76-a0f6-468639555184 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Szepesv ´ari, Algorithms for Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 73fcb46f-c3b1-4e43-866c-c2630f4dd833 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Linear least-squares algorithms for temporal difference learning,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f90470b3-a611-4d59-91af-bc6f3ee59f2d · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization A convergent O(n) temporal-difference algorithm for off-policy learning with linear function approximation,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ac13deb6-978c-4372-8152-6ae0dd191a98 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Fast gradient-descent methods for temporal- difference learning with linear function approximation,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6f565f6f-6aa6-4de7-9fa3-e22375eeb500 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization The ode method for convergence of stochastic approximation and reinforcement learning,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 20c35761-fc08-47d0-90af-40a44c7d1cc0 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite sample analyses for TD (0) with function approximation,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6939bf3f-a554-4063-a6a5-98a4180b16b3 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite sample analysis of two-time scale stochastic approximation with applications to reinforcement learning,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 66367c32-08dd-40ce-8b9a-ab70528672e5 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Linear stochastic approximation: How far does constant step-size and iterate averaging go?
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 06eb515a-d082-401a-80c6-2b3e17f9add6 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization A finite time analysis of temporal difference learning with linear function approximation,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 95e39187-8e19-4362-8185-d6a6c666981e · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-time error bounds for linear stochastic approximation and TD learning,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 45ba3c6c-be93-46d6-b446-428f7ee9b6b2 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-sample analysis for SARSA and Q-learning with linear function approximation,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 51467f7d-61b6-49b9-8f4f-a80679da323c · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Characterizing the exact behaviors of temporal difference learning algorithms using markov jump linear system theory,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ef182439-bc13-4652-9d4d-4822cfd586f7 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-sample analysis of proximal gradient TD algorithms,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b780a0a6-e9a2-45ac-b7ea-a3f85231aff6 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Robust stochastic approximation approach to stochastic programming,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a4617d73-7f75-4c8e-b1a1-63a25af517d5 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite sample analysis of the GTD policy evaluation algorithms in Markov setting,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ac25571a-6ed9-426b-8388-51be3f3b9244 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Convergent TREE BACKUP and RETRACE with function approximation,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 73ea2fba-c87f-4977-903a-97b03dd3cf02 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multi-agent temporal-difference learning with linear function approximation: Weak convergence under time-varying network topologies,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e08e85de-756e-4c96-a584-477976830884 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Gossip algorithms: Design, analysis and applications,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1f620643-cab7-4efa-aa6c-0eb9ee15de60 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-Time Performance of Distributed Two-Time-Scale Stochastic Approximation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b2d4b097-a69b-4af9-8cd6-1134b27a7fa0 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization On the averaged stochastic approximation for linear regression,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e3bf95e0-e298-4e82-b8c1-75be7ed12b99 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Fully decentralized multi-agent reinforcement learning with networked agents,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6eb6318d-1818-47b9-8b8a-758f95e86eb9 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multi-Agent Reinforcement Learning: A Selective Overview of Theories and Algorithms
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15053b9a-4992-476a-9986-4bf793619ce4 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Decentralized Multi-Agent Reinforcement Learning with Networked Agents: Recent Advances
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 331c0120-082c-4297-934e-f0b439308d19 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Optimization for reinforcement learning: From a single agent to cooperative agents,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 96093c6c-b5ea-4845-a011-2fd34165f988 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Dual averaging for distributed optimization: Convergence analysis and network scaling,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 322c32f7-041e-45b6-a7ab-24056a097dee · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization A proximal-gradient homotopy method for the sparse least-squares problem,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2e4ff128-f5a6-4c47-98e5-4da431be0da5 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Information-theoretic lower bounds on the oracle complexity of convex optimization,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d71e7535-4f8d-4caf-a653-28f8b97eda96 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Primal-Dual Distributed Temporal Difference Learning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d20b27ad-6586-460c-9ef3-4e68192c65e8 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Ergodic mirror descent,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 44f75023-312d-4267-a165-1604f6da2615 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08f1fd21-5507-4b8b-aed7-64c44c070a52 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Distributed subgradient methods for multi-agent optimization,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e3dc3f03-1020-4ec2-b5d0-00c6b926807e · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Distributed strongly convex optimization,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0ce185fb-4ea1-4eb2-9816-287ff1cb412e · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization RSG: Beating subgradient method without smoothness and strong convexity,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e032b661-ff55-49d1-a624-f209749af128 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Homotopy smoothing for non-smooth problems with lower complexity than O (1/ϵ),
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0ed1903b-d18b-4bdb-a93c-06ad7f2944f6 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Solving non-smooth constrained programs with lower complexity than O (1/ε): A primal-dual homotopy smoothing approach,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f47f105f-00c4-48e8-8a6e-76ebdf94aca7 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Online convex programming and generalized infinitesimal gradient ascent,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 20700199-4016-4498-9da4-b56091d6d5f1 · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Subgradient methods for saddle-point problems,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 66a736dc-51d0-4107-95b9-88c0bfe1fc8f · outbound
Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Optimum bounds for the distributions of martingales in Banach spaces,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
No inbound Pith citation observations are available.