Pith. sign in

Paper Citation Record · LEDGER

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization

As of 14 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:1908.02805.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.02805 v4

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T14:42:28.270299Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact6
  • verified fuzzy46
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e15827cd-0928-4329-b202-777b519b224e · outbound

This paper cites an unresolved cited work.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T14:42:28.021233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:42:28.021233Z digest=sha256:91abae2bf11551cffecfaff2387b7798c02896dbf65b965deec53bfa89729157

Observation 28d695ab-85b1-4a7f-9e67-87ff82746cfb · outbound

This paper cites Learning to predict by the methods of temporal differences,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Learning to predict by the methods of temporal differences,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T14:42:28.026903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:42:28.026903Z digest=sha256:14fa8b96de1755ce64424e0e55a25f003dc13096e5a8bb348566bf714cf45a44

Observation b85d0643-9bca-4e5c-aaf1-4cba95ea484a · outbound

This paper cites an unresolved cited work.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:42:29.156338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.031474Z digest=sha256:c42f601bff1197c324454a09240aad3260d8356c7853b2b5d38e15869c5223d2

Observation f097c121-60d1-48e4-9657-2d4f9196d5ed · outbound

This paper cites Residual algorithms: reinforcement learning with function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Residual algorithms: reinforcement learning with function approximation,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.142212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.035942Z digest=sha256:31875fea772b8969dbf1dae281aac30e89da25e3938e7d000ad5d97bb3618b06

Observation 012a80d8-3e06-4301-8e8d-96b9f264a07b · outbound

This paper cites Human-level control through deep reinforcement learning,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Human-level control through deep reinforcement learning,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.128048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.040138Z digest=sha256:11d3634c15ed59b629b8ad0e99b405992ed77fc7369d52e97a40c3f9c09c8781

Observation 7fa4b313-3b5a-4079-8380-02ec11b83324 · outbound

This paper cites Mastering the game of Go with deep neural networks and tree search,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Mastering the game of Go with deep neural networks and tree search,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.113770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.044430Z digest=sha256:cee3c1e466da4d6c3cf1a58a97d7d84f85a501372beb220cafce870e47fb0503

Observation 92d3b55c-5b72-454e-ac5f-94f028458fe6 · outbound

This paper cites Residential energy management in smart grid: A Markov decision process-based approach,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Residential energy management in smart grid: A Markov decision process-based approach,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.098963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.048757Z digest=sha256:9b5ecb5ee16688bff95a30de7fedb239dc39a9fbb95e9ab67ba2cdcbb9d2777d

Observation e95ae937-a098-4317-aebf-beff5d4a8211 · outbound

This paper cites Multiagent reinforcement learning for urban traffic control using coordination graphs,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multiagent reinforcement learning for urban traffic control using coordination graphs,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.083452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.052409Z digest=sha256:3b96fbf0863298aa6fe88e8e22d9f67c4ee58794ad3e36f0cc5fb6a90e7e7267

Observation fbb5bb53-4125-4ee2-b78b-e8f348687ee4 · outbound

This paper cites A distributed actor-critic algorithm and applications to mobile sensor network coordination problems,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization A distributed actor-critic algorithm and applications to mobile sensor network coordination problems,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.067997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.056249Z digest=sha256:91f62f846bdd258f245e0dea94a58a3772f3a7d3c8dc91d11a4aadae00e869ec

Observation 49766629-fc3d-454e-9cc1-9b5052ce5a04 · outbound

This paper cites Reinforcement learning in robotics: A survey,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Reinforcement learning in robotics: A survey,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.053696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.060569Z digest=sha256:bb172576caee18801bf66dab1268142d5c6a5c003951663f4b604bf83bb56463

Observation 9b700b1d-9cf8-455e-b181-e9d553b1c303 · outbound

This paper cites Distributed policy evaluation under multiple behavior strategies,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Distributed policy evaluation under multiple behavior strategies,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.040159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.065089Z digest=sha256:bf7cae261f7fcd9e530e5c7cfb42d8f4be567627df58cc10fafdfdad0f3af04b

Observation b2b6d37e-4d05-4f0a-bb5e-c999ed00be54 · outbound

This paper cites Distributed reinforcement learning via gossip,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Distributed reinforcement learning via gossip,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.026525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.069673Z digest=sha256:af313ac6ba62f26cdaeee6408c19544c2cf5305f827e3782f0a5e3387f1a8506

Observation a3990ef9-fa7d-433a-a8e8-0f316331d07f · outbound

This paper cites Primal-dual algorithm for distributed reinforcement learning: distributed GTD,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Primal-dual algorithm for distributed reinforcement learning: distributed GTD,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.012073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.073776Z digest=sha256:ce4f8082ed33ee13c18d0697d03bc1dda9b726ddaa0c8aeef913abde5d00ad67

Observation 517665fd-0af4-4b0c-9e51-34bec6416c86 · outbound

This paper cites Multi-agent reinforcement learning via double averaging primal-dual optimization,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multi-agent reinforcement learning via double averaging primal-dual optimization,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.997808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.078023Z digest=sha256:7926c01ec0b4e201cee198741d42f0cc8bae4ca8d57fa2a1a3f0486a2c5b0e53

Observation f08ca55e-3645-4e75-8322-1e38b754e677 · outbound

This paper cites Multi-Agent Fully Decentralized Value Function Learning with Linear Convergence Rates.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multi-Agent Fully Decentralized Value Function Learning with Linear Convergence Rates

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:42:28.457039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.082359Z digest=sha256:53704762f43c14d4f94645eca5994280e54bc7fbe4a02e942346f3e069d7e026

Observation 5a347f68-9944-4d32-b06b-89d9b2ca8741 · outbound

This paper cites Finite-time analysis of distributed TD(0) with linear function approximation on multi-agent reinforcement learning,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-time analysis of distributed TD(0) with linear function approximation on multi-agent reinforcement learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.983791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.086840Z digest=sha256:4e31f5fa6b242244dd80c26c4082094acd8e81864f03b536e3dd530bd8df4d08

Observation d057ba14-535b-4eea-a1e0-a34e572e6da5 · outbound

This paper cites Finite-Time Performance of Distributed Temporal Difference Learning with Linear Function Approximation.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-Time Performance of Distributed Temporal Difference Learning with Linear Function Approximation

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:42:28.436052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.091067Z digest=sha256:8843dc8c8472bebd2b4f6723470973c3e2447b44f756fb38c3a429dcc83bf56f

Observation 2f048e6b-ac78-4cbd-8448-70eadbe156a0 · outbound

This paper cites Finite-Sample Analysis of Decentralized Temporal-Difference Learning with Linear Function Approximation.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-Sample Analysis of Decentralized Temporal-Difference Learning with Linear Function Approximation

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:42:28.415808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.095985Z digest=sha256:b95630b0592890ae280dc484fcebef8647d1baad4af916175814e360dc9664e4

Observation 59102187-fb98-4eef-81bb-b01ed09342aa · outbound

This paper cites Fully Asynchronous Policy Evaluation in Distributed Reinforcement Learning over Networks.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Fully Asynchronous Policy Evaluation in Distributed Reinforcement Learning over Networks

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:42:28.392356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.100700Z digest=sha256:a4baa54f576480728df12048e60454fa9aa7e8dc3c1959756aa1bfd3f8d79ce1

Observation f0da8c8e-5c83-45ba-a38c-8d6038b48645 · outbound

This paper cites An analysis of temporal-difference learning with function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization An analysis of temporal-difference learning with function approximation,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.969568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.105575Z digest=sha256:28fce776a8f6d0d25f696b27e164dc598fd8b77ff7aa8f790602e33d11dd45b8

Observation 4e0211f5-e45b-4b76-a0f6-468639555184 · outbound

This paper cites Szepesv ´ari, Algorithms for Reinforcement Learning.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Szepesv ´ari, Algorithms for Reinforcement Learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.954808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.110004Z digest=sha256:1f0308919a398d86f6ad14eb45b9af48559a0e4912462e38d6e9b63606eab26c

Observation 73fcb46f-c3b1-4e43-866c-c2630f4dd833 · outbound

This paper cites Linear least-squares algorithms for temporal difference learning,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Linear least-squares algorithms for temporal difference learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.940962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.114330Z digest=sha256:f4df6b7a22dc72ddd3b42349f50ffc81d741b90701339e400793f1a744b62a60

Observation f90470b3-a611-4d59-91af-bc6f3ee59f2d · outbound

This paper cites A convergent O(n) temporal-difference algorithm for off-policy learning with linear function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization A convergent O(n) temporal-difference algorithm for off-policy learning with linear function approximation,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.927720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.118634Z digest=sha256:f0eba776f81f8e5fe6d7af357bf8012ecb6cfaed158345795110a95f5bca7299

Observation ac13deb6-978c-4372-8152-6ae0dd191a98 · outbound

This paper cites Fast gradient-descent methods for temporal- difference learning with linear function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Fast gradient-descent methods for temporal- difference learning with linear function approximation,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.914294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.122954Z digest=sha256:7f945e9a0f2742a3aeb72df080781ee68535e7062dc3269a6f9186a4e0ba9ad0

Observation 6f565f6f-6aa6-4de7-9fa3-e22375eeb500 · outbound

This paper cites The ode method for convergence of stochastic approximation and reinforcement learning,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization The ode method for convergence of stochastic approximation and reinforcement learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.900843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.127491Z digest=sha256:fe24bcfc839d4384d7baf898163dfb6642481cf70cc7991fc80fc954e9346bf6

Observation 20c35761-fc08-47d0-90af-40a44c7d1cc0 · outbound

This paper cites Finite sample analyses for TD (0) with function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite sample analyses for TD (0) with function approximation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.884765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.132222Z digest=sha256:a2c9a4038b11d4fb462b5962b6519634c969da84290bbde37214908346c46d51

Observation 6939bf3f-a554-4063-a6a5-98a4180b16b3 · outbound

This paper cites Finite sample analysis of two-time scale stochastic approximation with applications to reinforcement learning,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite sample analysis of two-time scale stochastic approximation with applications to reinforcement learning,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.868871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.136958Z digest=sha256:1913c9e09da83c6d8b83225de412a8919ec0debe04875839bcff8dc672af28fc

Observation 66367c32-08dd-40ce-8b9a-ab70528672e5 · outbound

This paper cites Linear stochastic approximation: How far does constant step-size and iterate averaging go?.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Linear stochastic approximation: How far does constant step-size and iterate averaging go?

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.853601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.141580Z digest=sha256:e0d040caffa1d316e7f8d356293a9036b066eb9e0be082bc76713dc371ce30c1

Observation 06eb515a-d082-401a-80c6-2b3e17f9add6 · outbound

This paper cites A finite time analysis of temporal difference learning with linear function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization A finite time analysis of temporal difference learning with linear function approximation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.839187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.145862Z digest=sha256:0bab406ee9d71883846dfa247b597ce96dd5e56f58c25c1eabc620a2cca85a9b

Observation 95e39187-8e19-4362-8185-d6a6c666981e · outbound

This paper cites Finite-time error bounds for linear stochastic approximation and TD learning,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-time error bounds for linear stochastic approximation and TD learning,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.825063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.150011Z digest=sha256:1733f601c47b6c2662d44db268639daae3a967d88a1b5a4fe737f3de43fba820

Observation 45ba3c6c-be93-46d6-b446-428f7ee9b6b2 · outbound

This paper cites Finite-sample analysis for SARSA and Q-learning with linear function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-sample analysis for SARSA and Q-learning with linear function approximation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.810675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.154518Z digest=sha256:c7847e130256c328c0d7dea42dc1e26752d10ef42702868b8a134807be1dc118

Observation 51467f7d-61b6-49b9-8f4f-a80679da323c · outbound

This paper cites Characterizing the exact behaviors of temporal difference learning algorithms using markov jump linear system theory,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Characterizing the exact behaviors of temporal difference learning algorithms using markov jump linear system theory,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.794956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.158742Z digest=sha256:5631bb36d8bb15b8c690949b83268ddd7e399e0572bcfb47afad7953758bb7c5

Observation ef182439-bc13-4652-9d4d-4822cfd586f7 · outbound

This paper cites Finite-sample analysis of proximal gradient TD algorithms,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-sample analysis of proximal gradient TD algorithms,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.779789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.163018Z digest=sha256:a9b865084c66376adf0f1c52a3d2f0542f7a72521fbdc58f74fec3053a6282c9

Observation b780a0a6-e9a2-45ac-b7ea-a3f85231aff6 · outbound

This paper cites Robust stochastic approximation approach to stochastic programming,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Robust stochastic approximation approach to stochastic programming,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.764325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.166628Z digest=sha256:7b15dc0cc91e3781dfc436f2bf2ed6851723bfd7a376f1e878ff416f3574f94d

Observation a4617d73-7f75-4c8e-b1a1-63a25af517d5 · outbound

This paper cites Finite sample analysis of the GTD policy evaluation algorithms in Markov setting,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite sample analysis of the GTD policy evaluation algorithms in Markov setting,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.751541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.170230Z digest=sha256:fc01f70fcadd94f82ba8ad689af0b3f5520d29f3f54f26cb6635a33a3dbf9dbc

Observation ac25571a-6ed9-426b-8388-51be3f3b9244 · outbound

This paper cites Convergent TREE BACKUP and RETRACE with function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Convergent TREE BACKUP and RETRACE with function approximation,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.738015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.174649Z digest=sha256:27a5a58c0818f894a36e496891e0686aee4793f06bc3595b93d8399f321b08f3

Observation 73ea2fba-c87f-4977-903a-97b03dd3cf02 · outbound

This paper cites Multi-agent temporal-difference learning with linear function approximation: Weak convergence under time-varying network topologies,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multi-agent temporal-difference learning with linear function approximation: Weak convergence under time-varying network topologies,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.723701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.178323Z digest=sha256:58d87f9bbfb65699d997b531551402ea67c0fafe1f91c9e2fb3a2ddc24c98ce7

Observation e08e85de-756e-4c96-a584-477976830884 · outbound

This paper cites Gossip algorithms: Design, analysis and applications,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Gossip algorithms: Design, analysis and applications,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.708438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.182268Z digest=sha256:66e6810aab5e2bb2004f5761df4bd245f7c16ec6acc6697b2b3b0eca89386d83

Observation 1f620643-cab7-4efa-aa6c-0eb9ee15de60 · outbound

This paper cites Finite-Time Performance of Distributed Two-Time-Scale Stochastic Approximation.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-Time Performance of Distributed Two-Time-Scale Stochastic Approximation

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:42:28.371015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.187587Z digest=sha256:565ac44257982aeceb44ec45d10268ca55cd6ed60abcb6275c3919e681be2785

Observation b2d4b097-a69b-4af9-8cd6-1134b27a7fa0 · outbound

This paper cites On the averaged stochastic approximation for linear regression,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization On the averaged stochastic approximation for linear regression,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.693809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.191638Z digest=sha256:1dc677fb08a595c717f0dbce62c046a3fd791516eb3b8d9d57f65dc397454489

Observation e3bf95e0-e298-4e82-b8c1-75be7ed12b99 · outbound

This paper cites Fully decentralized multi-agent reinforcement learning with networked agents,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Fully decentralized multi-agent reinforcement learning with networked agents,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.678415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.195009Z digest=sha256:076fb54982b55f51499f3dc0953d90a5ca006977ca8461e62b58002d72a3a5bb

Observation 6eb6318d-1818-47b9-8b8a-758f95e86eb9 · outbound

This paper cites Multi-Agent Reinforcement Learning: A Selective Overview of Theories and Algorithms.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multi-Agent Reinforcement Learning: A Selective Overview of Theories and Algorithms

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-14T14:42:28.199222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:42:28.199222Z digest=sha256:9bc33d81f038f631e8001041fb78c1fa4abdfb80e09fc86b6cdebad578b0c11d

Observation 15053b9a-4992-476a-9986-4bf793619ce4 · outbound

This paper cites Decentralized Multi-Agent Reinforcement Learning with Networked Agents: Recent Advances.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Decentralized Multi-Agent Reinforcement Learning with Networked Agents: Recent Advances

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:42:28.332117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.204080Z digest=sha256:a719d16db545b692d5b008ffe0a66010056f79fa44877c25b18379b1aa80f89d

Observation 331c0120-082c-4297-934e-f0b439308d19 · outbound

This paper cites Optimization for reinforcement learning: From a single agent to cooperative agents,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Optimization for reinforcement learning: From a single agent to cooperative agents,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.661986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.208814Z digest=sha256:2f7563f5c5f717c1950b6bacb5668b4eec4cb5a3edbb9a46fb4fa26dc99792c4

Observation 96093c6c-b5ea-4845-a011-2fd34165f988 · outbound

This paper cites Dual averaging for distributed optimization: Convergence analysis and network scaling,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Dual averaging for distributed optimization: Convergence analysis and network scaling,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.646492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.212999Z digest=sha256:b9b2cfb6e07b62db3fa258f8ad124bc3913c668ef014e5ad11ac22c54c95c708

Observation 322c32f7-041e-45b6-a7ab-24056a097dee · outbound

This paper cites A proximal-gradient homotopy method for the sparse least-squares problem,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization A proximal-gradient homotopy method for the sparse least-squares problem,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.629867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.217396Z digest=sha256:a25b6fc685125b7e1c4018c4fda0d20894fb1058229edeb048d66643fa16cbf6

Observation 2e4ff128-f5a6-4c47-98e5-4da431be0da5 · outbound

This paper cites Information-theoretic lower bounds on the oracle complexity of convex optimization,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Information-theoretic lower bounds on the oracle complexity of convex optimization,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.615945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.221610Z digest=sha256:8e7ec26a440b83ec12bd76750532443b1c0132a17d59d6aa2a1d15f496e29ba6

Observation d71e7535-4f8d-4caf-a653-28f8b97eda96 · outbound

This paper cites Primal-Dual Distributed Temporal Difference Learning.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Primal-Dual Distributed Temporal Difference Learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-14T14:42:28.226368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:42:28.226368Z digest=sha256:0f05f636c6fc3c8ec7f96a7db546c7d8723ce19f67bc04bee0d2dadddbbefb74

Observation d20b27ad-6586-460c-9ef3-4e68192c65e8 · outbound

This paper cites Ergodic mirror descent,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Ergodic mirror descent,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.600588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.230720Z digest=sha256:0006eb9655c99286bc7bf087b1e78f6cd5602d72a5e883ee39593828ec2b479b

Observation 44f75023-312d-4267-a165-1604f6da2615 · outbound

This paper cites an unresolved cited work.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-14T14:42:28.234977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:42:28.234977Z digest=sha256:fb58f4b529ae3bbede9d32a100b794e7c5427744c1768168595d90a6a7dea0ce

Observation 08f1fd21-5507-4b8b-aed7-64c44c070a52 · outbound

This paper cites Distributed subgradient methods for multi-agent optimization,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Distributed subgradient methods for multi-agent optimization,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.574614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.239573Z digest=sha256:76bc75cb283a27bb2eb125b373d8d614b1d2b764b5bbca3f783e651ec8b01548

Observation e3dc3f03-1020-4ec2-b5d0-00c6b926807e · outbound

This paper cites Distributed strongly convex optimization,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Distributed strongly convex optimization,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.560560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.244768Z digest=sha256:a647bc430b36e3fff25bf0eefa60b73510e48e951836d6344ca40cccbadc6a01

Observation 0ce185fb-4ea1-4eb2-9816-287ff1cb412e · outbound

This paper cites RSG: Beating subgradient method without smoothness and strong convexity,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization RSG: Beating subgradient method without smoothness and strong convexity,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.545976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.249153Z digest=sha256:8c89f64ba74218a3d5b6d4c3e9361f2e019c2bf3840586a843b736d58b5f0fa6

Observation e032b661-ff55-49d1-a624-f209749af128 · outbound

This paper cites Homotopy smoothing for non-smooth problems with lower complexity than O (1/ϵ),.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Homotopy smoothing for non-smooth problems with lower complexity than O (1/ϵ),

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.530768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.253551Z digest=sha256:33a3faeb73d191f6e377d9b69549f7493d5ed307804172816ae055512b48ca9d

Observation 0ed1903b-d18b-4bdb-a93c-06ad7f2944f6 · outbound

This paper cites Solving non-smooth constrained programs with lower complexity than O (1/ε): A primal-dual homotopy smoothing approach,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Solving non-smooth constrained programs with lower complexity than O (1/ε): A primal-dual homotopy smoothing approach,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.515981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.257995Z digest=sha256:27a96aa0065ae92f9de576e027185cfaddacc0b16b6f5cb854a84e9eb71b90a1

Observation f47f105f-00c4-48e8-8a6e-76ebdf94aca7 · outbound

This paper cites Online convex programming and generalized infinitesimal gradient ascent,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Online convex programming and generalized infinitesimal gradient ascent,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.500723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.261827Z digest=sha256:96af8e99c5c29f3aacebbd1c6b1105bc9bc77fe31a81520a0e83a8cbe8b28e84

Observation 20700199-4016-4498-9da4-b56091d6d5f1 · outbound

This paper cites Subgradient methods for saddle-point problems,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Subgradient methods for saddle-point problems,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.486704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.265746Z digest=sha256:87a53bb4a4e6a2e946e6349bae1eb80bb8b2b5fac52660b005a89f4509185091

Observation 66a736dc-51d0-4107-95b9-88c0bfe1fc8f · outbound

This paper cites Optimum bounds for the distributions of martingales in Banach spaces,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Optimum bounds for the distributions of martingales in Banach spaces,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.472142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:42:28.270299Z digest=sha256:8d4fb0b46d9a14ebe3749aa46329bfdf4af998f3155b3eab001e2f9468329add

Pith citing papers

No inbound Pith citation observations are available.