Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:03:21.469609Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 3 inbound Pith citation observations for arXiv:2506.01052.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:03:21.469609Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-27T23:43:27.932309Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-03T23:29:02.034641Z
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1d5643d9-2025-4b22-9964-92ad4c933866 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation A finite time analysis of temporal difference learning with linear function approximation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ec333438-c1ed-4c55-b275-bc94098c4c92 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Deepdriving: Learning affordance for direct perception in autonomous driving
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7bd7d890-9da1-41fa-85d6-530cb310fb9b · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Cutkosky and Francesco Orabona
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fb67af19-7fe8-4d7d-9fec-36f5bb844bc7 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Finite sample analysis of two-timescale stochastic approximation with applications to reinforcement learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9cb3a3f7-445c-465b-8da7-cf2debe6104b · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Logarithmic Sobolev inequalities for finite Markov chains
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d564c3d4-6ac3-42a7-8cd8-2e3b0f6a4abb · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 471abd51-65e4-4fd7-bbfe-277c76423a81 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation DoG is SGD 's best friend: A parameter-free dynamic step size schedule
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ff5df0d8-eb78-42cc-8c04-56e2ec90063b · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation On TD (0) with function approximation: Concentration bounds and a centered variant with exponential convergence
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 47003276-022e-4df4-808f-7b7d96485781 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Stochastic approximation: a survey
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1c95acc9-e3c2-4b24-932a-9deb21d9f9ad · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Linear stochastic approximation: How far does constant step-size and iterate averaging go? In International conference on artificial intelligence and statistics, pages 1347--1355
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5497b0d8-739c-46cc-8cb7-49e1d57eb3c7 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Markov chains and mixing times, volume 107
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01a6ae64-ac52-43fb-8979-810e7d79c305 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Temporal difference learning as gradient splitting
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 82719736-fc3b-4a88-9abb-063adc43b8ef · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Reinforcement Learning: Foundations
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4e820a99-e309-4ebe-bf3d-7a0fce6df899 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation A simple finite-time analysis of TD learning with linear function approximation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3856d66e-0213-4226-b285-cf6cf4cc1925 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Approximate Temporal Difference Learning is a Gradient Descent for Reversible Policies
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0553c78a-6c65-4b45-86b8-ca08b2f7abe0 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Parameter-free Stochastic Optimization of Variationally Coherent Functions
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64eb2c5d-7ba6-4fd2-9eb7-e224ef1934c0 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Mastering the game of go with deep neural networks and tree search
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 59e42b4e-b53a-4c23-b908-54615f8be824 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Finite-time error bounds for linear stochastic approximation and td learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f01bc406-cc4d-4117-9bd7-f630f9287375 · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Learning to predict by the methods of temporal differences
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c1d018b-922a-4076-8883-a4599b8389ae · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Analysis of temporal-diffference learning with function approximation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ffd0b19-b5d0-4ba1-8a6a-ec3367bdd4dd · outbound
A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4896404f-231f-44ae-8383-243f7c9a77c6 · inbound
Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2a3fd8fd-90a7-4e5a-903a-76b475ae6caf · inbound
Fast and Robust Convergence Rate for TD(0) with Linear Function Approximation, Universal Learning Steps and I.I.D. Samples A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0ef1ca93-3dd2-41be-8378-60cf046f8f3e · inbound
A Diffusion Approximation for Temporal-Difference Learning with Linear Features under Markovian Noise A Robust $\widetilde{\mathcal{O}}(1/\sqrt{T})$ Rate for Unprojected TD Learning with Linear Function Approximation
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.