Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2007.09055.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T04:36:12.787317Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-08T22:35:40.699101Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 97df3bbe-b358-41e8-9e0f-eba154358607 · inbound
What Matters in Learning from Offline Human Demonstrations for Robot Manipulation Hyperparameter Selection for Offline Reinforcement Learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 503d20eb-01b0-463b-a778-cf34efdc8a83 · inbound
Simulating Errors in Touchscreen Typing Hyperparameter Selection for Offline Reinforcement Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f965292-2506-4043-9daa-6e2b00cb5c13 · inbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Hyperparameter Selection for Offline Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f69e0b2-8fea-4d05-92fa-65b796d5717b · inbound
Fully Offline Reinforcement Learning Hyperparameter Selection for Offline Reinforcement Learning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61c75cfd-8bf8-4d42-94d8-4f1ad775cd03 · inbound
Learning to Evaluate Autonomous Behaviour in Human-Robot Interaction Hyperparameter Selection for Offline Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff3857be-c107-487d-924d-b1acc5806e36 · inbound
Accelerating Detailed Routing Convergence through Offline Reinforcement Learning Hyperparameter Selection for Offline Reinforcement Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95d05ae8-6b0e-49c3-add4-668a5af5f3e9 · inbound
Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning Hyperparameter Selection for Offline Reinforcement Learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 89e47c35-418a-4152-b106-d5f27aa3e2b0 · inbound
Sample-efficient inductive matrix completion with noise and inexact side-information Hyperparameter Selection for Offline Reinforcement Learning
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ce6e5a52-5af1-444e-9dc0-b7455dfa9314 · inbound
Autoregressive Diffusion World Models for Off-Policy Evaluation of LLM Agents Hyperparameter Selection for Offline Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 61327d7e-f6e5-4e51-b616-4ddf7b986675 · inbound
Autoregressive Diffusion World Models for Off-Policy Evaluation of LLM Agents Hyperparameter Selection for Offline Reinforcement Learning
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c22ebbc-b5a0-47a6-9c31-fe65c5bccc80 · inbound
Some Essential Constructive Foundations for Systems and Control Hyperparameter Selection for Offline Reinforcement Learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 68bb7b42-0d07-4fb7-b15d-0f90912b8842 · inbound
$\text{DT}^2$: Decision-Targeted Digital Twins Hyperparameter Selection for Offline Reinforcement Learning
Reference 129
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3b6f9a96-ec65-4ca4-923b-9cc60bf59f8b · inbound
Strategic Bargaining in Multi-Buyer Markets: Reinforcement Learning from Verifiable Rewards for LLM Negotiations Hyperparameter Selection for Offline Reinforcement Learning
Reference 175
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e951cdb5-a131-4551-ba6f-3b7b55ee904c · inbound
Active Offline-to-Online Reinforcement Learning Hyperparameter Selection for Offline Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.