Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2308.12270.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T11:39:56.118369Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T12:38:07.182042Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation b5079cf3-34dc-4d7f-bafa-83893cb2a5af · inbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Language Reward Modulation for Pretraining Reinforcement Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bda703ef-bb9d-483f-997b-2915c27a2829 · inbound
Preference VLM: Leveraging VLMs for Scalable Preference-Based Reinforcement Learning Language Reward Modulation for Pretraining Reinforcement Learning
Reference 2004
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9ad4aef-6a45-4b4e-a471-bad9b521ad4b · inbound
World Model Implanting for Test-time Adaptation of Embodied Agents Language Reward Modulation for Pretraining Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1408bd9c-6cc0-4a5c-976e-adf6f535fd0e · inbound
Exploratory Retrieval-Augmented Planning For Continual Embodied Instruction Following Language Reward Modulation for Pretraining Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64bcc828-ea4b-4093-93a6-e0bc5277943e · inbound
Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons Language Reward Modulation for Pretraining Reinforcement Learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c7651e20-adac-4601-9eb8-77299d18c095 · inbound
CoRe: Combined Rewards with Vision-Language Model Feedback for Preference-Aligned Reinforcement Learning Language Reward Modulation for Pretraining Reinforcement Learning
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5b32704d-960c-4be5-b95b-8601d7dc8a62 · inbound
Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Language Reward Modulation for Pretraining Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.