Pith. sign in

Paper Citation Record · LEDGER

Temporal Difference Learning for Model Predictive Control

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 41 inbound Pith citation observations for arXiv:2203.04955.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2203.04955 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 41 of 41 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:46:58.397406Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:00:06.256105Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b86049ce-e128-480b-999e-31cd066c7fcb · inbound

DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning cites this paper.

DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning Temporal Difference Learning for Model Predictive Control

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-17T16:06:09.742215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-17T16:06:09.448517Z digest=sha256:07823c1177c51ffd0bcb0e83f53726a9fa4c3e3f1629b1444d7db424eb592d44

Observation eab4e933-d8a5-4a34-a61d-f518ae8a57d6 · inbound

WoMAP: World Models For Embodied Open-Vocabulary Object Localization cites this paper.

WoMAP: World Models For Embodied Open-Vocabulary Object Localization Temporal Difference Learning for Model Predictive Control

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:46:58.397406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:46:58.397406Z digest=sha256:3d5f11fd040f8eba5c4d43e3d0e5f6bb2956a4fa9fedef84bffb58c80b8c7cf0

Observation a750b801-2ecd-4803-8de8-21424d59e889 · inbound

BiTrajDiff: Bidirectional Trajectory Generation with Diffusion Models for Offline Reinforcement Learning cites this paper.

BiTrajDiff: Bidirectional Trajectory Generation with Diffusion Models for Offline Reinforcement Learning Temporal Difference Learning for Model Predictive Control

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:52:15.261991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T10:48:28.980868Z digest=sha256:8abdaffe9a851ecb40058afff49171c1f81211e30901b4bb32c3577475645c58

Observation ae0cd45e-f4f7-48f4-b121-c83b93de9fc5 · inbound

Real-Time Execution of Action Chunking Flow Policies cites this paper.

Real-Time Execution of Action Chunking Flow Policies Temporal Difference Learning for Model Predictive Control

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:18:51.670797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T14:18:51.613045Z digest=sha256:de71c86360248619dc093bca603b7b36e186871b5e54101c154e9cc1acae48e0

Observation 5754e964-70a6-4abd-a996-3b33939c6594 · inbound

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning cites this paper.

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning Temporal Difference Learning for Model Predictive Control

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:33:50.866977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T00:33:50.471804Z digest=sha256:675fd86cb2e51a093136f43b6e19fed9804e9fafc1e8002fcc8c2f3cdf6b4189

Observation 1a33e769-05e3-4b25-9260-985923d1599b · inbound

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning cites this paper.

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning Temporal Difference Learning for Model Predictive Control

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:17:14.161907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T09:15:45.511104Z digest=sha256:dbbb3b3b90867fdf4c5db62bf2479cd03cb28959638b4009cd39ad9b6c94f386

Observation 2499a6b0-ac18-4471-ab79-133415da982d · inbound

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond cites this paper.

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond Temporal Difference Learning for Model Predictive Control

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:22.074261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:22.074261Z digest=sha256:877d62a04af71f05d56a87fbc2f3bc07228f77cfdf1ae73d43195c31ce554b87

Observation ffd82329-753f-467b-93d4-035e47ca5a57 · inbound

Investigating Lagrangian Neural Networks for Infinite Horizon Planning in Quadrupedal Locomotion cites this paper.

Investigating Lagrangian Neural Networks for Infinite Horizon Planning in Quadrupedal Locomotion Temporal Difference Learning for Model Predictive Control

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:28.994599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:28.994599Z digest=sha256:20ee135c3f1961c2938503951f2b42c95ea30e27a29eec57649baa5efd23c4b9

Observation 758f57f7-b5a3-47c9-8015-bdaf72ee23d4 · inbound

M3PO: Massively Multi-Task Model-Based Policy Optimization cites this paper.

M3PO: Massively Multi-Task Model-Based Policy Optimization Temporal Difference Learning for Model Predictive Control

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:48.993986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:48.993986Z digest=sha256:edb2637690a7f4048758e30c4f29e56d82c0361562def970ae22483b47251890

Observation b20647d5-91c5-4072-a4c9-6280a9e409cc · inbound

Sample-Efficient Reinforcement Learning Controller for Deep Brain Stimulation in Parkinson's Disease cites this paper.

Sample-Efficient Reinforcement Learning Controller for Deep Brain Stimulation in Parkinson's Disease Temporal Difference Learning for Model Predictive Control

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T19:12:58.596909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:12:58.596909Z digest=sha256:465e9637aabe1dc518b900d01850df55279deb72b75be187f6535a63b1de8dac

Observation 6539a871-cc50-41c3-9916-4241e83ac012 · inbound

Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution cites this paper.

Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution Temporal Difference Learning for Model Predictive Control

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T23:06:33.769454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:06:33.769454Z digest=sha256:89a13fb1243164d3fb2433bfd6575e6bb82e662b5581dc90d7a16f9e077f7623

Observation a0943cec-78f7-44e1-958e-ca6abdf9043f · inbound

Arnold: a generalist muscle transformer policy cites this paper.

Arnold: a generalist muscle transformer policy Temporal Difference Learning for Model Predictive Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T16:41:53.486675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:41:53.486675Z digest=sha256:3815997fb836d7f3b047ebe584a9cc34ea7b79b87057925f0d0ad58528846ca5

Observation 7b387953-eda6-4620-9c97-6611d230f78c · inbound

High-Precision and High-Efficiency Trajectory Tracking for Excavators Based on Closed-Loop Dynamics cites this paper.

High-Precision and High-Efficiency Trajectory Tracking for Excavators Based on Closed-Loop Dynamics Temporal Difference Learning for Model Predictive Control

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-18T15:31:33.478521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T15:30:19.253933Z digest=sha256:b87736af6312e411cdb07ca3052a7a24184adc88f5ddf6bfb40ee695a4db0da1

Observation b0bb3a76-b285-49d5-9221-c98dfe250bd5 · inbound

Model-Based Reinforcement Learning under Random Observation Delays cites this paper.

Model-Based Reinforcement Learning under Random Observation Delays Temporal Difference Learning for Model Predictive Control

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:36:28.673612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T14:35:04.201252Z digest=sha256:5f2a60ff8b20604c5559fc0c54fb7af58689a0240aed335b393dd0da120b045d

Observation 02524fc5-bb06-4565-9c41-dd5ae15b49d1 · inbound

D2 Actor Critic: Diffusion Actor Meets Distributional Critic cites this paper.

D2 Actor Critic: Diffusion Actor Meets Distributional Critic Temporal Difference Learning for Model Predictive Control

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:35:29.447709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T07:31:27.330951Z digest=sha256:24e641dc8e6d5097abe0211bcc3537298cd36aac59bdedd9e50470ce562822f4

Observation 1e85fd0c-eeb7-4b03-b1b0-084afa60a574 · inbound

Ctrl-World: A Controllable Generative World Model for Robot Manipulation cites this paper.

Ctrl-World: A Controllable Generative World Model for Robot Manipulation Temporal Difference Learning for Model Predictive Control

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-16T01:14:10.399192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T01:14:10.174044Z digest=sha256:ed75f0b003e6f75c0e4e6c0c58d7fdd1174a97182c94de996644155a0d9937b6

Observation e07f23f2-a065-4dc2-a677-93eb3386fc5f · inbound

Next-Latent Prediction Transformers Learn Compact World Models cites this paper.

Next-Latent Prediction Transformers Learn Compact World Models Temporal Difference Learning for Model Predictive Control

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-25T07:25:29.528570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T07:21:04.684347Z digest=sha256:20282ab3ae4732285004bcaee15126d3dcc7d264f7abcb283db153817d163604

Observation 5557c7da-06d9-48d5-97af-6e146eac1cb0 · inbound

Next-Latent Prediction Transformers Learn Compact World Models cites this paper.

Next-Latent Prediction Transformers Learn Compact World Models Temporal Difference Learning for Model Predictive Control

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T23:27:55.490000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:27:55.490000Z digest=sha256:3684bc34796d886c9bddff2bacfeeec52f3e68de94cdc9d0c7c556fa5f7416ee

Observation 5ebb0c64-7a51-49cc-8dd9-3cc23dc5e40b · inbound

CoRL-MPPI: Enhancing MPPI With Learnable Behaviours For Efficient And Provably-Safe Multi-Robot Collision Avoidance cites this paper.

CoRL-MPPI: Enhancing MPPI With Learnable Behaviours For Efficient And Provably-Safe Multi-Robot Collision Avoidance Temporal Difference Learning for Model Predictive Control

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T22:45:05.283619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:45:05.283619Z digest=sha256:28511b016bbbb20dfd11bb10fa29cf369728a720eac3bfcdbc914db8db2514e9

Observation bc98658b-637c-4452-9ece-d4ffce8d7d84 · inbound

Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning cites this paper.

Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning Temporal Difference Learning for Model Predictive Control

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T14:50:12.913336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T14:50:12.804707Z digest=sha256:672798a7a090ed8d02cc89f02bd42fbb8eb55b140a68bdabeaa2fe25733d40cf

Observation eafb0b50-f16f-42ca-8105-9781c824625a · inbound

Beyond Test-Time Memory: State-Space Optimal Control for LLM Reasoning cites this paper.

Beyond Test-Time Memory: State-Space Optimal Control for LLM Reasoning Temporal Difference Learning for Model Predictive Control

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-15T12:09:16.468055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:09:16.468055Z digest=sha256:2eb34df1a37a8cbb8cb6c95528c51ecb41f880406539fba0d9ce925b61c31a07

Observation e8b490cc-ed7f-414a-ba21-c9bfa54c004e · inbound

RAY-TOLD: Ray-Based Latent Dynamics for Dense Dynamic Obstacle Avoidance with TDMPC cites this paper.

RAY-TOLD: Ray-Based Latent Dynamics for Dense Dynamic Obstacle Avoidance with TDMPC Temporal Difference Learning for Model Predictive Control

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.799296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T08:43:35.563513Z digest=sha256:72545473db769e4369b9a9ed96c258f1697b6df433e3614efe59bc526b2111c3

Observation f352925b-cc14-4e46-81e1-416b17436e9b · inbound

TRAP: Tail-aware Ranking Attack for World-Model Planning cites this paper.

TRAP: Tail-aware Ranking Attack for World-Model Planning Temporal Difference Learning for Model Predictive Control

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:56:04.901730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:15:45.407680Z digest=sha256:43d4d40924e3894d4362aff0cbbfcbc42b38f22b8c9b33c75546f8a1d7d9f91d

Observation 4ff10b9a-d7d1-40fe-975d-2e3dd7cc63b3 · inbound

Learning Visual Feature-Based World Models via Residual Latent Action cites this paper.

Learning Visual Feature-Based World Models via Residual Latent Action Temporal Difference Learning for Model Predictive Control

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:40:51.947552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T01:37:27.658078Z digest=sha256:4acda2dcfafd6f406cc6fbfee1b6005e56b5a7667ee076403c8adc5fa70eaee7

Observation 0fb90d78-f436-4320-a716-5141f08363e2 · inbound

JEDI: Joint Embedding Diffusion World Model for Online Model-Based Reinforcement Learning cites this paper.

JEDI: Joint Embedding Diffusion World Model for Online Model-Based Reinforcement Learning Temporal Difference Learning for Model Predictive Control

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:37:51.941434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T19:37:19.404335Z digest=sha256:f7146843d66dd4e57bc0981d3b569871199ed4d436723a8f9bada3224dbc4d90

Observation 5deb802f-3537-44be-a08e-e9e8d8a2486d · inbound

Matrix-Space Reinforcement Learning for Reusing Local Transition Geometry cites this paper.

Matrix-Space Reinforcement Learning for Reusing Local Transition Geometry Temporal Difference Learning for Model Predictive Control

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:03:28.775037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T02:03:11.540545Z digest=sha256:4c339023211aebb69acf635c31798baa171b4f820184b492dbb6e9042a37127e

Observation 1aad3508-dc94-454c-9196-d633a5efbdb8 · inbound

Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making cites this paper.

Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making Temporal Difference Learning for Model Predictive Control

Reference 300

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:59:02.097841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T20:54:31.025488Z digest=sha256:505391570b04804ca4c1060e66d864eb7f2485d1eddaba694fe90ff0255dfa9f

Observation 5221a487-d671-40b3-8abe-9c1d1dc5c585 · inbound

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation cites this paper.

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation Temporal Difference Learning for Model Predictive Control

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:01:20.259352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T08:57:19.834179Z digest=sha256:18586af9800a34ae54a47d05c560db58323e5c0f6b9a62ada3599310bd7491a7

Observation 6467fad2-d804-4929-a4f4-601c7fa178d1 · inbound

Scaling World-Model Reinforcement Learning Through Diffusion Policy Optimization cites this paper.

Scaling World-Model Reinforcement Learning Through Diffusion Policy Optimization Temporal Difference Learning for Model Predictive Control

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:01.393138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T22:46:27.179341Z digest=sha256:3a9a3be0af4bde85a2b8dfd8b892558220791a99253ae43cf7491aa5654aa8fb

Observation 3d7662a0-b317-4104-9406-3d811556df83 · inbound

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient cites this paper.

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient Temporal Difference Learning for Model Predictive Control

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:03:48.589473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T17:34:41.053725Z digest=sha256:9622afca3f73049abc45fa35eeae2beddda355664f2a4a19f81002e356072bc0

Observation 9a77346a-24c1-43bc-920f-d581620dccf6 · inbound

MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models cites this paper.

MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models Temporal Difference Learning for Model Predictive Control

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.836554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T07:46:28.469913Z digest=sha256:fbee69b8020c83f0242520af67cf41dcde373edf0234a96857e60903131b7287

Observation d0ccad1b-1419-47ed-ab0a-ebab28666818 · inbound

Enhancing Human-Likeness in Reinforcement Learning Agents via Hierarchical Macro Action Quantization cites this paper.

Enhancing Human-Likeness in Reinforcement Learning Agents via Hierarchical Macro Action Quantization Temporal Difference Learning for Model Predictive Control

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:32:44.205977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T22:29:34.616531Z digest=sha256:43349dc68e623e6605aa8d1ed419ae61b5c184cf8aa1862634141835e1d71d27

Observation a9e4387f-ba8b-4e27-9052-eb84d5f43e5c · inbound

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning cites this paper.

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning Temporal Difference Learning for Model Predictive Control

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:03.124660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T09:48:27.652901Z digest=sha256:20af89ac8322b5819cb9c4a6735758431cbd0e29a3b035c4c895d934a8e00155

Observation 62e5c0bf-1b8e-44de-ad89-bc8f631f0feb · inbound

Solving Markov Decision Processes with Future Information via MPC cites this paper.

Solving Markov Decision Processes with Future Information via MPC Temporal Difference Learning for Model Predictive Control

Reference 84

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T19:00:06.257841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-25T22:00:02.577707Z digest=sha256:0e598f750b1d7931625d65ab74f4821cb1aac95e672cca83a38e1f1e52d5fe64

Observation 9dddc824-db97-47c5-8e68-30ebc3cdc0da · inbound

Valdi: Value Diffusion World Models cites this paper.

Valdi: Value Diffusion World Models Temporal Difference Learning for Model Predictive Control

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-02T15:47:05.609873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-02T15:46:50.507626Z digest=sha256:1ce3c598e1d0c427158c45d469912d249ca759de356a6087e24b9da13dd0b328

Observation debe30af-548b-4ee4-8683-a8205d23c246 · inbound

DriftWorld: Fast World Modeling through Drifting cites this paper.

DriftWorld: Fast World Modeling through Drifting Temporal Difference Learning for Model Predictive Control

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T00:24:04.583755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:24:04.583755Z digest=sha256:f9433cd56d682d8dd20288a231d379d808efc9f663af3025ca5d017a33171645

Observation 4c724199-6b02-4ee6-a461-50feb15d93dc · inbound

Reinforcement Learning: From Algorithms To Foundation Models cites this paper.

Reinforcement Learning: From Algorithms To Foundation Models Temporal Difference Learning for Model Predictive Control

Reference 254

Resolution
unresolved
no resolver link, observed 2026-08-01T17:45:21.168260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:45:21.168260Z digest=sha256:03faea034f85b9b90a122e243dcb1a4524839881141a64def329a2215566955c

Observation 42a18446-7aa8-464e-8d94-6ef3f37aafbe · inbound

Toward Goal-Agnostic Joint-Embedding Predictive Control of Partial Differential Equations cites this paper.

Toward Goal-Agnostic Joint-Embedding Predictive Control of Partial Differential Equations Temporal Difference Learning for Model Predictive Control

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T12:24:20.765658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:24:20.765658Z digest=sha256:32c73574ac0ff8c5a1154e69447ce9ae912fd561b6062641504413b1278e0211

Observation 7de8c998-94cb-4d96-98a8-5f39e1582169 · inbound

When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual Auditing of Frozen Locomotion Policies cites this paper.

When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual Auditing of Frozen Locomotion Policies Temporal Difference Learning for Model Predictive Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T06:32:51.940717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:32:51.940717Z digest=sha256:0eaecd611ededc2f00fee616f542bf45064964841cee51d9f3b1f27397be4599

Observation 8fd80f4c-4625-4ccb-bb8f-d3e3a4e7c2e8 · inbound

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models cites this paper.

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models Temporal Difference Learning for Model Predictive Control

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T04:59:31.588147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:59:31.588147Z digest=sha256:9bc3d538915efe991e7ec5b28715599b55ec337c8d607a1cb1bd851705db424a

Observation 3e8be49d-0c0e-4320-9bcf-87242aebceb4 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Temporal Difference Learning for Model Predictive Control

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:32.767376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:32.767376Z digest=sha256:8b6934c7fea04c2724c3ab8bf6b7cda20ecd38575cbbe0eb6c8cedabdfc3caa4