Pith. sign in

Paper Citation Record · LEDGER

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint

As of 10 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 3 inbound Pith citation observations for arXiv:2502.03550.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.03550 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:38:30.745532Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:27:30.170831Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T09:56:27.801679Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f56e1c6f-7fa0-45ef-98e2-c52492be8718 · outbound

This paper cites Model-Based Offline Planning.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Model-Based Offline Planning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.124877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.124877Z digest=sha256:e67cac6533832c2247ea8308e6d8292dc964e140084308405e46deb95e11bd3f

Observation c2bbd487-d4d9-40bd-8fdc-7979a37f7ef1 · outbound

This paper cites Bertsekas.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Bertsekas

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.996494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.192382Z digest=sha256:aed5375c818872f57846e6b274589598ea4132539db13e2ca0ff29d95482807c

Observation 519b3369-66e9-4e7d-bf7e-dffe0a053f9c · outbound

This paper cites Bhardwaj, A.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Bhardwaj, A

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.980648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.197110Z digest=sha256:803d86118cfead0392a83231d9bf9d4bb7671f6628296785170a5297de2a955f

Observation 08590fe3-6236-479f-a724-0f7540a02bd6 · outbound

This paper cites an unresolved cited work.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:38:31.964782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.202166Z digest=sha256:ad61490b3a1717a8983a0922de1f5bd8a14c33d4e18e7bff553d0259569fe166

Observation cabca893-2be3-4a28-9cfc-4f077aa2c774 · outbound

This paper cites an unresolved cited work.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.207033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.207033Z digest=sha256:f9318c58336b01dcbf9e342e21c634216f2548de721d68d2aa9423a07a73c892

Observation 5a428069-57d7-4dc9-b4a3-e2abcd1a6b02 · outbound

This paper cites Fujimoto and S.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Fujimoto and S

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.919520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.211877Z digest=sha256:67f6dc2c7354e0d913b0d76476c0f11b2a909f0a7a4b23a4cef5f2cba6df6edf

Observation 7442f3a2-d4a2-4579-9e17-3af9a8ab19eb · outbound

This paper cites Fujimoto, H.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Fujimoto, H

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.829642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.218712Z digest=sha256:86224fec1f710dd42fd8e2c0f014ffa72b8ac789e6fab7da4480f077c0835929

Observation 8feb9364-cea6-4e24-bf8e-c535ff65f48e · outbound

This paper cites Fujimoto, D.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Fujimoto, D

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.747103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.223435Z digest=sha256:6b9deb4288094b7e237744821d59e743231c5ca454d044e0f0519a3c4923f58c

Observation 8b590690-0025-41e2-9905-07b1a5960926 · outbound

This paper cites Extreme Q-Learning: MaxEnt RL without Entropy.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Extreme Q-Learning: MaxEnt RL without Entropy

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.228255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.228255Z digest=sha256:eb29282dc002e20c1c5e41cce4b278ebb28eee7530bdd520eb342dcc65d2a464

Observation e5977cab-fc46-48e4-b042-c3811fc215c3 · outbound

This paper cites Haarnoja, A.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Haarnoja, A

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.675738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.232877Z digest=sha256:1ab32dcdb2805a983073be0990368eb7d4c683e9e28408c9113cada454e3472d

Observation 70b6a54d-f06b-48c5-ad77-27857d432087 · outbound

This paper cites Hafner, T.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Hafner, T

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.237424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.237424Z digest=sha256:2e51202688cfa216c93eb61c9cc4d94397e4e25d8a7e2d2c026caa5f355074ae

Observation a660f189-2834-4629-9f96-021408a568d7 · outbound

This paper cites Mastering Atari with Discrete World Models.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Mastering Atari with Discrete World Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.242136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.242136Z digest=sha256:5165345140fffbeded57fd733aacf691d8e94e23eee21c268c4cfbf18adedb4c

Observation 401fa314-6b9f-47b7-a504-5b3bfd1767f3 · outbound

This paper cites Mastering Diverse Domains through World Models.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Mastering Diverse Domains through World Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.257078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.257078Z digest=sha256:a2659f7b2f96c6d98de0b0250e53900ddd42c1294d90be23a3d7ae4ad61b693d

Observation f4b7eb98-7dfd-4f39-9e84-d717639c59f9 · outbound

This paper cites TD-MPC2: Scalable, Robust World Models for Continuous Control.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint TD-MPC2: Scalable, Robust World Models for Continuous Control

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.266687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.266687Z digest=sha256:86c00250180d5fee60678b26fd3c71695ac0c1be345231337204881f59e08a72

Observation 4249080f-d4b3-4d86-a741-110d2eac392b · outbound

This paper cites Temporal Difference Learning for Model Predictive Control.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Temporal Difference Learning for Model Predictive Control

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.271605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.271605Z digest=sha256:f44e0eca04da8c52c1e813b529f8682647a5512e2da2c8ef48e1a11c998d1216

Observation eeaeff53-9369-48a2-8fdb-d34285396c27 · outbound

This paper cites IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.276596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.276596Z digest=sha256:c6a5fe40642bbdbe669707c387abc8ba695e82cfe334760980caaf7bb9ec2fab

Observation ae0332d3-7a15-42c5-a263-eb211c34bf9c · outbound

This paper cites Janner, J.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Janner, J

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.642661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.281062Z digest=sha256:669781b3c87b8e22fdf49ab6e8e0d2d019cfcc14901c8737204b082a2d4c883b

Observation 6141572d-b3b7-4136-bcbb-60c913ca303b · outbound

This paper cites Kabzan, L.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Kabzan, L

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.626186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.285679Z digest=sha256:d1b212ce88082a1cc5b62c8aa67489d26ec64872b4535b265de7cf03e4972d7b

Observation 42cae131-7b0f-4643-bdf9-76a84831b285 · outbound

This paper cites Katsigiannis and N.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Katsigiannis and N

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.609507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.289823Z digest=sha256:235db4f04b93cf71d73213e78ca7502cd831228c13aa49f740f3471fa8e7b080

Observation 9c18f673-7e58-471c-aa6e-cb5990781dc0 · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Offline Reinforcement Learning with Implicit Q-Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.294338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.294338Z digest=sha256:d7260661256c38d04e63716279cf9ed30798269e450c39970e565a0e3fb0324d

Observation 7a85f7ab-7872-4eb4-9af3-31ab4ef8bcb1 · outbound

This paper cites Kumar, J.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Kumar, J

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.594121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.298917Z digest=sha256:2868bbef546123f4b89edbaa5aea3366fda70b63e6004c426744b5028e0c91f8

Observation 83656bc0-ea89-49bf-9cc9-47ae1768b611 · outbound

This paper cites Kumar, A.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Kumar, A

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.579052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.303175Z digest=sha256:f75635854794bd477d302bda602507482a2600780d8dc3306f3f1b3fa1c93af8

Observation 2e7546d1-e634-4b71-b34c-7210e38c3061 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.339026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.339026Z digest=sha256:55163aad084945d42a08ea956b6ad05328b3a39856ae097b17bab12ade6436ca

Observation 029326ba-1ce2-4930-958f-0a7ecdd9392c · outbound

This paper cites an unresolved cited work.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.371033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.371033Z digest=sha256:f04bc78059b505f7e6b64ad6e93996102f9517fe1093932de317a572850c233b

Observation e0318fbd-c343-4565-ab6c-4468e4f59fd5 · outbound

This paper cites Littman and A.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Littman and A

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.563079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.410665Z digest=sha256:6a6e698aeb54317bba323cabc001dc04582f9222f87cb8724a6bc5a75548e128

Observation 7348b843-5152-447b-b0ab-b73dd3d03181 · outbound

This paper cites Plan Online, Learn Offline: Efficient Learning and Exploration via Model-Based Control.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Plan Online, Learn Offline: Efficient Learning and Exploration via Model-Based Control

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.455729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.455729Z digest=sha256:9dcf652d2a68c190d61a06fa5516fad3231064e0a1a4fb703583bcdd970626a1

Observation 1e4ef373-e834-469e-920c-4bac2418e548 · outbound

This paper cites an unresolved cited work.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:38:31.546528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.487926Z digest=sha256:72f9dd8445a0d4ae3dcc98bb3f1af68f839b8092a6471569d9d1df2b54ab8d89

Observation 5f696695-2beb-46e4-b953-857243b74732 · outbound

This paper cites an unresolved cited work.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:38:31.530900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.524226Z digest=sha256:c3563d1a47a3d4b4141f94dafd183f48b9d457531f1b88e4cf6e7ebe44f0f6d8

Observation f4b860dc-19a2-4a85-a1c0-4331ef5d6a08 · outbound

This paper cites AWAC: Accelerating Online Reinforcement Learning with Offline Datasets.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint AWAC: Accelerating Online Reinforcement Learning with Offline Datasets

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.572672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.572672Z digest=sha256:dc7a81b091bafc7fa4efcfc6c65d9986f6f859f9b38a7f54345506ceec41bf15

Observation 79e80661-4606-431d-bf29-13365e22ffd8 · outbound

This paper cites Nakamoto, S.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Nakamoto, S

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.515069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.609097Z digest=sha256:69e2c1146fd0c45d53552d9496769fbd3c6f0e3756babf93042af5a6a952d5e9

Observation d7bf83f9-7ea1-4056-8267-c873f50e8fe9 · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.689021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.689021Z digest=sha256:eb2afa7ef5129452e969897c856eaca49f4f43d2093d48d24a32ad7af653dd20

Observation 630c9d3d-4412-45d3-b40b-f0db30e070d4 · outbound

This paper cites Trust Region Policy Optimization.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Trust Region Policy Optimization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.694082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.694082Z digest=sha256:f7262c1b3d1a86c478e1c9f4bc5ec67d3be874599ccb5f766dfdd4de348080bf

Observation 1294ccb0-cd74-46f5-b3e4-88d81c668ee6 · outbound

This paper cites HumanoidBench: Simulated Humanoid Benchmark for Whole-Body Locomotion and Manipulation.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint HumanoidBench: Simulated Humanoid Benchmark for Whole-Body Locomotion and Manipulation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.698954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.698954Z digest=sha256:9e082d296cf40d7c9df2f35f9c59a5c2cafa88309c4c14364592bc205b7fd4ab

Observation cce74b20-cde2-4ce9-946f-f98bd742c5a1 · outbound

This paper cites Sikchi, W.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Sikchi, W

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.472133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.703425Z digest=sha256:faef75b305f44883a18ef8b798f49726faa0c8508684644d2c628824aa82a3e1

Observation 95bbe34e-3893-4893-a074-af28b11983b2 · outbound

This paper cites an unresolved cited work.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:38:31.321794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.707899Z digest=sha256:ca3d846660fd56a6ccec8976c6747ae830b0abf6781858927700608b5cfc0be4

Observation e714c118-1a4f-4184-81fb-10adb79753f1 · outbound

This paper cites an unresolved cited work.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.711895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.711895Z digest=sha256:a9ef8f34b41cbc12ea688868208e5fd8582c6d05887ed653c83e363a4deb8dd7

Observation 524f49e3-1015-4ff9-8ba5-bf0351f5b9fb · outbound

This paper cites an unresolved cited work.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.716371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.716371Z digest=sha256:1be57406e6429da168dc7f99f9826209540bea3417cb0adc6f6fca69a3942728

Observation 03b7e3bf-b9ad-4a06-9d3d-d41cd7ab7fc9 · outbound

This paper cites DeepMind Control Suite.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint DeepMind Control Suite

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.724743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.724743Z digest=sha256:785fa5c9acc21275455b94e98711e2d395db08f6808d3fff3b9f3c8406eceb67

Observation 8823a377-3131-4459-bc00-2a4113ebc097 · outbound

This paper cites Thrun and A.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Thrun and A

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.265681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.730792Z digest=sha256:e721dd3b38d257d25a1f994464919ed0ca0c5480cb0e5443624f4ee2f538daa8

Observation aff67faf-75cd-4427-8aa2-a7851e2150eb · outbound

This paper cites Williams, N.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Williams, N

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:38:31.250709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:38:30.734739Z digest=sha256:a8bf6ca5ed1b858459f805f1e0f69b2ec92592d51a406a4f42fce32bdca34cab

Observation 5e5b0fcc-f58c-4872-b15f-75026891178c · outbound

This paper cites Diffusion Model Predictive Control.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Diffusion Model Predictive Control

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.745532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.745532Z digest=sha256:d5eb5b4a480877b0dea2c807b37fd41bfab8013e14dedb50f9cf627181081c6b

Pith citing papers

Observation a772cadb-ca4b-44c8-844c-52e155ff9de9 · inbound

DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion cites this paper.

DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:27:30.170831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:27:30.170831Z digest=sha256:bab49e31e56fa829c9bde1ec9997b791683d86043e540fa9ce061eacf437dbfb

Observation 75bf9623-f153-4404-a424-e19b8526d9cf · inbound

RAY-TOLD: Ray-Based Latent Dynamics for Dense Dynamic Obstacle Avoidance with TDMPC cites this paper.

RAY-TOLD: Ray-Based Latent Dynamics for Dense Dynamic Obstacle Avoidance with TDMPC TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.805022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T08:43:35.563513Z digest=sha256:af701d7691b32af21f99253d8e81683e3ea09d673a1130165e78f933d446fa0c

Observation 7c62a9e8-584d-419c-b612-6138dcb44332 · inbound

Dream-MPC: Gradient-Based Model Predictive Control with Latent Imagination cites this paper.

Dream-MPC: Gradient-Based Model Predictive Control with Latent Imagination TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T18:11:05.433394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T16:31:35.396809Z digest=sha256:b2632099f9e3e12cf12200f3968589bc7fd7794f3bde7ab48d214d804f8436c5