Pith. sign in

Paper Citation Record · LEDGER

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models

As of 12 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2608.06799.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06799 v1

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:41:43.912335Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

29 of 29 outbound references displayed

  • verified exact2
  • verified fuzzy2
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c0ab036a-f408-45a3-84df-45276b015293 · outbound

This paper cites Ctrl-World: A Controllable Generative World Model for Robot Manipulation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Ctrl-World: A Controllable Generative World Model for Robot Manipulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.805937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.805937Z digest=sha256:255a912589416eea545b9065f29e2aad26ea0cb57ece90284b0a7540ca45f872

Observation ff021cfc-2268-44d6-ab40-4170ef71fb38 · outbound

This paper cites World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models World Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.810581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.810581Z digest=sha256:1291297d2a2f90403d6b860060b50b7f3ccbc11e9f14f29041d347ff48c882c8

Observation 3b91ce94-53c9-4e2e-976c-7b955099b888 · outbound

This paper cites World Model for Robot Learning: A Comprehensive Survey.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models World Model for Robot Learning: A Comprehensive Survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.819932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.819932Z digest=sha256:a784a667348022b7a44789aa4ef45e136408e9d019e4145f7133d685aa26b2f2

Observation dcd1afdb-0f74-4bc3-9709-a984eff04d8b · outbound

This paper cites GAIA-1: A Generative World Model for Autonomous Driving.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models GAIA-1: A Generative World Model for Autonomous Driving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.824533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.824533Z digest=sha256:b4c7eeaedf9f5064a85ecc12134175f9913210dbd147d61cb793c5744b77463a

Observation 5f92cb20-abac-4f4a-8b4b-aac060782734 · outbound

This paper cites PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.829126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.829126Z digest=sha256:09fb9884fc8770ff4217007dbb2abe9fa8c016df6b39bcf6e88b93ccb5c09703

Observation 1fc97768-26d9-43a6-be44-ef0a5a4f0c8e · outbound

This paper cites Robots pre-train robots: Manipulation-centric robotic representation from large-scale robot datasets.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Robots pre-train robots: Manipulation-centric robotic representation from large-scale robot datasets

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:41:44.680657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:41:43.833447Z digest=sha256:c61d583201e5d15bdc64cf683b448539883201a126e3f3ddd4dc8f5b19675a39

Observation 7ea2f6d8-a26b-4621-8ca7-0e96105cb8e0 · outbound

This paper cites Contrastive Representation Regularization for Vision-Language-Action Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Contrastive Representation Regularization for Vision-Language-Action Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.838078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.838078Z digest=sha256:c58bfe38e04a8d8f9c62bb3fee44c66b155cb564188fd1a76a5057a0571a1356

Observation af3f1281-6c88-4867-b91b-fddb33a74a6e · outbound

This paper cites Predictive but Not Plannable: RC-aux for Latent World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Predictive but Not Plannable: RC-aux for Latent World Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.842384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.842384Z digest=sha256:a8e41bbc25b5661fe5d405384a9b61a3c67615e0669303bc77c1860d1a7406b3

Observation 2bb77745-38f2-4b85-b1d7-189a61b52d8a · outbound

This paper cites Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.846637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.846637Z digest=sha256:7ccb0ef7e280d6e9ed4b8b9e4fa4f0c1813c5e1557a8043e412abdb5352659e2

Observation 51beead7-cf0f-4c5a-bc12-9285f117fa83 · outbound

This paper cites LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.851098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.851098Z digest=sha256:d70922c1cabc93ce2177c28e5a8997a1f52ae64ebf041b34d1272118c22ce8ee

Observation aedfe381-ac0e-4ae4-9857-f3162503142e · outbound

This paper cites LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.855678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.855678Z digest=sha256:cc564a4053d8ca439b93ee05ce57a8e8b68f6d2b80919c1b04c59e2b318bf1c1

Observation 643e5636-e55b-4e10-bad6-760151b930b4 · outbound

This paper cites V-JEPA 2.1: Unlocking Dense Features in Video Self-Supervised Learning.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models V-JEPA 2.1: Unlocking Dense Features in Video Self-Supervised Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.859391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.859391Z digest=sha256:939887b3379aceb35cf1f0d244d08de18099f31143d61a3918973a7b32159c53

Observation 67a071b4-b42c-4f8a-adcc-b42a81e085d4 · outbound

This paper cites Latent Geometry Beyond Search: Amortizing Planning in World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Latent Geometry Beyond Search: Amortizing Planning in World Models

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:41:44.410479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:41:43.863650Z digest=sha256:ebd7ebac15224fdd12bf13d3dcce08fec4794acf3727c994a3fb35074cbed102

Observation 10945e6d-0db1-408e-b76b-d182625c2ba1 · outbound

This paper cites LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.867336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.867336Z digest=sha256:cdf4394e75f9b8ec0214bb5f579639d06b31f929253f342a9da259b7a8b97eeb

Observation ec375ccf-128a-4bf5-a12b-5cc8b3c77bcf · outbound

This paper cites Reconstruction or Semantics? What Makes a Latent Space Useful for Robotic World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Reconstruction or Semantics? What Makes a Latent Space Useful for Robotic World Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.871583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.871583Z digest=sha256:dc8753a01fdf06b1bc24b49371e862ed5ef1dd6c30765d7993ed0641aff3efa0

Observation 2efa1da1-fa8e-4df0-a553-911bd1a2e13f · outbound

This paper cites OGBench: Benchmarking offline goal-conditioned RL.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models OGBench: Benchmarking offline goal-conditioned RL

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:41:44.668012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:41:43.876052Z digest=sha256:803d6eaa0c38e7b76f3e6ad3dff1a0505f9d03a71fab766c56c6ed0b806984ed

Observation 8c808325-6709-4310-8b71-c8b8d37aec99 · outbound

This paper cites an unresolved cited work.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:41:44.655923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:41:43.880350Z digest=sha256:39df8c1bd506d107882180145aefbbc0ec60e62147e28c741348328cf34ab79a

Observation 5ab9f07f-3493-4110-82c5-0dec7157c14d · outbound

This paper cites GigaWorld-0: World models as data engine to empower embodied AI.arXiv preprint arXiv:2511.19861,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models GigaWorld-0: World models as data engine to empower embodied AI.arXiv preprint arXiv:2511.19861,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.884472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.884472Z digest=sha256:d62200bdf9ed41eb0d84e8c590633e7829ce4eee2e23b4c86e6824e3fca1a48f

Observation fa3a742b-7e23-4cc6-bc98-485f524eae8d · outbound

This paper cites GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:41:44.293651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:41:43.888041Z digest=sha256:dad281579c8380f924360bccbb41edfd3ea9f9c5c7dbe470bea4cf287663e56d

Observation 100d3085-3228-4b68-934e-1b92b13bcb18 · outbound

This paper cites Beyond language modeling: An exploration of multimodal pretraining.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Beyond language modeling: An exploration of multimodal pretraining

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.891873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.891873Z digest=sha256:149ce70eb00091d76bfb10f3b93826dc17399b57a4a61df82e7b2e5749867514

Observation b2b4d677-4f69-4894-89b2-530598783472 · outbound

This paper cites Open-world hand-object interaction video generation based on structure and contact-aware representation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Open-world hand-object interaction video generation based on structure and contact-aware representation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.895990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.895990Z digest=sha256:080de94fea65b94a9c2653fc4992109ef953b45694924b81ac347ff5099dd378

Observation fd1f48ef-37a7-4918-ab92-6647f9a4f302 · outbound

This paper cites UniDriveDreamer: A single-stage multimodal world model for autonomous driving.arXiv preprint arXiv:2602.02002,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models UniDriveDreamer: A single-stage multimodal world model for autonomous driving.arXiv preprint arXiv:2602.02002,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.899958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.899958Z digest=sha256:4ea6803defe69f31fe39eb59ed79caa52fa61305ac0e38bed63b558e1c74cba0

Observation a6211a8a-2fbd-4454-aef9-41f3e0c4ef1e · outbound

This paper cites FlowVLA: Visual chain of thought-based motion reasoning for vision-language-action models.arXiv preprint arXiv:2508.18269,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models FlowVLA: Visual chain of thought-based motion reasoning for vision-language-action models.arXiv preprint arXiv:2508.18269,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.904159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.904159Z digest=sha256:1b79047797eb429da7f89bc4155a7be80e1ed217b7f7d80c21ccd67086d83603

Observation 4de75839-f51f-4d07-8fc8-2261a689b7ea · outbound

This paper cites DualCoT-VLA: Visual-linguistic chain of thought via parallel reasoning for vision-language-action models.arXiv preprint arXiv:2603.22280,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models DualCoT-VLA: Visual-linguistic chain of thought via parallel reasoning for vision-language-action models.arXiv preprint arXiv:2603.22280,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.908418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.908418Z digest=sha256:daecf4dc67bbaf58350f8d3e18188098f8501db21a257bdd2dc6fe6f729a717c

Observation 45e2f42a-1cc4-4f11-9eae-04f23fe607f5 · outbound

This paper cites DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.912335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.912335Z digest=sha256:cc5146608166736edc01f6702ffe8f800560298b6775b526d34fc0a322479497

Observation b57a9edb-0419-45a3-ae8c-8411cb48279c · outbound

This paper cites Mastering Diverse Domains through World Models.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Mastering Diverse Domains through World Models

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.815090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.815090Z digest=sha256:a1112e28c9e260241d36f483d424a9b2e1ca5b69f8837c9eef590d10d3c31576

Observation cec43895-c873-4d31-8f07-e2c3850ce7af · outbound

This paper cites V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.793178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.793178Z digest=sha256:819faa1899570dc218dd5e616054cc856000864f1c02a7b926d186d701f2807c

Observation bfdf302f-39df-4e2f-a3ec-b7a8a32bc3f2 · outbound

This paper cites Why ai systems don’t learn and what to do about it: Lessons on autonomous learning from cognitive science.arXiv preprint arXiv:2603.15381,.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Why ai systems don’t learn and what to do about it: Lessons on autonomous learning from cognitive science.arXiv preprint arXiv:2603.15381,

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.797534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.797534Z digest=sha256:ad5f67ca386b8ae41c3a51f80ae055f3069427c008b272738538c0003711a8c1

Observation faa62ea0-a425-480d-a657-9e928ba511f5 · outbound

This paper cites Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation.

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-10T20:41:43.801307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:41:43.801307Z digest=sha256:7066a9814496de5102b1e756f249c2c77a345dd0a54d88ee0cff53750f4c7d03

Pith citing papers

No inbound Pith citation observations are available.