Pith. sign in

Paper Citation Record · LEDGER

Learning and Leveraging World Models in Visual Representation Learning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2403.00504.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.00504 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T05:06:02.521473Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T22:54:01.434942Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ab95dbd2-51b2-4d76-b5d3-f2ec68c15207 · inbound

HEP-JEPA: A foundation model for collider physics using joint embedding predictive architecture cites this paper.

HEP-JEPA: A foundation model for collider physics using joint embedding predictive architecture Learning and Leveraging World Models in Visual Representation Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T00:14:59.569131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:14:59.569131Z digest=sha256:2a90a60383f18f9a71ada2b2fec2d2c72a89ae313f610e83b15ab43146912c87

Observation c9d461e5-a4b4-466e-8fa4-a76e70df02ef · inbound

Flopping for FLOPs: Leveraging equivariance for computational efficiency cites this paper.

Flopping for FLOPs: Leveraging equivariance for computational efficiency Learning and Leveraging World Models in Visual Representation Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T20:07:43.743498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:07:43.743498Z digest=sha256:2c686fb6f8e1393ff63b6c8dab7da64bf9d3e58d4315c44cb3ac2ba4bde68a4a

Observation 666131fe-3f33-46f3-bc47-c1fc4814381d · inbound

EgoAgent: A Joint Predictive Agent Model in Egocentric Worlds cites this paper.

EgoAgent: A Joint Predictive Agent Model in Egocentric Worlds Learning and Leveraging World Models in Visual Representation Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T17:46:20.572357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:46:20.572357Z digest=sha256:4ad2f267ff6473df2eed7de2e30dacca8151da006e535faebf51dd744286ea69

Observation 50b6e713-4e3c-4b7d-976e-3493833bb0b2 · inbound

Galileo: Learning Global & Local Features of Many Remote Sensing Modalities cites this paper.

Galileo: Learning Global & Local Features of Many Remote Sensing Modalities Learning and Leveraging World Models in Visual Representation Learning

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T21:51:15.876843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:51:15.876843Z digest=sha256:35b618e912a1229f651af196220fba7bad5d48775151b4692315231e4368b944

Observation f79ec40c-044a-4dbb-bf12-8ec53645cdff · inbound

Toward Embodied AGI: A Review of Embodied AI and the Road Ahead cites this paper.

Toward Embodied AGI: A Review of Embodied AI and the Road Ahead Learning and Leveraging World Models in Visual Representation Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:08.579604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:08.579604Z digest=sha256:6eb447b3a642525cc8dcb4443aec0fa7d3bdb95fa3eb13a42082a1b486e5b38c

Observation 3bb57699-b973-40f7-90e2-5082aa8596bf · inbound

WorldPrediction: A Benchmark for High-level World Modeling and Long-horizon Procedural Planning cites this paper.

WorldPrediction: A Benchmark for High-level World Modeling and Long-horizon Procedural Planning Learning and Leveraging World Models in Visual Representation Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T10:50:50.711002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:50:50.711002Z digest=sha256:74a9de15881ea82a73dae1e38a82e950a2fb5db3442dc14e7f1f724ead250695

Observation 55e27c9b-9eed-403d-9abe-847c6b28278f · inbound

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning cites this paper.

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning Learning and Leveraging World Models in Visual Representation Learning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:17:14.031185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T09:15:45.511104Z digest=sha256:dd7268470506f45adc9e7f1fda2726a0c5b8993a705d54fe1fac7cce15b4d6f7

Observation fe9b8ed9-e92e-44c5-8d57-8c22feb906e1 · inbound

A Survey of State Representation Learning for Deep Reinforcement Learning cites this paper.

A Survey of State Representation Learning for Deep Reinforcement Learning Learning and Leveraging World Models in Visual Representation Learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:35.739408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:34:35.739408Z digest=sha256:7369024bcdab5bf6a2104aef08a4e1f792ac460cdf9b7993b13144909586c9f3

Observation 2142393f-cdf3-4047-b9e9-62bba6d2dbd4 · inbound

Adaptive Bayesian Single-Shot Quantum Sensing cites this paper.

Adaptive Bayesian Single-Shot Quantum Sensing Learning and Leveraging World Models in Visual Representation Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T15:15:59.550313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:15:59.550313Z digest=sha256:d8f8529b90516eedadb584c5dd7d3f4db34060c625b37452aee6816402940152

Observation 8c4b7613-9c3b-405e-9ea1-38a9af354aa6 · inbound

Edge General Intelligence Through World Models and Agentic AI: Fundamentals, Solutions, and Challenges cites this paper.

Edge General Intelligence Through World Models and Agentic AI: Fundamentals, Solutions, and Challenges Learning and Leveraging World Models in Visual Representation Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T21:02:04.066353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:02:04.066353Z digest=sha256:07507d8bebe3268002aef7f3a0c9cc77746809fca08fd782a922fefafbb437db

Observation d514b024-71e7-4bcf-ae08-9e502997257c · inbound

GeoWorld: Geometric World Models cites this paper.

GeoWorld: Geometric World Models Learning and Leveraging World Models in Visual Representation Learning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:40:03.305541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T11:39:15.308355Z digest=sha256:dddc869e506b218fcf2c42e1296647aa34fde95194dad634f807df17ddfe91d8

Observation af72e11f-7161-4f5c-84ef-5a4d88bde945 · inbound

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction cites this paper.

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction Learning and Leveraging World Models in Visual Representation Learning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:50:05.116144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T14:46:18.645009Z digest=sha256:81a2399a43669d34cc1720b6527d1746d03b047e4804334576f20501a7211bf2

Observation ea28cad9-d805-4107-88ef-2ed58f723149 · inbound

The Lov\'{a}sz Local Lemma: Foundations and Applications cites this paper.

The Lov\'{a}sz Local Lemma: Foundations and Applications Learning and Leveraging World Models in Visual Representation Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-15T13:23:57.354325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:23:57.354325Z digest=sha256:3a3faa0177c24af662c7e02a74e21ae0ab9b557d0b7fdd900fc988a4ce132c61

Observation 77608f92-3d42-4d3f-84fc-aca1fb80b64e · inbound

World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry cites this paper.

World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry Learning and Leveraging World Models in Visual Representation Learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-13T14:05:26.303000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T14:05:26.303000Z digest=sha256:4df04ce7bb7184bf963e4e8fb533531fec78dddbad9b029a667deac8ffd6648d

Observation bfe5b756-c5e9-489b-b8cb-06e6f31b28ed · inbound

Toward Consistent World Models with Multi-Token Prediction and Latent Semantic Enhancement cites this paper.

Toward Consistent World Models with Multi-Token Prediction and Latent Semantic Enhancement Learning and Leveraging World Models in Visual Representation Learning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:55:49.618449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:28:33.350491Z digest=sha256:b07166d33897aab96d708f213618b9954b65847ab94c99e32992a08e160ee441

Observation 2e316451-1292-403a-9c68-42a607d3890f · inbound

Text-Conditional JEPA for Learning Semantically Rich Visual Representations cites this paper.

Text-Conditional JEPA for Learning Semantically Rich Visual Representations Learning and Leveraging World Models in Visual Representation Learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:51:30.669233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T17:58:43.650308Z digest=sha256:8a5c56289a5f70c41ddc1f513ade96e972637b6fc1baff42893695e12816ab4a

Observation 97cf2300-0297-4178-aee3-a1a1c256227f · inbound

AeroJEPA: Learning Semantic Latent Representations for Scalable 3D Aerodynamic Field Modeling cites this paper.

AeroJEPA: Learning Semantic Latent Representations for Scalable 3D Aerodynamic Field Modeling Learning and Leveraging World Models in Visual Representation Learning

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:36:09.268171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T15:53:39.074115Z digest=sha256:ce964d4444dc19160e119d8ee7ee964b0f4143343ac628192a16ce768e968fd7

Observation 0ac542cf-51b9-4b4e-ab99-c002c070c6be · inbound

Three-in-One World Model: Energy-Based Consistency, Prediction, and Counterfactual Inference for Marketing Intervention cites this paper.

Three-in-One World Model: Energy-Based Consistency, Prediction, and Counterfactual Inference for Marketing Intervention Learning and Leveraging World Models in Visual Representation Learning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T02:45:56.722986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T02:45:08.315949Z digest=sha256:bd7efc4cd74d871b293fe6f99d666ffa5a7f33b6089d194ebff5dd8f294b1dd3

Observation e15ea76d-f0a4-4cbf-891e-a37ec1d21076 · inbound

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation cites this paper.

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation Learning and Leveraging World Models in Visual Representation Learning

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:01:20.317075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T08:57:19.834179Z digest=sha256:157b8ca5757a0eb2f7217900436e7bee82a7b48ef3056e3fa7ab17912b3e8bc7

Observation 06811cd5-d8df-4689-8b43-e50da9c52169 · inbound

Scaling World-Model Reinforcement Learning Through Diffusion Policy Optimization cites this paper.

Scaling World-Model Reinforcement Learning Through Diffusion Policy Optimization Learning and Leveraging World Models in Visual Representation Learning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:01.436526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T22:46:27.179341Z digest=sha256:e7cdd73f72d1c9dd3de798891950c94438451fbe5b0d7ce2caa52e45dae0b3c1

Observation b0a6b739-b7b2-4d13-b0b2-535669071ff7 · inbound

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient cites this paper.

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient Learning and Leveraging World Models in Visual Representation Learning

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:03:48.594049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T17:34:41.053725Z digest=sha256:b4dff5749d78338c52edae2465609eb6171e6555c1cfff6a6f192a24022eae42

Observation 2574f0cd-bd8d-4b30-aaf4-6750e568f37e · inbound

Separating Representation from Reconstruction Enables Scalable Text Encoders cites this paper.

Separating Representation from Reconstruction Enables Scalable Text Encoders Learning and Leveraging World Models in Visual Representation Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T22:20:52.422988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T22:20:52.422988Z digest=sha256:0cc52201d58ba74930e438f381001dd795b260c0f2c97c70124a7fadaf9420f2

Observation ec1dd9e5-264d-4417-b06f-3f8a9553fca9 · inbound

Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation cites this paper.

Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation Learning and Leveraging World Models in Visual Representation Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T18:42:36.975241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:42:36.975241Z digest=sha256:c372b85788b17e5bb50f89af34ebd8be8b4d6790c54aab7b3bde6c3b8662a3a4

Observation d1c13e67-c5d0-4a44-868a-e50868c7a153 · inbound

SR-JEPA: Learning Predictive Latent State in 3D Scenes cites this paper.

SR-JEPA: Learning Predictive Latent State in 3D Scenes Learning and Leveraging World Models in Visual Representation Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T23:47:03.495059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:47:03.495059Z digest=sha256:74ad84de83591e73a7c191b31916bbf5df0f156aa16a89f906afea5e62517a33

Observation 700efa35-9b9f-459d-b5a9-e33df4471b86 · inbound

Support Operation Factorization: Compositional Readout of Frozen Vision Encoders under Controlled Interventions cites this paper.

Support Operation Factorization: Compositional Readout of Frozen Vision Encoders under Controlled Interventions Learning and Leveraging World Models in Visual Representation Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:26:05.346601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:26:05.346601Z digest=sha256:a5594663bb331bd374127837d594ffce3fb2adbe4bc00d72b77b3ab5ecab4d8a

Observation 8cc7a71d-ba13-4532-8ecf-be3043235c94 · inbound

UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling cites this paper.

UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling Learning and Leveraging World Models in Visual Representation Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:02.521473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:02.521473Z digest=sha256:b8cc6f080282cac5f38280067178bb0ba1d68db1aea3773feaf53c3c7406ad21