Pith. sign in

Paper Citation Record · LEDGER

VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2505.23656.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23656 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T10:35:06.650603Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:59:58.033217Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b5a4a27d-6d6a-47e5-9999-9b24a1c0b084 · inbound

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling cites this paper.

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 94

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T05:17:06.569867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-19T05:13:28.767788Z digest=sha256:515ba3a11142d75712a2e1112d5d0ad968527cd1498267b7da4e660e24e0488b

Observation b39a40a2-396c-4989-9eeb-1d6698f25ca4 · inbound

Aligning Perception, Reasoning, Modeling and Interaction: A Survey on Physical AI cites this paper.

Aligning Perception, Reasoning, Modeling and Interaction: A Survey on Physical AI VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:01:13.996334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T09:56:36.716680Z digest=sha256:ef1e2b5300ff29e06f1a59d41729e4b98f16d054d211ba4783e02ca7d5d63927

Observation 347d065b-bf02-47cf-93e4-dc32cde8e675 · inbound

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models cites this paper.

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:00:27.218289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:57:57.263574Z digest=sha256:07e307a0f1e288acce664323e99a1d7fd36bf820648fc3b88790e4853c665cdc

Observation 69637c36-5420-458d-a69b-d39cf17d89a6 · inbound

Transition Matching Distillation for Fast Video Generation cites this paper.

Transition Matching Distillation for Fast Video Generation VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-03T10:35:06.650603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:35:06.650603Z digest=sha256:e402fdb8d2bf70e9943657ded737ab0f0e3741e804eea1e4658e0d1b0067a059

Observation fea40679-2fe9-42ba-a4d7-980a8fee9702 · inbound

Olaf-World: Orienting Latent Actions for Video World Modeling cites this paper.

Olaf-World: Orienting Latent Actions for Video World Modeling VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T01:20:05.776188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:20:05.776188Z digest=sha256:dbd43ebdd1579b9824672eb5717924b4e5a427fefad55bea383c48902fe93294

Observation f5a7bc71-0409-43f1-ab5f-f5fe57d6eed1 · inbound

Under One Sun: Multi-Object Generative Perception of Materials and Illumination cites this paper.

Under One Sun: Multi-Object Generative Perception of Materials and Illumination VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-13T22:08:47.493022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:08:47.493022Z digest=sha256:d63ba6598e76db8fcb02edaee9d3fd0cb50116f0d5c108b59f1f3095af145f76

Observation b8350745-ef14-44b6-b0c1-2fb281f9fa2e · inbound

Human Cognition in Machines: A Unified Perspective of World Models cites this paper.

Human Cognition in Machines: A Unified Perspective of World Models VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 221

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:12:26.004914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T08:12:15.663761Z digest=sha256:89a38cdea6554b293a99f1615d4641310100177aa379126d93dd80c69d757623

Observation ddfce7a8-fb29-4369-8fe4-64e89c4cea8f · inbound

Divide and Conquer: Decoupled Representation Alignment for Multimodal World Models cites this paper.

Divide and Conquer: Decoupled Representation Alignment for Multimodal World Models VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:46:02.916958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:51:12.604102Z digest=sha256:770acc3e2b38d232fd0a38648e088bb78662109b9e0bdab5ed5ddf7d37f92509

Observation cfba5c5f-3f70-4db3-a916-26d354c6e17d · inbound

Divide and Conquer: Decoupled Representation Alignment for Multimodal World Models cites this paper.

Divide and Conquer: Decoupled Representation Alignment for Multimodal World Models VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:57:27.925208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-02T23:56:39.865433Z digest=sha256:2414c6bbd1480eb35fd17f70bf97ca8ab51fa58bcff203435aabb9af8a9e0cbe

Observation d844438a-eb74-4af3-bbda-a7a2a3b2ab97 · inbound

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models cites this paper.

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T02:25:54.965159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T02:21:52.861714Z digest=sha256:4d6f08a06c5db66befdfd0452a8f5ec36413bf1343e7ff9d24838bb9bbfc9730

Observation bef583c8-5140-4de0-8c28-eb5ae6021bc8 · inbound

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models cites this paper.

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:15:07.975899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T23:13:31.195562Z digest=sha256:66c31db78c5f60dbf9e0a9c2c72d4651f7ae0c7ded6fee2e56d4ca98229a80e2

Observation dee2aa1f-2305-4f4a-bc05-e1b9005fd7ac · inbound

Improved Baselines with Representation Autoencoders cites this paper.

Improved Baselines with Representation Autoencoders VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:43:15.208972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T11:40:14.358108Z digest=sha256:8979024a6c524be9c62dcee71b0946978c051a078cf545d944d764f968826294

Observation cf386c95-a14b-4de8-8545-7dfc4da47dfc · inbound

GeoFlow: Enforcing Implicit Geometric Consistency in Video Generation cites this paper.

GeoFlow: Enforcing Implicit Geometric Consistency in Video Generation VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:08:13.552484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T11:06:09.367559Z digest=sha256:1dd1cb8e62db58d469f094469a3b4842c08a176854f7c305a5d5f7109cc47f1b

Observation e4378f01-936c-4241-8c9f-b3310203a143 · inbound

Spatial Gram Alignment for Ultra-High-Resolution Image Synthesis cites this paper.

Spatial Gram Alignment for Ultra-High-Resolution Image Synthesis VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T05:29:39.647375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T05:26:29.331382Z digest=sha256:6ecd79b6332357a850c834a1ac58bf6cde9754c18dc0acf3a22357d3218357b5

Observation 5a6330e7-96b9-40a6-b95f-eeea49b81e39 · inbound

GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation cites this paper.

GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:36:39.673333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-25T05:34:03.684055Z digest=sha256:d975d4b0fc4d6d142512b16b8326c82c985bf01b67d2920ce6c32059bef4abaf

Observation 06816102-79a0-4900-aef0-2137740776ec · inbound

GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation cites this paper.

GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:59.244185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T16:45:15.955954Z digest=sha256:a8b1561650d788dbe61274d3e19cc9ea148277d98a8c18cb743cc67827ed4ac9

Observation 69bacd6a-cb6c-4e45-b530-b31d7f79b389 · inbound

Tempered Self-Similarity Alignment for Physically Plausible Video Generation cites this paper.

Tempered Self-Similarity Alignment for Physically Plausible Video Generation VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:44:38.419321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T11:39:06.597513Z digest=sha256:869bfb0bb4288c9ed79ea0ec4fd7c5c12de0c3b86416d143ca2be6b425d675be

Observation 8a7e44a6-7118-4dbf-aefb-16e6097fecee · inbound

VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization cites this paper.

VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:26:17.292794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T15:26:21.284810Z digest=sha256:5428aeaed91d60e93d5ff52219aa68be6830eb95b698ef13dfeb276ff529e55e

Observation 41590636-0f5c-4401-8ca2-229780b66fb2 · inbound

VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization cites this paper.

VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:44:36.814378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T10:38:22.619277Z digest=sha256:a454a1c11189a2c33bd99fb341b58f30ddfab569a24df5581ca55399cb462864

Observation 82558469-6c06-4310-a168-d1df9ffc5073 · inbound

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment cites this paper.

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:16:44.766781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T07:02:37.291472Z digest=sha256:a546316a0d23f962669a2c2794ed436c8bf43e2100183aee13b4ed3327f4dbb3

Observation fa4e91e3-0bc4-46b1-9ce4-87de4c77cc3e · inbound

MaineCoon: Pursuing A Real-Time Audio-Visual Social World Model cites this paper.

MaineCoon: Pursuing A Real-Time Audio-Visual Social World Model VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:48:55.711321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T01:09:02.214590Z digest=sha256:46d001a7b6fb7429391926f31bc7e740c09e38dd4eee340c592381fddf2671ec

Observation a3d93cff-ae59-400c-9522-85b5c75c3b6a · inbound

DiffusionBench: On Holistic Evaluation of Diffusion Transformers cites this paper.

DiffusionBench: On Holistic Evaluation of Diffusion Transformers VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:59:58.035148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-26T00:06:11.951205Z digest=sha256:38ece817e84f0ad9bbb68dde311c2896ceacb559fd4b5237609e69efc9ce5f9e

Observation 0cf101dc-2c4b-4fee-ad47-32c22d7f180f · inbound

PhysRAG: Enhancing Physics-Awareness in Video Generation via Retrieval-Augmented Generation cites this paper.

PhysRAG: Enhancing Physics-Awareness in Video Generation via Retrieval-Augmented Generation VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 90

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:19:51.068810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T05:16:53.011837Z digest=sha256:9828131978cfea953acef1c65a76f334eb1cd4a3c76158666c83eb290202510b

Observation 41115953-6c40-4c0f-b9ca-6988c3897b11 · inbound

PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation cites this paper.

PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-06-29T20:03:56.910145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T04:34:38.286863Z digest=sha256:c28692405289240154dfaf38e4bf9d518f883649c10ed8475ba855524d4a7f6d

Observation 9af6b0f9-eacb-41c7-a90e-2ac286a2674c · inbound

Aura: Consistent Multi-Subject Video Generation via VLM-Grounded Semantic Alignment cites this paper.

Aura: Consistent Multi-Subject Video Generation via VLM-Grounded Semantic Alignment VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-11T20:11:31.576642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T20:11:31.576642Z digest=sha256:573727b97901efa20a5d91b66a608f7a0aba6f8d63e7738ff463abf15ed71e6d

Observation 6d6d83d4-61b0-4f0e-a988-27a6bc4da859 · inbound

Enhancing Video Physical Consistency via Role-aware Joint Training and Modality-decoupled Denoising cites this paper.

Enhancing Video Physical Consistency via Role-aware Joint Training and Modality-decoupled Denoising VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 43

Resolution
malformed identifier
no resolver link, observed 2026-07-11T15:44:22.210969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T15:44:22.210969Z digest=sha256:3c064eb5971941a37ee1aabab6b53332efb18a6b4dd2cfacfc866083ac407c1a

Observation 9b091439-78e3-4a07-94a1-48dbf571c7d3 · inbound

VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders cites this paper.

VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T02:51:47.212234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:51:47.212234Z digest=sha256:23b448cd3a637c983a998d6df7b975922cf4c38bda118a8d1a6d84ab933ede97

Observation 52c22c41-0074-482b-bf14-43e8adf5a7f5 · inbound

AgentHOI: Multi-Agent Reasoning for Human-Object-Interaction Video Generation via Implicit Representation Alignment cites this paper.

AgentHOI: Multi-Agent Reasoning for Human-Object-Interaction Video Generation via Implicit Representation Alignment VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-01T05:27:20.293209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:27:20.293209Z digest=sha256:c317a66d34f1cf0ad1c91c7d83a76a7bd1b8da3d9ea64867081224829b3a2093