Pith. sign in

Paper Citation Record · LEDGER

VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2503.23368.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.23368 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:34:16.094906Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:20:05.840081Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2a742e0c-8837-4f4f-8bb2-41edbf20e9c1 · inbound

A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality cites this paper.

A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T18:51:31.233075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:51:31.233075Z digest=sha256:5df425df59cbb49c847fc63728756062c9de34be43912a247b55f1f34d22e8a8

Observation 84627d0a-877e-42ef-af92-e3f490eb5a14 · inbound

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility cites this paper.

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:56:24.394265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T12:55:42.679016Z digest=sha256:059cf9df12ea7ee9d81762ba50c014997920af268712d2ab5e1bb9916a9c3680

Observation ff32deee-3011-4c9c-815f-bc563eef382b · inbound

Self-Refining Video Sampling cites this paper.

Self-Refining Video Sampling VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:40:14.482113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T14:37:57.167882Z digest=sha256:c97fc0c652a2ded886e43e17256d4c02e29f4611802acbdb9634b972cb092e2f

Observation b431a42f-d677-4464-a42b-c033990e3954 · inbound

Vision Language Models Cannot Reason About Physical Transformation cites this paper.

Vision Language Models Cannot Reason About Physical Transformation VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-15T13:27:51.848177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:27:51.848177Z digest=sha256:81b7f5f6fe8c36736d4238afc7ebee1e0823c5de0003976ea7275c58c2971852

Observation 79bd86d6-9257-4fc5-bd67-4474dcee7d62 · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 300

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:05:51.719526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:e805b196a1ded9b57fc00fae6a18f7b18be993fdd3fc14fd2021130ae8952dc8

Observation a083338f-2e36-480f-9393-1af2799bc958 · inbound

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving cites this paper.

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 137

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:31:26.419010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T04:13:37.421188Z digest=sha256:0553249c731c4422d7f4f8848117ab5c2422903f73eddac06a4026a244ab9d1a

Observation 494e971f-48f1-4b80-bfb2-dec0392206a3 · inbound

OptiWorld: Optimal Control for Video World Generation under Physical Constraints cites this paper.

OptiWorld: Optimal Control for Video World Generation under Physical Constraints VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:32:35.539164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T19:02:51.848742Z digest=sha256:660598c1ca03b3cee3605628412470e242bce1699d12bf3b9c7726a6d7423875

Observation 8cbd19d9-eb58-466c-ac65-eeab1a9da6da · inbound

Physics Question Scene Graph: Fine-grained Evaluation of Physical Plausibility in Text-to-Video Generation cites this paper.

Physics Question Scene Graph: Fine-grained Evaluation of Physical Plausibility in Text-to-Video Generation VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:20:05.841727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-25T21:33:38.643889Z digest=sha256:08b534f442cb0ba98f48c778d477d4013a7059d6df044b52e5718d490857aaba

Observation 32225ca1-f466-4642-acc8-ea40ea9fd754 · inbound

muSync-GS: Physics-Synchronized Driving Video Synthesis for Weather and Geometric Road Hazards cites this paper.

muSync-GS: Physics-Synchronized Driving Video Synthesis for Weather and Geometric Road Hazards VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T00:34:16.094906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:34:16.094906Z digest=sha256:5a4494a8df4a1089e8f3be31b3a76b80567a7566af22bea93d9d8b015cbc56b6