Pith. sign in

Paper Citation Record · LEDGER

VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 55 inbound Pith citation observations for arXiv:2503.06800.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.06800 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 55 of 55 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T05:13:22.343005Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:19:51.020711Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7f912f0d-3f95-48cd-954e-45de6f0ca019 · inbound

Generative Physical AI in Vision: A Survey cites this paper.

Generative Physical AI in Vision: A Survey VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 201

Resolution
unresolved
no resolver link, observed 2026-08-10T18:53:00.854751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:53:00.854751Z digest=sha256:2178b9f349f8068c5850bc5410d4fcc19b46964eec3374da2f389347cfe47a9f

Observation 101b1ba2-fbfd-4999-b886-9c911a6af147 · inbound

VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models cites this paper.

VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:05.096694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:05.096694Z digest=sha256:15727dea4eccc6c6ed8057ea38cd4e6a8f45db94f70550c0b9f4642cb9fa4f29

Observation 5952afe4-b42d-4238-8458-97351e4b5921 · inbound

"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models cites this paper.

"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:50.339759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:50.339759Z digest=sha256:7d5db00bbfa2615e6a8dfc7e34693e59334e431618122b2cd599adcdf825c46b

Observation 5408ffb6-a523-4df4-b901-5475e57ec79a · inbound

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models cites this paper.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:00.441134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:00.441134Z digest=sha256:b70b41bd1bccb47415d836cdf36ed357bbc8588d5c786c0e84266f96386964c3

Observation 8814fc66-f840-41c2-83fd-f5297fd85d97 · inbound

Human Preference-Aligned Concept Customization Benchmark via Decomposed Evaluation cites this paper.

Human Preference-Aligned Concept Customization Benchmark via Decomposed Evaluation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:59:25.198610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:59:25.198610Z digest=sha256:07ac58e5d65b69ed827a69ff8b20405b5c47098f410194cd9059e503507bf609

Observation 3512bfe1-4b8d-4c9d-99a2-9424d273ac60 · inbound

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility cites this paper.

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:56:24.343732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-18T12:55:42.679016Z digest=sha256:3eb7ca927965c1308a26d1bbf97afec4c97e632142801773d9d297d3d9eb3969

Observation feea1a15-9db2-4ab4-8253-1d0e571742b9 · inbound

World Simulation with Video Foundation Models for Physical AI cites this paper.

World Simulation with Video Foundation Models for Physical AI VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-12T23:01:13.918325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-12T23:01:13.546110Z digest=sha256:8284748e3912b6cb9d22af367c44a9c0536ccbdac31efef91717ad543195a723

Observation 500a90b3-6903-4d27-8f9b-9b60434e040c · inbound

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models cites this paper.

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:59:08.743022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-17T05:55:11.495430Z digest=sha256:151548ef21cb6c7196dae89de7d46f09ca8c1e9b653f3d39325a56c7c1ff0ad1

Observation 2eaf1f6a-91c2-4d27-b10a-c018831d97ce · inbound

Generative Action Tell-Tales: Assessing Human Motion in Synthesized Videos cites this paper.

Generative Action Tell-Tales: Assessing Human Motion in Synthesized Videos VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T19:11:47.098371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:11:47.098371Z digest=sha256:e14d37fb298125fe18aaf6b1f91cdc4ecb8d9990b702f3106618595ce3e176dc

Observation 7661578d-9147-4d74-b3f5-cc93263bf7bd · inbound

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models cites this paper.

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:00:27.237634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-21T17:57:57.263574Z digest=sha256:5f8e3e84e4a55eb10bc4a1b765aa8c75d2bd9b5f80cd0def85e4c8da58349262

Observation 8a6a0aeb-f9d3-42e4-80ff-be1b46426c19 · inbound

ProPhy: Progressive Physical Alignment for Dynamic World Simulation cites this paper.

ProPhy: Progressive Physical Alignment for Dynamic World Simulation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:08:47.998118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-17T01:05:08.087136Z digest=sha256:2d332bfa6532a4df25f976b1ebc51d306f69d023fd99c153dbee379c4b7b5a76

Observation 98814dd8-c6bb-49dc-8692-50fc179782af · inbound

Agentic Physical AI toward a Domain-Specific Foundation Model for Nuclear Reactor Control cites this paper.

Agentic Physical AI toward a Domain-Specific Foundation Model for Nuclear Reactor Control VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:00:23.976691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-21T16:57:19.490074Z digest=sha256:9a441d54fcc81e81467ebcb58b7e8b1f9e0e85f9e60e1dd23fc6c0fbb895a303

Observation 4ed3be5e-8e4f-4170-a9ab-4d69a8c8e3c6 · inbound

Self-Refining Video Sampling cites this paper.

Self-Refining Video Sampling VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T14:40:14.478534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-21T14:37:57.167882Z digest=sha256:82e29fccf8f0bb97467bbe3ba4b6886b2cd41fade3b5157cfc07b129cf95cadf

Observation c657b1b5-45b9-42b6-b311-c1d46f519cb2 · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 170

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:05:51.471254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:d8017c6710cd750045f74f83d41248a84b7fb5bc9494d1de6998f6dc24db5087

Observation ddbbf14e-72a5-420a-abb7-34edde3c7d7f · inbound

MoRight: Motion Control Done Right cites this paper.

MoRight: Motion Control Done Right VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:26:01.050690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T17:38:03.776766Z digest=sha256:b4f12815bf795f36eabbe8c0ff70965b2c23fcb8a76f17f001a332fcbde4ded3

Observation 6f66a35d-6b39-4dba-991d-3817f187d81b · inbound

PhysInOne: Visual Physics Learning and Reasoning in One Suite cites this paper.

PhysInOne: Visual Physics Learning and Reasoning in One Suite VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:25:59.537119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T16:39:48.066744Z digest=sha256:57e6467903ce1025a9555b34915ddbd7852137ade0d8cdb01e3c85ed3d879d56

Observation 9864ebcd-f30e-4cc2-9c58-0e545b7f4650 · inbound

How Far Are Video Models from True Multimodal Reasoning? cites this paper.

How Far Are Video Models from True Multimodal Reasoning? VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:51:03.643358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T02:44:52.920816Z digest=sha256:76e00eb0c6752b9f886c96800f93ff673407ae56a3c1971caaac8ecb7f9aa486

Observation 199032ec-d394-47bc-b93f-e3e2d2bee3c2 · inbound

BRITE: A Benchmark for Reliable and Interpretable T2V Evaluation on Implausible Scenarios cites this paper.

BRITE: A Benchmark for Reliable and Interpretable T2V Evaluation on Implausible Scenarios VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:41:33.649224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-09T21:12:33.209353Z digest=sha256:1f243b46e8193c04a54025b03bf3df723830c320112ec43a3db7a8341c495471

Observation 236bef27-0804-4a57-a7eb-11e53c67c88e · inbound

Do Joint Audio-Video Generation Models Understand Physics? cites this paper.

Do Joint Audio-Video Generation Models Understand Physics? VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:50:56.767419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-11T02:12:04.230076Z digest=sha256:b9f8acfb159b26a6c545d83e55b023d1fad7215eff9792d7f2779f9cabb3619c

Observation 68ca9a9f-28a4-45f6-97c7-000e535ba148 · inbound

Do Joint Audio-Video Generation Models Understand Physics? cites this paper.

Do Joint Audio-Video Generation Models Understand Physics? VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:45:08.203083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:39:22.070629Z digest=sha256:5a3d5b445924f4c95bc73aeef37c6b40d2dd2da22d856ea1f8abc06158319051

Observation c4348ac0-7666-440b-b235-55241b7a45f7 · inbound

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models cites this paper.

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:40:54.609388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-11T02:21:52.861714Z digest=sha256:8e0cd9e4c160ec4fcf8fd269b19ae23362e7392799266f8d6eac0c18329e748b

Observation ae39e17c-04af-48de-bf8a-868a40164b10 · inbound

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models cites this paper.

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:15:08.007606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T23:13:31.195562Z digest=sha256:f9c218f80df07afad40b323d4e0420a7c56e78f7ebffc4a799e5478da727e539

Observation 42134aa3-572d-4f86-9217-752a2cba9445 · inbound

PhyGround: Benchmarking Physical Reasoning in Generative World Models cites this paper.

PhyGround: Benchmarking Physical Reasoning in Generative World Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:21:27.838360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-12T05:17:30.010064Z digest=sha256:02b554b5359bcaf420f2afc734a45d4acc520a1c46de9dd6320049941d046b89

Observation c6006889-d1ff-4cf3-818a-0b4733a448b9 · inbound

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating cites this paper.

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:02:22.162702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-13T06:00:31.582714Z digest=sha256:2e49eb11b11ec6609a3ee6708959752d6d6a256604e88e9b5c4b183faa983b46

Observation 2dac5e39-8a77-4e2b-a560-ced9b61bfc1b · inbound

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating cites this paper.

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:55:45.294459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T22:38:43.102769Z digest=sha256:28607149c736b4a9e39b7b146f98234c7fb23eeb12d118d736271f0a72523193

Observation b5c2cfd1-1775-4c66-bdd3-46d10217d19f · inbound

DriveCtrl: Conditioned Sim-to-Real Driving Video Generation cites this paper.

DriveCtrl: Conditioned Sim-to-Real Driving Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:15:04.778745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T21:06:06.548538Z digest=sha256:e0e4f4679d730873200f2545dd9ecf10aa2007ba9cd16ce2f77b0bc14ac1448f

Observation 881ec02c-7a65-4267-a803-556ef6b74eb5 · inbound

NEWTON: Agentic Planning for Physically Grounded Video Generation cites this paper.

NEWTON: Agentic Planning for Physically Grounded Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:13.762531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-20T10:56:22.343333Z digest=sha256:292cf1beb66f66be9b6a85d30caa94e935689318169ca58545b959a70cd23bb7

Observation a98e831c-1b5f-4ca3-91d7-63573d7e2c39 · inbound

PhyWorld: Physics-Faithful World Model for Video Generation cites this paper.

PhyWorld: Physics-Faithful World Model for Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.595472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:7c4332dd3d43aace536c3266bcd766f2dda6dc0bdf106d088b9d32f2d4c49937

Observation 5cd09e5e-1276-445c-953d-e2da461e1bdb · inbound

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models cites this paper.

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:40:23.190574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-25T04:39:22.400458Z digest=sha256:6aaeee68d7126e7f8df879fc598cf83ad758a3fcf2de6c2b66a2ac5a43e96a0b

Observation c747a6d3-4c52-4ddd-b01f-5d5acc06334a · inbound

LaMo: Self-Supervised Latent Motion Priors for Physical Realism in Video Generation cites this paper.

LaMo: Self-Supervised Latent Motion Priors for Physical Realism in Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:45:20.481958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-25T04:42:32.717968Z digest=sha256:de05e2b94dcd3093321bcc81c9be164fa8213999d35ca9249fee1ea97a0c100e

Observation 3916976f-d93f-406c-9aca-fa3e7b81f9c7 · inbound

Tempered Self-Similarity Alignment for Physically Plausible Video Generation cites this paper.

Tempered Self-Similarity Alignment for Physically Plausible Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:44:38.371299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T11:39:06.597513Z digest=sha256:8b3951fa15c767c2e075f42e7ad16232d25ab2a6a8fe84ec927f8c40305c201a

Observation 31fa3a2e-2dd7-4854-ba29-969d8e717dd3 · inbound

WBench: A Comprehensive Multi-turn Benchmark for Interactive Video World Model Evaluation cites this paper.

WBench: A Comprehensive Multi-turn Benchmark for Interactive Video World Model Evaluation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:14:02.269296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T22:57:08.381846Z digest=sha256:0dac227f74e0b1b6169bef1f6aa61545e09c801e075820270b889337edbd44b5

Observation 2dc8fd89-4e71-47a0-8c81-0f5c24d823d7 · inbound

What-If World: A Causal Benchmark for General World Models in Embodied Scenarios cites this paper.

What-If World: A Causal Benchmark for General World Models in Embodied Scenarios VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:50.403328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T18:23:22.987086Z digest=sha256:7f976f7c36f01a98abcfec029b5f46c202fb960066dbfd0a15034dcfeab52971

Observation 928a2668-cecb-4338-97a9-034a1af2aa1c · inbound

Proprio: Latent Self-Scoring and Inference-Time Refinement for Physically Plausible Video Generation cites this paper.

Proprio: Latent Self-Scoring and Inference-Time Refinement for Physically Plausible Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:03:26.633426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T12:55:24.689338Z digest=sha256:d884efbb5c78e9ff6cdfbcb6175b55877b04c4f8e04c817056bc3f468a5242b0

Observation b092f7a9-95a8-4f84-9bd4-cd2b18a83278 · inbound

YoCausal: How Far is Video Generation from World Model? A Causality Perspective cites this paper.

YoCausal: How Far is Video Generation from World Model? A Causality Perspective VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:33:15.610723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T08:27:03.674229Z digest=sha256:cce26b6be5c97e9b29fe78260cf43a7e2a9edc20baa996bdb9e2b6defd082e43

Observation 2a5c3853-3881-45ff-821b-facdf2267b1e · inbound

OptiWorld: Optimal Control for Video World Generation under Physical Constraints cites this paper.

OptiWorld: Optimal Control for Video World Generation under Physical Constraints VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:32:35.562695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T19:02:51.848742Z digest=sha256:9cc23c95d510c02a86ce776b768be26edd111e59b3ff84ad5479b0de024074af

Observation c199f755-4fc8-4cc0-8f1d-1a8b4870040a · inbound

MPMWorlds: Material-Point-Method Simulations for Inferring and Extrapolating Physical Dynamics cites this paper.

MPMWorlds: Material-Point-Method Simulations for Inferring and Extrapolating Physical Dynamics VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:16:24.684102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T12:19:27.221596Z digest=sha256:ffa2d0178b1619553e635dba7c10c111635fdaef699a4a5c70f4e4d9a14272c5

Observation d8879029-c4bb-428d-9030-f6d732fd5ae2 · inbound

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment cites this paper.

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:16:44.758956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T07:02:37.291472Z digest=sha256:39f31b5df5bb82795b73f8f441c5ea6c8bb5cfb49ff4ca2760e4a7ee3b102c06

Observation db7ea51f-8389-43de-bceb-5e4a862c72a0 · inbound

Current World Models Lack a Persistent State Core cites this paper.

Current World Models Lack a Persistent State Core VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:49:31.017133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T17:33:41.461245Z digest=sha256:1c5b01b6cd7c82bcb1f7d19c3ae3a39d8fda76526919822e50bd25cfb6adee6f

Observation 7e2c8aa6-6fec-46e8-a275-c5d9d023e5a7 · inbound

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation cites this paper.

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:09:45.415761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-26T09:04:23.965554Z digest=sha256:b9dd710457a401d9d1bda7799d887984762bf6afabd0e32517286a676e383ab2

Observation 03148a37-6e8a-4ca4-a273-bb1a937b2c8b · inbound

PhysRAG: Enhancing Physics-Awareness in Video Generation via Retrieval-Augmented Generation cites this paper.

PhysRAG: Enhancing Physics-Awareness in Video Generation via Retrieval-Augmented Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:19:51.022182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T05:16:53.011837Z digest=sha256:9e4d4229ee8621ef43ac16849727e84189c860b9a3776fad561a49cd71b57ca0

Observation 5b8ddd1a-9cf9-47f0-88fd-f3abfedb3e5a · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T18:25:57.779239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T02:03:45.564122Z digest=sha256:a496c751929f03cb104b3e948173849e598a07ec4a97db52d4a68c84980172db

Observation 776a5435-d44c-4177-9a53-4db6cf6b2e37 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:35:40.499609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-01T06:25:58.872140Z digest=sha256:d885242847ad4dca14208fa47272618aed30ac827882b1926566254220b540b7

Observation 9cd71609-d385-4009-b531-d27eaf95c46a · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:22.806256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-02T20:52:28.444524Z digest=sha256:4eaf1a9c04e130af027117d0b8be67ab4b54f14cd9a7870c6124979786101044

Observation fbfc599f-82d6-4ac7-9073-49c0a272493f · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T22:49:00.899176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-03T22:44:16.272541Z digest=sha256:36cc2e80b541b3a2c97591e436c3d7a2efa85db247d226949f64a30e0fb62824

Observation 06aee8cc-d380-4e35-81c5-407ff4d20c72 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T17:14:19.770867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T17:14:19.770867Z digest=sha256:8dcf311d53732310c88e39446901e801fe75e21cdcdb0b245af63f5807d139ba

Observation 911c1311-d029-402b-81d2-40a9d4c148af · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T10:00:09.187494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:00:09.187494Z digest=sha256:891a78bb1ca4243c36a34c76b02d3240833d4b8615c60b3c6ce0403328857b93

Observation 42136782-eba8-4cb4-8ca0-4f81411e752f · inbound

Apple-$\pi$: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence cites this paper.

Apple-$\pi$: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T21:07:22.906584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:07:22.906584Z digest=sha256:929c2bc3d2ae38277dc54dae46ecf0ee70f38e7437d22a8d559dc23508d5d904

Observation 087d6028-ce69-4bf7-bf1d-e1f662b0f4e2 · inbound

When Physical Preferences Meet Semantic Constraints: Physical and Semantic Direct Preference Optimization for Text-to-Video Generation cites this paper.

When Physical Preferences Meet Semantic Constraints: Physical and Semantic Direct Preference Optimization for Text-to-Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T19:33:03.502183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:33:03.502183Z digest=sha256:fe0ad70701dc9673fafd4e363e85e3b72ff892bc191310199bec61c266db847b

Observation 5920ead6-c2f8-48fb-9f7d-3462801e86f6 · inbound

Thinking in Video: Can Video Generators Really Reason About the Real World? cites this paper.

Thinking in Video: Can Video Generators Really Reason About the Real World? VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T17:47:06.272600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:47:06.272600Z digest=sha256:960abc2995c089b81322ecb0eb4c8dc6eb51acfd2eba9d6c69328a6e8b356487

Observation a3f15e8b-0fbc-4760-a54c-db2a3b675253 · inbound

Learning Explicit Physical Parameter Control and Benchmarking for Video Generation cites this paper.

Learning Explicit Physical Parameter Control and Benchmarking for Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T14:02:10.808123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:02:10.808123Z digest=sha256:2d876777303fa480e515df69876cd090fdefed5bb4cbb8aab6f4f48becb430d2

Observation a6a7e5c4-8726-45ce-a12b-a2e89635b934 · inbound

Physics-Grounded Fluid Video Generation with a Simulation Dataset and Dual-Stream Optical-Flow Supervision cites this paper.

Physics-Grounded Fluid Video Generation with a Simulation Dataset and Dual-Stream Optical-Flow Supervision VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T02:49:57.720734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:49:57.720734Z digest=sha256:0a0fd7d3606c656afd95f146dea408469133dc6934ad05d4a585e64a6d37a46e

Observation 73a0adac-4e87-49e6-b669-293c9e9c6aea · inbound

VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System cites this paper.

VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T08:28:04.629823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:28:04.629823Z digest=sha256:c1528e704f0f148673af9b30497064948c826abd9715dd4507f1b6d9875dec69

Observation 824e9a33-f5e3-4323-b33b-9064aae12a26 · inbound

GAUGE: A Measurement-Grounded Benchmark for Physical Fidelity in Simulation Engines and Video World Models cites this paper.

GAUGE: A Measurement-Grounded Benchmark for Physical Fidelity in Simulation Engines and Video World Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T20:36:25.525906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:36:25.525906Z digest=sha256:04ce7c09efdad739ddcc87a37820c8523d072427c7432bb5c302a89da1dcc05e

Observation 239471a0-f019-4c4e-bf3c-aaedcba1138e · inbound

Sci-VBench: Evaluating Knowledge- and Reasoning-Intensive Video Generation in Science Domains cites this paper.

Sci-VBench: Evaluating Knowledge- and Reasoning-Intensive Video Generation in Science Domains VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T05:13:22.343005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:13:22.343005Z digest=sha256:2b9ded8aa8234ae5a07de0dea5921991d9ef633a41a23956fba6e4c891cc2659