Pith. sign in

Paper Citation Record · LEDGER

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models

As of 17 August 2026, this Paper Citation Record lists 100 of 106 outbound references and 2 inbound Pith citation observations for arXiv:2412.11785.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11785 v3

Coverage vector

measured 100 of 106 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:39:19.337878Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:53:00.987266Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T18:53:01.110433Z

Reference resolution

100 of 106 outbound references displayed

  • verified exact0
  • verified fuzzy56
  • unresolved44
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d3418a8d-efac-44b3-ac3e-d1d4051bd916 · outbound

This paper cites Lumiere: A space-time dif- fusion model for video generation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Lumiere: A space-time dif- fusion model for video generation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.764031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.764031Z digest=sha256:32a4624a07c92fc22024c22110aa3318fa7f37fc17ce86161b63f1e3f0c0f535

Observation cfe2ae80-0330-4a02-bb49-f73a50882d81 · outbound

This paper cites CoPhy: Counterfactual learning of physical dynamics.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models CoPhy: Counterfactual learning of physical dynamics

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.769580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.769580Z digest=sha256:c87160c481645f40d60747378945bf60ccbf1fdd09fe5d163fe365e98dd4a582

Observation 1047c437-3a48-478d-aa48-82329a9b0054 · outbound

This paper cites Interaction networks for learning about objects, relations and physics.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Interaction networks for learning about objects, relations and physics

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.775330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.775330Z digest=sha256:f80827f85f009fa95fa482a35baf59ea43854525cbadcb0563bfa7319b9f2d21

Observation 7c92db3c-fea3-4314-bdea-807c8453234c · outbound

This paper cites Sutherland, Michael Arbel, and Arthur Gretton.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Sutherland, Michael Arbel, and Arthur Gretton

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.780815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.780815Z digest=sha256:d9dfa0c38929dfb925967be26dadcf270d8522cf4faa68057504dbf5dd6021bd

Observation 392eddf7-40ff-4f30-9b7a-3a9eaaa385b6 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.785697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.785697Z digest=sha256:f25471586894a5419fa5bd6f1a2f6daeaed0a415bd9e4d525cad653f0e375642

Observation d9dfbbed-b036-4cad-b21a-a0187051f9af · outbound

This paper cites Align your latents: High-resolution video synthe- sis with latent diffusion models.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Align your latents: High-resolution video synthe- sis with latent diffusion models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.791185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.791185Z digest=sha256:a8ffd7af3a3a86f8efa6676a97dc05700dd8a3ae02da014cc11a5e02675ef214

Observation ea3f3019-5bdc-4db0-86af-3d0f043ad8db · outbound

This paper cites an unresolved cited work.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.795866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.795866Z digest=sha256:7155d735e28567e46b6c28daa083883b14afcb3b0157ad957edc1bc415143a1e

Observation 68110bc5-b0cf-4056-8ec3-671853cf97ef · outbound

This paper cites Generative rendering: Controllable 4D-guided video generation with 2D diffusion models.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Generative rendering: Controllable 4D-guided video generation with 2D diffusion models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.801094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.801094Z digest=sha256:632643d9e7cfc247b0d861b5eada06723b1b041e813cb57b4b2c4aa58801863f

Observation 5a7fdc3d-dcd2-4e2d-b540-ddf19ef5e67a · outbound

This paper cites Realtime multi-person 2D pose estimation using part affin- ity fields.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Realtime multi-person 2D pose estimation using part affin- ity fields

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.806242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.806242Z digest=sha256:f8f50893d5dc8a3d577b0dcadf3e00e6f3971a3f7a59f16ca1d4fa0b67c5f7d4

Observation 877abcfb-67d6-4c21-a2e2-7f4ef9107ca4 · outbound

This paper cites Text2HOI: Text-guided 3D motion generation for hand-object interaction.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Text2HOI: Text-guided 3D motion generation for hand-object interaction

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.811206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.811206Z digest=sha256:58a7827686f9e65099f5798fabf368e3f21eee7989b9564dfc4e4a6f28b6ce15

Observation 1c47ca28-9fd8-4871-9d72-a8e279d2b687 · outbound

This paper cites Narang, Karl Van Wyk, Umar Iqbal, Stan Birchfield, et al.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Narang, Karl Van Wyk, Umar Iqbal, Stan Birchfield, et al

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.816252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.816252Z digest=sha256:30ecc1a41fc705848c4cb0c119861056820134ed26b0c9f60535193a353ca0aa

Observation 6899e66f-14c7-4dd8-b354-022f20d66d90 · outbound

This paper cites VideoCrafter1: Open Diffusion Models for High-Quality Video Generation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.820898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.820898Z digest=sha256:d92a946bce6d7f0a77783de387b8a73eddaf00a685de612548bad70074b74aca

Observation b0d78bb3-390c-4d14-877e-098f328a856d · outbound

This paper cites Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.826123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.826123Z digest=sha256:5e94310ea492213f1cbad0b5685c61737f38623592b72d8744d24cca22f61423

Observation f4f989aa-8832-4383-8f33-df115843f1f3 · outbound

This paper cites GanHand: Predicting human grasp affordances in multi-object scenes.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models GanHand: Predicting human grasp affordances in multi-object scenes

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.831794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.831794Z digest=sha256:8ddc32ce06e47409f959a9390b4eb4ea4d3ff86b5b9c7c22b706e6c6e12d16ef

Observation 08238dfe-ba2a-4327-aa2a-ae0443eb2450 · outbound

This paper cites CG-HOI: Contact-guided 3D human-object interaction generation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models CG-HOI: Contact-guided 3D human-object interaction generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.838145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.838145Z digest=sha256:19dd2dcf3cf47523973e4a69cabd77fb86d5a2f848dc2fde2b5b99d2639b5738

Observation cee148ce-1eca-4d4e-b1ae-bb00c0109061 · outbound

This paper cites Scaling recti- fied flow transformers for high-resolution image synthesis.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Scaling recti- fied flow transformers for high-resolution image synthesis

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.843788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.843788Z digest=sha256:f5ee25bcf429ccf6e7a6e847aa73a497809ed13e689134cd626b59ce6fddd61e

Observation 71a23daa-c3d6-4cbf-b45a-817cc7889bf6 · outbound

This paper cites Black, and Ot- mar Hilliges.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Black, and Ot- mar Hilliges

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.848903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.848903Z digest=sha256:b76effddcb50b49b8832daabf94ea3b1aaac59769d51b534e0a3930c5f5e38b3

Observation d56e87a9-6be1-413d-9c5a-0d21f5186d91 · outbound

This paper cites Black, and Otmar Hilliges.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Black, and Otmar Hilliges

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.854384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.854384Z digest=sha256:e62970981f0d9bf1c1baa67aa174ca785184a22bb35a4c674040579b1580fe57

Observation 5c6b6f97-0ac8-42fb-813d-6ef4736cd20f · outbound

This paper cites Phys- ically grounded vision-language models for robotic manip- ulation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Phys- ically grounded vision-language models for robotic manip- ulation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.860548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.860548Z digest=sha256:2cc89f043c152e7e32e4951b7b7d325236aeeed3192bbda42ead8b6b0ab01b41

Observation 6460b9d7-96e7-4c02-98bb-062c160eb37c · outbound

This paper cites Emu Video: Factor- izing text-to-video generation by explicit image condition- ing.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Emu Video: Factor- izing text-to-video generation by explicit image condition- ing

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.865226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.865226Z digest=sha256:8881f5f858bfad2b2331849520316ca1db0743c89d2b4e9f7604599023cfaa15

Observation 04590a54-3bf1-4331-a6b7-62e4703782b5 · outbound

This paper cites The something something video database for learning and evaluating visual common sense.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models The something something video database for learning and evaluating visual common sense

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.870989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.870989Z digest=sha256:c19e4eba7ab0375dcf2c0760917ec6556c83f8967bacf7697092d1d52393e446

Observation 85c40167-4767-4546-94b0-e4b11808199c · outbound

This paper cites Fuchs, Ingmar Posner, and Andrea Vedaldi.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Fuchs, Ingmar Posner, and Andrea Vedaldi

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.875669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.875669Z digest=sha256:070b8dc63a16df87a556d28d8ac23521ee0e5e7aee42031cf4237d3702d376db

Observation e7633788-82bd-4cb3-88c2-056a1d250be6 · outbound

This paper cites Seer: Language instructed video prediction with latent diffusion models.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Seer: Language instructed video prediction with latent diffusion models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.880958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.880958Z digest=sha256:dd5e1600304fc3eb03b153e63452b8a3a9a9c71f8032e68fa676bbe85f593881

Observation d0d81ae5-da92-453d-a88a-07a21ddc5843 · outbound

This paper cites Disentangling physi- cal dynamics from unknown factors for unsupervised video prediction.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Disentangling physi- cal dynamics from unknown factors for unsupervised video prediction

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.886509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.886509Z digest=sha256:0cb245f471d108720a322cec671587d1e394b8e652f0ea6ca51a36e126cd9b0f

Observation ed659e42-1d36-45f4-8971-84c2429940c2 · outbound

This paper cites I2V-Adapter: A general image-to- video adapter for diffusion models.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models I2V-Adapter: A general image-to- video adapter for diffusion models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.892351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.892351Z digest=sha256:1fc4174a0a912ddc5b724abbaf1ff5c2c7735471e34157ac033a79611d8b2de6

Observation 4396a781-aa7a-41c2-b08f-3497856924c6 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.897368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.897368Z digest=sha256:c83a0e4742da094b483af74c1fe2ff00df15f5cb1b4cad6bc83a7ccf32dc5caf

Observation ab2f2181-be79-4388-87a7-737ad9f097a8 · outbound

This paper cites Black, Ivan Laptev, and Cordelia Schmid.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Black, Ivan Laptev, and Cordelia Schmid

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.910659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.910659Z digest=sha256:6a264f6bb9b7909f496f1e3b55f57a794bd0746c185e6cf9bd4eb507c8b3c940

Observation 7db21885-b7f1-4cbf-8e3b-0fe944e49080 · outbound

This paper cites Towards unconstrained joint hand-object reconstruction from RGB videos.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Towards unconstrained joint hand-object reconstruction from RGB videos

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.916150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.916150Z digest=sha256:787595c1654057f3591f154df962c47a7a57caed175ad9319def1536ae719e35

Observation a2149235-bf70-4709-b52a-0b675e2d97c7 · outbound

This paper cites CameraCtrl: Enabling Camera Control for Text-to-Video Generation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models CameraCtrl: Enabling Camera Control for Text-to-Video Generation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.921248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.921248Z digest=sha256:efc7f387fe7e35a68619f2b752684e5b382d2a6f9cf8c65e566b20ed2bfdfe1f

Observation e331d359-561f-4baf-b955-ac5b15c986c6 · outbound

This paper cites GANs trained by a two time-scale update rule converge to a local nash equi- librium.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models GANs trained by a two time-scale update rule converge to a local nash equi- librium

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.926762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.926762Z digest=sha256:24b91a00833bc64ac0e58e77fbd292be8e595e076c98b04ef97eb305bc2e1fef

Observation 39832f06-c058-40c8-969d-2b367b986a9c · outbound

This paper cites Classifier-Free Diffusion Guidance.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Classifier-Free Diffusion Guidance

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.932193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.932193Z digest=sha256:ebcb63792aa0243860a0b01c793ddd8b77005a1c50004d71b6df327487e2cc64

Observation cfcb304a-f5fa-440c-8904-36732bbe5e82 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Imagen Video: High Definition Video Generation with Diffusion Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.937171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.937171Z digest=sha256:073fbe2bd84050bbe6fd07179732a19e6f4df7e1155532c1c276186f47317efe

Observation 8ea7967d-c5cf-43b9-b7b0-6c88071c69d8 · outbound

This paper cites an unresolved cited work.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.942224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.942224Z digest=sha256:760e93ed5b5424ea878529188af66be1ec81f15e45e59bea15b4233e571ddfa0

Observation 6b24d409-e72f-490a-9076-d062918f9a8a · outbound

This paper cites Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen- Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen- Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.950108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:18.947164Z digest=sha256:864c280bdc93680f72d1061990476018423fb5bf363a5d74aa424d4a26d9d9e3

Observation bbf0d752-7c39-4230-a603-57489fc549be · outbound

This paper cites Animate Anyone: Consistent and controllable image-to-video synthesis for character animation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Animate Anyone: Consistent and controllable image-to-video synthesis for character animation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.926277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:18.952983Z digest=sha256:76774c3d1078593eca6ae1d5f2fc1367f361ca4e23e1faafcccbe6f2f0091f6c

Observation 96b74d71-11fa-477e-a9f7-f1c27bdfb051 · outbound

This paper cites DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.958454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.958454Z digest=sha256:d122117f4c57bc4f8cfa4c0c7581c3907836208ee1d5e29601d25bbc8accb474

Observation 44d74678-ed12-4eed-8030-c8f3ec50453d · outbound

This paper cites Make It Move: Controllable image-to-video generation with text descriptions.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Make It Move: Controllable image-to-video generation with text descriptions

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.903602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:18.964136Z digest=sha256:ceb196c880703a2c124886f3e0a99a3ba644e3be18854769aac98e9d1c1de97b

Observation bdb0589d-a0ee-46a7-8df8-03a8525b262b · outbound

This paper cites VideoControlNet: A Motion-Guided Video-to-Video Translation Framework by Using Diffusion Model with ControlNet.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models VideoControlNet: A Motion-Guided Video-to-Video Translation Framework by Using Diffusion Model with ControlNet

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:18.968765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:18.968765Z digest=sha256:d32869cc1daffd872df047887f4e029c21a03396d90a5dfec42033910408c6ae

Observation 3d3076c7-ae2d-4cff-9041-5f95f4fa7b0f · outbound

This paper cites Filtered-CoPhy: Unsupervised learning of counterfactual physics in pixel space.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Filtered-CoPhy: Unsupervised learning of counterfactual physics in pixel space

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.886266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:18.976763Z digest=sha256:ddc63890f6ec9500644dab5aad1361877605b4e46e6e91debe9c8774ddb95670

Observation fe7cc2ee-02e9-4846-8a25-bcef32705135 · outbound

This paper cites Hand-object contact consistency reasoning for hu- man grasps generation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Hand-object contact consistency reasoning for hu- man grasps generation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.868005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:18.982281Z digest=sha256:611160f0800cbbc7017e7b0073ad85c1cc5672d5537741db5e429bca4188c98d

Observation e06730cc-7239-4b65-ab90-20cbba4a5840 · outbound

This paper cites Co- Tracker3: Simpler and better point tracking by pseudo- labelling real videos.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Co- Tracker3: Simpler and better point tracking by pseudo- labelling real videos

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.841173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:18.988304Z digest=sha256:467711adae224a93cf8a243a4c8cc34a75712f979a56fc70fb500311ec625f05

Observation 764a0178-f37f-42a0-b9d9-e3ef2e8d58ef · outbound

This paper cites DreamPose: Fashion image-to-video synthesis via stable diffusion.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models DreamPose: Fashion image-to-video synthesis via stable diffusion

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.818208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:18.993045Z digest=sha256:466221d0fca2c672cbf5df72446f7dd232c6e219a0a7deea5a413328aa3572d3

Observation f1d837d9-8b02-4b98-a017-f9e7c7b23b60 · outbound

This paper cites Elucidating the design space of diffusion-based generative models.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Elucidating the design space of diffusion-based generative models

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.798220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:18.998702Z digest=sha256:3f3eb72775fc32c656c480e8603defd7a87b8216f565f8ff6a4250a44671297f

Observation 253e989d-3adf-4900-a950-42d762e79977 · outbound

This paper cites Be- yond the Contact: Discovering comprehensive affordance for 3D objects from pre-trained 2D diffusion models.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Be- yond the Contact: Discovering comprehensive affordance for 3D objects from pre-trained 2D diffusion models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.782064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.004124Z digest=sha256:cb63efe668f94f8a8b844ac1a8939cc52a8e2d8aaf1ed99da228d147ee5117d0

Observation abaf9553-e632-4a5a-9c6d-3c027539aa69 · outbound

This paper cites Kingma and Jimmy Ba.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Kingma and Jimmy Ba

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.764816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.009188Z digest=sha256:8afc78a62d6d934fdd43479e1b0d4ee1b8b2d7e4947313db88c26fde1e699321

Observation db6e7b5f-2b48-463e-a010-34dac2601dd2 · outbound

This paper cites Neural relational inference for interacting systems.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Neural relational inference for interacting systems

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.747522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.014314Z digest=sha256:ea06ccc7159ba311c2ea9d80aa4677915fe900063721b22a627abea487a01108

Observation efc507ae-532c-4fd0-ad06-f1b19ffa28d5 · outbound

This paper cites Efros, and Krishna Kumar Singh.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Efros, and Krishna Kumar Singh

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.728118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.019506Z digest=sha256:87f0c3cfcc85b243b2b607eb4cff28c87671386d1d886f774464aaa909b7ce1b

Observation f8937270-a63e-494a-8ef8-3cd87d89f132 · outbound

This paper cites Learning phys- ical intuition of block towers by example.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Learning phys- ical intuition of block towers by example

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.710475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.024891Z digest=sha256:92fdc27299d0696e37db15a4ab0549679b7170917157c42cb95e68372c43e663

Observation 561ab2a0-e60b-43f4-9f4b-2a306264d261 · outbound

This paper cites Causal discovery in physical sys- tems from videos.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Causal discovery in physical sys- tems from videos

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.690701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.030684Z digest=sha256:97a702a4e8053dcbebec839b70e73da127ee4b7793c7471eed6156fc8e0d6b03

Observation 7d8036c1-990e-4d6a-97dd-a4d494a7d4e2 · outbound

This paper cites Freeman, Fr ´edo Du- rand, and Edward H.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Freeman, Fr ´edo Du- rand, and Edward H

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.667674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.036655Z digest=sha256:8c3ff6d72e00da391d32f77a6cd349c3540437f052b157a6293e197ac4757e2a

Observation 935f0295-c68a-4ea2-b151-f5a8d3443045 · outbound

This paper cites Mind’s Eye: Grounded language model reasoning through simulation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Mind’s Eye: Grounded language model reasoning through simulation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.649870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.042360Z digest=sha256:9d858bd3fff3350ff5bcda1866b6bec8294977555971b3279714e9dc2d546b58

Observation 9547195f-a1df-44e8-8fb1-01f2d87d2c28 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:19.050361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:19.050361Z digest=sha256:e071784cdfee47057e27461aae037823ad448ac2162e5760389a47fabd450998

Observation 1f289356-e20c-47e8-88fe-a38433831044 · outbound

This paper cites PhysGen: Rigid-body physics-grounded image-to-video generation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models PhysGen: Rigid-body physics-grounded image-to-video generation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.628310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.056069Z digest=sha256:7fca6f7ead7a5ebefa8f7b9fd1f5adcdf8d28482fcc75c3fd596d6803a037adc

Observation af7a849c-407e-4f7a-9c4e-1d73dacfc999 · outbound

This paper cites Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:19.067397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:19.067397Z digest=sha256:4424dd6b431c7ba2f8f6ed34ed7e5471b46865d33c234b2490b786cd37eaf03c

Observation 342d7559-8032-49c2-af18-dd2e720d297f · outbound

This paper cites Freeman, and Michael Rubinstein.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Freeman, and Michael Rubinstein

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.612293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.075934Z digest=sha256:15d423344eab81f1e893f104056fa1cf93081b656c81756143c1e6c8f21395cb

Observation 0aaf1c24-5671-4986-910b-e3eaf6389e09 · outbound

This paper cites Freeman, and Michael Rubinstein.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Freeman, and Michael Rubinstein

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.597050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.086578Z digest=sha256:bb9b60bf8dec486585b853137b26d014db8d60bf607ce9e1b138d211e3cbb379

Observation ab873ba0-6bfa-4d09-8c19-2de341986e1c · outbound

This paper cites UGG: Unified generative grasping.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models UGG: Unified generative grasping

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.580663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.092099Z digest=sha256:665ea441b0be9cad29a9b535b9f35666abb762a5726e8cb901750dad3a9636c9

Observation d98614a6-ed06-4376-8cc9-7ba0bf72c0b8 · outbound

This paper cites Bastiaan Kleijn.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Bastiaan Kleijn

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.560844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.101708Z digest=sha256:ef27abaede4d1c099cc2e9c56b67d7126f9796665a8c2189203e5abd31a42b59

Observation 711ab03c-fa7a-49d4-97fe-c773590a7d7f · outbound

This paper cites Something-Else: Compositional action recognition with spatial-temporal in- teraction networks.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Something-Else: Compositional action recognition with spatial-temporal in- teraction networks

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.544113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.108896Z digest=sha256:f57c0ac710c9af962a11f4e215d23ebf34bf928d42076cde737bf3a2572bb767

Observation 37ab1a3b-2c84-4ef8-b1a8-dc4c6df34a50 · outbound

This paper cites T2I-Adapter: Learn- ing adapters to dig out more controllable ability for text-to- image diffusion models.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models T2I-Adapter: Learn- ing adapters to dig out more controllable ability for text-to- image diffusion models

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.525883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.114338Z digest=sha256:5bd0b6586257783c5c1a8e2f7e1cec30a4424d19171050422fcf7a7cda8d36b8

Observation 10325b9a-8b15-4ee9-9a54-3dfcea798c4a · outbound

This paper cites Otaduy, Dan Casas, and Christian Theobalt.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Otaduy, Dan Casas, and Christian Theobalt

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.507223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.119244Z digest=sha256:13fc2c77a8f35176339f22325943f41ad8a8fc17ba969273b6073c88a01be2e9

Observation 2393e460-4787-47c0-b147-e9a415451454 · outbound

This paper cites Guibas, and Jimei Yang.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Guibas, and Jimei Yang

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.488190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.124427Z digest=sha256:78735e9d29701977049662998819b8dd7d44e99c27f72f2b83bf82bc8e41658d

Observation e18b4c74-3be2-494f-b7b0-883a62d9add7 · outbound

This paper cites 3D whole-body grasp synthesis with directional controllability.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models 3D whole-body grasp synthesis with directional controllability

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.470777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.130502Z digest=sha256:14bb570cc9233cd0ed9d9ae7c7781404a049ebbb6c367229a9c1476444b9b3d3

Observation df8117a5-c371-4031-8918-b929e9424d3b · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Learn- ing transferable visual models from natural language super- vision

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.449844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.137055Z digest=sha256:29fd36ec59687deb26f0e0aceb847805a5d73c106372055be852b2a68483929e

Observation 7310b31b-dd2d-4ffd-90b3-2fb47514e77a · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models SAM 2: Segment Anything in Images and Videos

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:19.143183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:19.143183Z digest=sha256:3610a7f85b14ea6a4a3c7b06181144d3c0d05a1053363e3e3d50b2bd0ce9eb43

Observation e7d496c0-634b-48c4-a049-ea2665979bcc · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models High-resolution image synthesis with latent diffusion models

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.429882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.149450Z digest=sha256:92722a091b7e75f8173782653717a7cad0deb1bd2340ac2d0ef2af8dc9a7a147

Observation 047ca493-fe0e-4c71-9830-a314cca7e823 · outbound

This paper cites an unresolved cited work.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:39:20.405702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.155008Z digest=sha256:9a2d3841764326cc6da5c4baf075630b420da5b49f143d1cf395db6ec19e34bc

Observation a9c038b3-427d-4c2e-bed5-2ec043b6701e · outbound

This paper cites Make-A-Video: Text-to-video genera- tion without text-video data.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Make-A-Video: Text-to-video genera- tion without text-video data

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.388946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.160871Z digest=sha256:cb11a0d784b86f01e5b9e443c8fa08733705337ea23b10d59fcbb72591fd0830

Observation cdf1ded8-0183-42b9-a8a5-6f08150382ef · outbound

This paper cites StyleGAN-V: A continuous video generator with the price, image quality and perks of StyleGAN2.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models StyleGAN-V: A continuous video generator with the price, image quality and perks of StyleGAN2

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.371605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.167841Z digest=sha256:967dec7659c7bb3b8f18d4bd86669826d4ba91af0fc1ad25d2e9425cc82b7b47

Observation 2ba07e1c-15f8-4346-88f2-3e60249b99a8 · outbound

This paper cites Controlling the world by sleight of hand.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Controlling the world by sleight of hand

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.351068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.173921Z digest=sha256:c7b931430d877f5fe6810e4518d61c04510530a9f2ee3d15e96160a475c3e2fa

Observation c6fa3ae1-f6e7-4c6a-ad43-e2c2ad660bcb · outbound

This paper cites Black, and Dim- itrios Tzionas.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Black, and Dim- itrios Tzionas

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.331756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.179124Z digest=sha256:f57c18a2261a4f06e6ff3499edd2d303ec1de169a2340516d999f87e0d8bac57

Observation 03b03836-fc8d-4899-8155-30cd7f9ecebf · outbound

This paper cites an unresolved cited work.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:39:20.315882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.183908Z digest=sha256:6d7f1994e89ee5b31cb60f093f1fa02c44f6b43b304d8316fd50ae85f349141b

Observation 2f1673af-1566-48a3-95cb-d9b74f1f3419 · outbound

This paper cites Collaborative learning for hand and object reconstruction with attention-guided graph convolu- tion.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Collaborative learning for hand and object reconstruction with attention-guided graph convolu- tion

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.299813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.188417Z digest=sha256:e6eda1c4911b7e5dba347c33060890b6de9f9f05a3e95d6b0f79b2e7af3f75ec

Observation 98a86d9d-79fa-41ff-96eb-e808f8e6e741 · outbound

This paper cites FVD: A new metric for video generation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models FVD: A new metric for video generation

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.283005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.193170Z digest=sha256:b6d082161579b79b9fc0f00e6183861b88ac45603a7e4bb0fef1c7805a5c9c22

Observation 197d1c15-cf7f-472c-b230-b658a3608c9b · outbound

This paper cites SV3D: Novel multi-view synthesis and 3D generation from a single image using la- tent video diffusion.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models SV3D: Novel multi-view synthesis and 3D generation from a single image using la- tent video diffusion

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.267415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.198079Z digest=sha256:c07bcae1f106271070b10865d1819e6b48b9f5c2cc039693c424a340dc69649c

Observation da03acfd-53ed-4d21-ab99-df06ad51ca69 · outbound

This paper cites ModelScope Text-to-Video Technical Report.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models ModelScope Text-to-Video Technical Report

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:19.202146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:19.202146Z digest=sha256:14ca424d1e4ab6179548b252ca326ed6e1183de09a4bc9fd1d399e0491da9606

Observation 8996d7f5-2d9e-4a27-8402-297b8575f3b9 · outbound

This paper cites Boximator: Gen- erating rich and controllable motions for video synthesis.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Boximator: Gen- erating rich and controllable motions for video synthesis

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.249497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.206971Z digest=sha256:7336dd54dee6d983f716c37ca64f4aa2a55c2495de4deb487016e32ddaaec8a3

Observation d162aa25-dcf0-4508-96c6-0072c589e49b · outbound

This paper cites VideoComposer: Compositional video syn- thesis with motion controllability.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models VideoComposer: Compositional video syn- thesis with motion controllability

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.232889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.211988Z digest=sha256:8713ac0171ee99cbb140c27bc8550f4a5bbaa012f4a33578ac0d4211fe0bc558

Observation 944308d2-5532-47d1-9aa6-1a83a0d13ee6 · outbound

This paper cites Bovik, Hamid R.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Bovik, Hamid R

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.214271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.217261Z digest=sha256:d394632a8cac1ab9867107dd5ea886c2b5f037a21eac33f5b657f7770bcea097

Observation 450ea379-997f-47dd-bb50-d535e7398b40 · outbound

This paper cites MotionCtrl: A unified and flexible motion controller for video generation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models MotionCtrl: A unified and flexible motion controller for video generation

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.198381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.222391Z digest=sha256:fc25ed39f560b042054df09a43246419b50be8f862130b65f2eceec6cf647b48

Observation b85c1936-0d8f-425a-92ec-8ace81d4ef69 · outbound

This paper cites Visual Interaction Networks: Learning a physics simulator from video.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Visual Interaction Networks: Learning a physics simulator from video

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.182241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.227163Z digest=sha256:7976360a9b2a297cfecafe99767488001e20daffa23d0b57dfe6a98fc11bc8e8

Observation f8613616-e1ba-41ef-bf7d-81918716570d · outbound

This paper cites Lim, Bill Freeman, and Joshua B.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Lim, Bill Freeman, and Joshua B

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.162123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.232086Z digest=sha256:b7c957c0dfbcaf6b93077ae3d2504a3b4626d72fcaa2fc9d635f8957eef764bb

Observation cfc11c7f-cb4a-46be-8249-27510d7d5e9e · outbound

This paper cites Lim, Hongyi Zhang, Joshua B.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Lim, Hongyi Zhang, Joshua B

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.127224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.239000Z digest=sha256:5b5c62561ff6948841e624eca93f02933cbff2e2d6c63385e6b5ad3d1ec1c8d5

Observation 1f539313-5d33-4eb1-abe8-b76cf9be9953 · outbound

This paper cites Learning to see physics via visual de- animation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Learning to see physics via visual de- animation

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.100046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.244442Z digest=sha256:78cbfce701486cd79bca54a01d13c7db232d6ae1f9b40d461febd8ec37b5374b

Observation ca462425-f46b-40b9-bd93-79911c87fbe5 · outbound

This paper cites Tune-A-Video: One-shot tun- ing of image diffusion models for text-to-video generation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Tune-A-Video: One-shot tun- ing of image diffusion models for text-to-video generation

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.083562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.248950Z digest=sha256:2b4074350587792b1ba0df7cef408a718b0b94d10222d9a1707dc0b401ad7680

Observation 781c6eaf-2aee-4131-aebf-f782facdcad2 · outbound

This paper cites THOR: Text to Human-Object Interaction Diffusion via Relation Intervention.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models THOR: Text to Human-Object Interaction Diffusion via Relation Intervention

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:19.253823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:19.253823Z digest=sha256:3e2588c99c0346f0277d5e2c5fd934b36c3fc836648310c3a52bb3ac5360ccb9

Observation 341ea06d-76f9-4ff6-8116-0eb1bc42d6ba · outbound

This paper cites DragAnything: Motion control for any- thing using entity representation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models DragAnything: Motion control for any- thing using entity representation

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.067753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.260126Z digest=sha256:cfe839d10d6fe59476579a47c923074fe877e6b1c66235c34a81cd8074010224

Observation cca6ec1b-f5f4-425b-bfb8-19e75666b174 · outbound

This paper cites Template free reconstruction of human- object interaction with procedural interaction generation.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Template free reconstruction of human- object interaction with procedural interaction generation

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.051109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.266404Z digest=sha256:cf6a4604fbeabb4c3b8b436423133aa76707a5d13014921558c17cd07bd0af6b

Observation 4e8decf4-9717-4a8c-b796-9cee67f6d0db · outbound

This paper cites DynamiCrafter: Animating open-domain images with video diffusion priors.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models DynamiCrafter: Animating open-domain images with video diffusion priors

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.030255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.272074Z digest=sha256:84942a4d23a190f78f133fa367b6c2bfa3579eb61efe4165746d32e11d5173aa

Observation 092f572f-9468-42a4-9cd3-2ec8b939b767 · outbound

This paper cites InterDiff: Generating 3D human-object interactions with physics-informed diffusion.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models InterDiff: Generating 3D human-object interactions with physics-informed diffusion

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:20.006994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.277607Z digest=sha256:5823e28eb4d9300db98e677d6396e8f679393a0f2c28aa6ada6afbfdb61aad0c

Observation 6dd22cdd-bed9-4210-bf5b-7d9f03253e44 · outbound

This paper cites MagicAnimate: Temporally consistent human image animation using diffusion model.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models MagicAnimate: Temporally consistent human image animation using diffusion model

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:19.988374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.283281Z digest=sha256:33cac11b1cc2eff505642a55897cd2085b2774f9e57f1a700475384a9d1fe291

Observation 1484aae8-d68b-44bf-b9b6-15d84a728146 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:19.288475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:19.288475Z digest=sha256:2232cd16b03fc5a96b8bf6ecb11ad9baf258d31e20ad4f5f469d93c038836a43

Observation 206b6445-4046-4f92-916e-1d46bb631eb0 · outbound

This paper cites Space-time diffusion features for zero-shot text-driven motion transfer.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Space-time diffusion features for zero-shot text-driven motion transfer

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:19.970417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.293351Z digest=sha256:9cb0544c90a560152bd418366ce9f345ebb1a41ae93b9c1da8a4eb61154f39a1

Observation 301d1a07-b9f8-42d7-a76b-f855d00d6b3c · outbound

This paper cites Diffusion-guided reconstruction of everyday hand-object interaction clips.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Diffusion-guided reconstruction of everyday hand-object interaction clips

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:19.949336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.298361Z digest=sha256:8ad9cced85fde0951e15b4de7cbacb9fd958303d8817e418f98cf27c6cd5629a

Observation 5439ebbb-ddb0-493e-9a31-18e7049545df · outbound

This paper cites Affordance Diffusion: Synthesizing hand-object in- teractions.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Affordance Diffusion: Synthesizing hand-object in- teractions

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:19.929245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.304187Z digest=sha256:4bede97bda5d275ff89d9a3d5583d88b33c85438cd760c62cb26569f9bdb252d

Observation 5b086ef3-67e6-4f29-9f77-c2d581004990 · outbound

This paper cites G-HOP: Generative hand-object prior for interaction reconstruction and grasp synthesis.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models G-HOP: Generative hand-object prior for interaction reconstruction and grasp synthesis

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:19.906283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.312537Z digest=sha256:5616eaceab79097d5a53cdf0012775d32c485f9050a750b9dd06e343befbf50a

Observation 49deb138-56ad-414f-89a9-021ae3d1c25e · outbound

This paper cites Tenenbaum.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Tenenbaum

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:19.886324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.321729Z digest=sha256:24451d6f07e8207001891210a4d972a81c742a777661e0f8ca555b85fd0a9fef

Observation 50f23fd7-7d6e-4e95-bb9c-9aadeb36d66d · outbound

This paper cites DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-11T14:39:19.327485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:39:19.327485Z digest=sha256:6fdb4634230421255cd6c502b47eddfd27420538a6ce2f3fdfca81ab8e4b7841

Observation f9e33b66-f84b-4c43-8dc7-c0501edef737 · outbound

This paper cites Video probabilistic diffusion models in projected latent space.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Video probabilistic diffusion models in projected latent space

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:19.869934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.332886Z digest=sha256:b46871580f41b534771d443d6b73d0b0bafb03ca2b4a1474470f871be5ebef54

Observation 6c73c87f-0c2f-4b75-b654-0a74ca417bb2 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

InterDyn: Controllable Interactive Dynamics with Video Diffusion Models Adding conditional control to text-to-image diffusion models

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:39:19.852193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T14:39:19.337878Z digest=sha256:69d305f7b05b7a0c4536dca58185b5eae06792ed0a79fcbb7455d656f387dedc

Pith citing papers

Observation b806d032-a812-4621-9c4f-104a04d5cdde · inbound

Generative Physical AI in Vision: A Survey cites this paper.

Generative Physical AI in Vision: A Survey InterDyn: Controllable Interactive Dynamics with Video Diffusion Models

Reference 252

Resolution
verified exact
local_arxiv, observed 2026-08-10T18:53:01.116142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T18:53:00.987266Z digest=sha256:39d9634a33b934c97b3a69f8a2b27568093d2a3ed4c67cb8828cc468931a83e2

Observation 9a4cb098-bed9-4131-b721-3a24e55fcf06 · inbound

Precise Action-to-Video Generation Through Visual Action Prompts cites this paper.

Precise Action-to-Video Generation Through Visual Action Prompts InterDyn: Controllable Interactive Dynamics with Video Diffusion Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T19:14:51.841623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:14:51.841623Z digest=sha256:c1bca2f32f85d9d00705b12e3ce312eb862c45c175b8632bdf00c0696b35027c