Pith. sign in

Paper Citation Record · LEDGER

AnyI2V: Animating Any Conditional Image with Motion Control

As of 8 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 1 inbound Pith citation observation for arXiv:2507.02857.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02857 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:25:29.212390Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-03T16:19:51.348169Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:28:38.518422Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact3
  • verified fuzzy25
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1fed75c3-b20c-4f6b-96ac-78b1a717cd86 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

AnyI2V: Animating Any Conditional Image with Motion Control Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:24.234995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:24.234995Z digest=sha256:4dbe961801db3181dce49f981f9f77d878b388e7a32c22ac21f9e617608a7bc9

Observation 88234fc6-27ef-4f33-83cd-dc35b5f64f2a · outbound

This paper cites Align your latents: High-resolution video synthesis with latent diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control Align your latents: High-resolution video synthesis with latent diffusion models

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:34.460141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:24.323877Z digest=sha256:a7257dea5570f3dd9df58ada4f0495892ac6cd10190ace1d228c57db5d6f7ea0

Observation efdf9c10-8737-47a9-a29a-4edda0d43b45 · outbound

This paper cites A unified 3d human motion synthesis model via conditional variational auto-encoder.

AnyI2V: Animating Any Conditional Image with Motion Control A unified 3d human motion synthesis model via conditional variational auto-encoder

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:34.328860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:24.413431Z digest=sha256:de0d1d61309aec897c93d620e21230015da229cd3ae6302ee71411153023878b

Observation 47fb6f8e-8c13-427a-8c2c-c82e2e41c664 · outbound

This paper cites VideoCrafter1: Open Diffusion Models for High-Quality Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:24.503120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:24.503120Z digest=sha256:6626dc604250d13efd0fd710321506ec12cb06ace25c28808d13a2c8e1f465bc

Observation 537051c3-c895-478e-8dab-e699d1e73474 · outbound

This paper cites Videocrafter2: Overcoming data limitations for high-quality video diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control Videocrafter2: Overcoming data limitations for high-quality video diffusion models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:34.189521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:24.594083Z digest=sha256:96c0f30cea2418396b6c3f34f6be9f7c2a42dfce1f4a06829751ea67298ba902

Observation 78f7342e-fe75-436f-8721-6dfd29aad377 · outbound

This paper cites MeViS: A large-scale benchmark for video segmentation with motion expressions.

AnyI2V: Animating Any Conditional Image with Motion Control MeViS: A large-scale benchmark for video segmentation with motion expressions

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.997665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:24.679423Z digest=sha256:23e1c43b26c3395bbd90c5b572ec85f0334255e8dc3420e1cb2d583863b28a4a

Observation 8991ac9c-0e5e-49cd-86b0-9fe1b6ca30f8 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

AnyI2V: Animating Any Conditional Image with Motion Control AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:24.769350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:24.769350Z digest=sha256:da6e48d059716005df4d1980e77926d424c6330e895b871f11305b5358b08a1c

Observation 35b70089-05a0-4fa1-8c2b-22c36faedc0e · outbound

This paper cites Sparsectrl: Adding sparse controls to text-to-video diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control Sparsectrl: Adding sparse controls to text-to-video diffusion models

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.854495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:24.884766Z digest=sha256:8ddd0bbf8faebe122c0e15951da3e887d891f36fe68758bf7b81e0246a92ecb8

Observation 84520573-1143-4c07-b930-d1d95e56baa1 · outbound

This paper cites CameraCtrl: Enabling Camera Control for Text-to-Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control CameraCtrl: Enabling Camera Control for Text-to-Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:24.979478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:24.979478Z digest=sha256:42867f67fc57137ed37611d96481a43e591deb36ea02a4c2f07137c35767c02d

Observation 00467e71-d160-4147-9adc-40188cf446b9 · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:25.057827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:25.057827Z digest=sha256:52161007af9c1eda5e31d07e8d934ff136a0622304102c1a9598d76c01bb1a8a

Observation c424eb1b-8aa8-43b4-960f-411cbbdaba4e · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

AnyI2V: Animating Any Conditional Image with Motion Control Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:25.140993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:25.140993Z digest=sha256:17ca1d807c0c103b2848f4be366c3840d59ea0ef7be1fabc041329524894d98f

Observation 4a3287ac-ad55-4777-a748-4f898526342c · outbound

This paper cites Video diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control Video diffusion models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.684728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:25.240688Z digest=sha256:b1a644083b95dddfcbca00cfb893d0acc69e81b78b84a635e30d4b3d3310dd89

Observation c3b78684-d66f-476e-bafb-695ab7fa9ace · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

AnyI2V: Animating Any Conditional Image with Motion Control LoRA: Low-Rank Adaptation of Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:25.318105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:25.318105Z digest=sha256:8102f14a3bd0bc4e64e33826b41eea555371a0105698d27ff02b25100587eb36

Observation 9fddd0ee-b922-46d1-9b73-51842255c65b · outbound

This paper cites Cocktail: Mixing multi-modality control for text-conditional image generation.

AnyI2V: Animating Any Conditional Image with Motion Control Cocktail: Mixing multi-modality control for text-conditional image generation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.505365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:25.380684Z digest=sha256:c23a58cd20723dd9fbfbbb97e7b570d69e99513733a8134ae23dfd92ea9900f6

Observation 475fe6c3-411c-4d7c-be8a-a8000066e717 · outbound

This paper cites VideoControlNet: A Motion-Guided Video-to-Video Translation Framework by Using Diffusion Model with ControlNet.

AnyI2V: Animating Any Conditional Image with Motion Control VideoControlNet: A Motion-Guided Video-to-Video Translation Framework by Using Diffusion Model with ControlNet

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:25.466570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:25.466570Z digest=sha256:f717a90ac12a25ad0de77413568899ce12947d526163d23e4cc4e82288a73fbf

Observation bc66d003-8b0c-4769-bb91-666dc3eb6548 · outbound

This paper cites Arbitrary style transfer in real-time with adaptive instance normalization.

AnyI2V: Animating Any Conditional Image with Motion Control Arbitrary style transfer in real-time with adaptive instance normalization

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.367870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:25.558706Z digest=sha256:0a9733d8cbab6634c3ef0939040650fb5c92718179b36b7b5f2d760418f49cdd

Observation d19d4810-6fc5-4c9f-a904-61b8c942574f · outbound

This paper cites Cotracker: It is better to track together.

AnyI2V: Animating Any Conditional Image with Motion Control Cotracker: It is better to track together

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.236726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:25.679527Z digest=sha256:6f71eff54e6c8a7b24f987e8386a39e787ddb094097b17f2668170380cc4ab4a

Observation 79574eae-be01-4584-ad52-ea507b4c5876 · outbound

This paper cites Text2video-zero: Text- to-image diffusion models are zero-shot video generators.

AnyI2V: Animating Any Conditional Image with Motion Control Text2video-zero: Text- to-image diffusion models are zero-shot video generators

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:25.815397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:25.815397Z digest=sha256:f526dcd1e2f244a12c943fb79ec47a792944335d8d37f552dbc179c56dd59100

Observation b8490783-78ba-40ff-bd46-73591113df79 · outbound

This paper cites DiffBlender: Composable and Versatile Multimodal Text-to-Image Diffusion Models.

AnyI2V: Animating Any Conditional Image with Motion Control DiffBlender: Composable and Versatile Multimodal Text-to-Image Diffusion Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:25.914025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:25.914025Z digest=sha256:a78aa2dfb50ffd1882f977f662587907f2045858e4ab1e29de04e1bce2334b35

Observation f0a79e52-683e-471f-834b-a8b00b4bf615 · outbound

This paper cites Compose and Conquer: Diffusion-Based 3D Depth Aware Composable Image Synthesis.

AnyI2V: Animating Any Conditional Image with Motion Control Compose and Conquer: Diffusion-Based 3D Depth Aware Composable Image Synthesis

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:25:29.851895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:26.042201Z digest=sha256:faf8cdea1759cc1675f8dc2d66843380849abc4f698f5a7683b3b715ad47592c

Observation d3b52798-77ba-4689-a296-1a3266b80d7b · outbound

This paper cites Image Conductor: Precision Control for Interactive Video Synthesis.

AnyI2V: Animating Any Conditional Image with Motion Control Image Conductor: Precision Control for Interactive Video Synthesis

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:26.170396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:26.170396Z digest=sha256:54d4de86232f4413aa455f4236beb0e577e70de9476af311a08e9e87c0774533

Observation 5feeba51-a878-426d-afe4-e1017b07d665 · outbound

This paper cites LOVECon: Text-driven Training-Free Long Video Editing with ControlNet.

AnyI2V: Animating Any Conditional Image with Motion Control LOVECon: Text-driven Training-Free Long Video Editing with ControlNet

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:25:29.716487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:26.239158Z digest=sha256:32f61ccda337ad557a77936c456266704d436317fdc4532e6e33220e2aa2f0a5

Observation 00524d34-72b6-4ca3-a573-6b5d1407a5dc · outbound

This paper cites Trailblazer: Trajectory control for diffusion-based video generation.

AnyI2V: Animating Any Conditional Image with Motion Control Trailblazer: Trajectory control for diffusion-based video generation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:33.012633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:26.320914Z digest=sha256:73d35e0ce431cde076beeb7f6eee296bc16c1227232dc5d18d8c8e38eb589381

Observation e796bd06-2fb9-498a-8b85-e3ccb432bbb7 · outbound

This paper cites Some methods for classification and analysis of multivariate observations.

AnyI2V: Animating Any Conditional Image with Motion Control Some methods for classification and analysis of multivariate observations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:32.881148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:26.403459Z digest=sha256:1741245bc49f5b36df391e3fe73bf01d4600097679c64b176fd51a7103858840

Observation 661d2f45-91d9-4fc7-9bda-a2cced54d8c5 · outbound

This paper cites Large-scale video panoptic segmentation in the wild: A benchmark.

AnyI2V: Animating Any Conditional Image with Motion Control Large-scale video panoptic segmentation in the wild: A benchmark

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:32.735514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:26.474151Z digest=sha256:9830c814e0aeeddf67555a4bdcc5fc4ff8bc954c730605307f82dc608a8d2617

Observation d73ee21d-b3cd-43af-97c1-0bea7ea1e6ca · outbound

This paper cites Freecontrol: Training-free spatial control of any text-to-image diffusion model with any condition.

AnyI2V: Animating Any Conditional Image with Motion Control Freecontrol: Training-free spatial control of any text-to-image diffusion model with any condition

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:32.562331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:26.553258Z digest=sha256:388f8c6419ef551179d109f7000e9af1b1ff4b2cfc04a93af36eb8e5d1c2b677

Observation 8526d6aa-bf06-4f6b-85a0-d2c9887d5ca2 · outbound

This paper cites SG-I2V: Self-Guided Trajectory Control in Image-to-Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control SG-I2V: Self-Guided Trajectory Control in Image-to-Video Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:26.611739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:26.611739Z digest=sha256:01f4ab699463c6bef5d94f477e36f4c7683aee1e016f2632a40069241296ccb7

Observation 0afba912-7f85-4603-bfd0-d1ec67db126d · outbound

This paper cites Mofa-video: Controllable image animation via generative motion field adaptions in frozen image-to-video diffusion model.

AnyI2V: Animating Any Conditional Image with Motion Control Mofa-video: Controllable image animation via generative motion field adaptions in frozen image-to-video diffusion model

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:32.433431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:26.685839Z digest=sha256:7d2b123d6abedb415c70f5a1a9dc2cea4613e904c8c0c0952ed7a1786059464c

Observation 92e86803-561b-43a2-b383-9cdfb26ad7d8 · outbound

This paper cites Drag your gan: Interactive point-based manipulation on the generative image manifold.

AnyI2V: Animating Any Conditional Image with Motion Control Drag your gan: Interactive point-based manipulation on the generative image manifold

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:32.218979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:26.760190Z digest=sha256:3ee3c361ae4a7a815ea20b5c303de0ecf3162038fe5e6b175f8715440b37dc20

Observation 38b44cb9-bcc3-40a5-8436-4c7176e5d212 · outbound

This paper cites UniControl: A Unified Diffusion Model for Controllable Visual Generation In the Wild.

AnyI2V: Animating Any Conditional Image with Motion Control UniControl: A Unified Diffusion Model for Controllable Visual Generation In the Wild

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:26.818684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:26.818684Z digest=sha256:5c044085cbca5b7822307e4cc7bf7180fbafc1cd155154f76d9d8aeb1c88437a

Observation 42402254-5ab4-4997-be70-b199df2bf15c · outbound

This paper cites FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models.

AnyI2V: Animating Any Conditional Image with Motion Control FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:26.881733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:26.881733Z digest=sha256:669c7a10ffb1b4659061cc790c0db30df7918bf176964bbc71ce002230b6f953

Observation baed88d4-efb6-4686-a27d-15a66c5355a1 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control High-resolution image synthesis with latent diffusion models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:26.900124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:26.900124Z digest=sha256:186597d325a4fb01ef05172e098885479f5a1b31431e95e7133dd60cf4286bdf

Observation fd175d70-34e2-448a-b6f4-b2577bf1b0ce · outbound

This paper cites Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling.

AnyI2V: Animating Any Conditional Image with Motion Control Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:32.068727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:26.970264Z digest=sha256:2d0b99e6ca5f3de81e69fec6cb43fa8b6de006f62ccd3417594eea827174373c

Observation 15eb4f40-dd58-403c-b2c4-ed011779c61b · outbound

This paper cites Dragdiffusion: Harnessing diffusion models for interactive point-based image editing.

AnyI2V: Animating Any Conditional Image with Motion Control Dragdiffusion: Harnessing diffusion models for interactive point-based image editing

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:31.963843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:27.078636Z digest=sha256:e43f2736ce9c1e13e1cc311ab7223b8d4e7aa9f333712f9535fa2f11dcf30a54

Observation 40b743f7-bc70-444f-8b57-053c8fc6fc83 · outbound

This paper cites A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models.

AnyI2V: Animating Any Conditional Image with Motion Control A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:27.267209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:27.267209Z digest=sha256:8bcd60a6b333f9a5b38f47bcc1ccb6aff844db4fd64e8d2d9e7a442742fd5ece

Observation 32df7593-0501-4d21-bc58-60be5e71da60 · outbound

This paper cites Free-form motion control: A synthetic video generation dataset with controllable camera and object motions.

AnyI2V: Animating Any Conditional Image with Motion Control Free-form motion control: A synthetic video generation dataset with controllable camera and object motions

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:27.408167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:27.408167Z digest=sha256:6a672d5dec77c8dd3e058f1c4d8d69f1b0bead099b82b4b72fb97c1a7052c0f3

Observation d3201c98-aac7-4366-8a95-f8f259a0d87d · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

AnyI2V: Animating Any Conditional Image with Motion Control Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:27.446167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:27.446167Z digest=sha256:b8bddeb06c7cff870876d42968ad8c2160a5cd1a638dc20940ea87ca34bee288

Observation b5292019-0b83-4721-be2b-ae7ff2d5fb28 · outbound

This paper cites Denoising Diffusion Implicit Models.

AnyI2V: Animating Any Conditional Image with Motion Control Denoising Diffusion Implicit Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:27.533479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:27.533479Z digest=sha256:d050be3b756c2e158dea93471ed07249fff8959f4e06da6fbc5aafe793472b94

Observation f6a8d720-8439-483d-a44e-83d33bed5481 · outbound

This paper cites Anycontrol: create your artwork with versatile control on text-to-image generation.

AnyI2V: Animating Any Conditional Image with Motion Control Anycontrol: create your artwork with versatile control on text-to-image generation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:31.798937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:27.623943Z digest=sha256:8efc0b0ea50fe3ea6e7c77179d27225709d48b7e7535b90dea56c76a14ff95f0

Observation 27c0b93b-146f-400e-91c5-0aa28818b3fd · outbound

This paper cites Plug-and-play diffusion features for text-driven image-to-image translation.

AnyI2V: Animating Any Conditional Image with Motion Control Plug-and-play diffusion features for text-driven image-to-image translation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:31.654118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:27.774197Z digest=sha256:a72a341a3ca1a18445668c091c6a3b0cd3f8fbbf09f13e94926e687ce5f94bb3

Observation d3771b2c-ec8f-4c5b-af71-83e19304621e · outbound

This paper cites ModelScope Text-to-Video Technical Report.

AnyI2V: Animating Any Conditional Image with Motion Control ModelScope Text-to-Video Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:27.903864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:27.903864Z digest=sha256:95f196652828f5baac660f7c6dfc92f0ff0fccb410048b48136ba8aea725eb32

Observation ef2b81e3-e487-4b92-a6da-fb15b3fc5765 · outbound

This paper cites Boximator: Generating Rich and Controllable Motions for Video Synthesis.

AnyI2V: Animating Any Conditional Image with Motion Control Boximator: Generating Rich and Controllable Motions for Video Synthesis

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.032439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.032439Z digest=sha256:c61234fc40c3e5b4bcefae718c98179906997ba89276c7afc96cd1abc610f2a7

Observation d856d6b6-8afc-4eee-a71d-a74b60ec6519 · outbound

This paper cites Videocomposer: Compositional video synthesis with motion controllability.

AnyI2V: Animating Any Conditional Image with Motion Control Videocomposer: Compositional video synthesis with motion controllability

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:31.526763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:28.197918Z digest=sha256:9bfbf07bcd1a446f27dd67b5c448be988f7b4cf10513827eb3e7c9c92af20870

Observation 37898861-1e15-4fec-b544-e391ee84143d · outbound

This paper cites Lavie: High-quality video generation with cascaded latent diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control Lavie: High-quality video generation with cascaded latent diffusion models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:31.177524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:28.282240Z digest=sha256:bb5cf3e2766e214d486c0363faa30ce3a8a7950840c7dff059aee3f800d7a3b8

Observation c9840de4-56e8-4fd6-9acd-8ff050e23d5c · outbound

This paper cites ObjCtrl-2.5D: Training-free Object Control with Camera Poses.

AnyI2V: Animating Any Conditional Image with Motion Control ObjCtrl-2.5D: Training-free Object Control with Camera Poses

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.341243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.341243Z digest=sha256:5b33fecc4920fd32a8d0a67f9e858294b773af4a8424c9178d631f56dd6308e6

Observation 91b33beb-3f62-4ddb-969e-f0e49dd75f68 · outbound

This paper cites Motionctrl: A unified and flexible motion controller for video generation.

AnyI2V: Animating Any Conditional Image with Motion Control Motionctrl: A unified and flexible motion controller for video generation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:30.904545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:28.415133Z digest=sha256:91b4e498f155ca56f442f66215e2fc0c868c16a45c53fe95096c3f29e024e73a

Observation 396d7879-b606-43e1-a74c-6cba50ac1ae1 · outbound

This paper cites Principal component analysis.

AnyI2V: Animating Any Conditional Image with Motion Control Principal component analysis

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:30.631228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:28.525885Z digest=sha256:f96edbbbe2f7ef0f696e2d6b9087c82071b38983c86458ee1c4f03d2e1df9832

Observation 39f036ab-5588-464d-8392-59af76ab3bf2 · outbound

This paper cites MotionBooth: Motion-Aware Customized Text-to-Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control MotionBooth: Motion-Aware Customized Text-to-Video Generation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.608683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.608683Z digest=sha256:3938ba7ac4fd51baa628d34edaf6eacd259dac99732517b6a7c6dd172f962984

Observation 15b79da1-3c97-477f-b11c-507de6af68ec · outbound

This paper cites Draganything: Motion control for anything using entity representation.

AnyI2V: Animating Any Conditional Image with Motion Control Draganything: Motion control for anything using entity representation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:30.383936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:28.647601Z digest=sha256:dc367cf4924e7cb1e965b20fe950f6e35df8671361bbe4b22751191039abed1d

Observation d8787f87-8def-4463-8b24-f76278b73f0c · outbound

This paper cites Video Diffusion Models are Training-free Motion Interpreter and Controller.

AnyI2V: Animating Any Conditional Image with Motion Control Video Diffusion Models are Training-free Motion Interpreter and Controller

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.728755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.728755Z digest=sha256:773ffc4a4f63c94d6e1720ff3400fb80ae5a0215b07f5393ff2a11b3f3eb66a0

Observation 4ab67a9c-61ea-49c4-a657-7d1bc1f651b0 · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

AnyI2V: Animating Any Conditional Image with Motion Control Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.790311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.790311Z digest=sha256:ae089596c8af91d8c869aaf328df40151b6d8a745b71c96ef198f0deb2aead2d

Observation ec32a8f0-e334-4054-882a-570f449f538f · outbound

This paper cites Direct-a-video: Customized video generation with user- directed camera movement and object motion.

AnyI2V: Animating Any Conditional Image with Motion Control Direct-a-video: Customized video generation with user- directed camera movement and object motion

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:25:30.116321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:28.847157Z digest=sha256:3dd7bc5a6dbf7be2270951067fbc55091ac54230e4ca2a32fdb0b59d1ef0ac60

Observation 695822a8-cc50-409c-8cf6-d9736a99a273 · outbound

This paper cites DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory.

AnyI2V: Animating Any Conditional Image with Motion Control DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.946126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.946126Z digest=sha256:852ae2a7aba1facd8db1ab20c4c4e3d93e97977bb234fdf78df1492cdb6daf62

Observation 5b748225-0f62-4468-8f00-57e762518ba8 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

AnyI2V: Animating Any Conditional Image with Motion Control Adding conditional control to text-to-image diffusion models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:28.992068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:28.992068Z digest=sha256:aa98d13028f558e84be11f6c911061543a1ff41a36373cc1631515ab3928c594

Observation 1dc34cde-99a8-4f00-8daa-3face3f52316 · outbound

This paper cites ControlVideo: Training-free Controllable Text-to-Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:29.057301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:29.057301Z digest=sha256:9ee7369df41c5df8880ca93e1d84b77c8ba6f571fd2b17efa56df5cc8c7421e2

Observation 1351d6ee-1961-4baa-95d3-85db563283d7 · outbound

This paper cites Tora: Trajectory-oriented Diffusion Transformer for Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control Tora: Trajectory-oriented Diffusion Transformer for Video Generation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:29.127199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:29.127199Z digest=sha256:617b4d9e6ee6a4c623e9c25fef084327b1331edf8cb426f529f6ccaf99495931

Observation cdcdc99d-47e6-4ea1-8476-9548690d5220 · outbound

This paper cites TrackGo: A Flexible and Efficient Method for Controllable Video Generation.

AnyI2V: Animating Any Conditional Image with Motion Control TrackGo: A Flexible and Efficient Method for Controllable Video Generation

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:25:29.416173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T20:25:29.212390Z digest=sha256:2c38cb7597fd8bb566508029dee9c38f3e4e604781804615169bbb65a69f00e0

Pith citing papers

Observation bd2c33f3-ad0b-4654-8b84-00132bccf2f5 · inbound

QWERTY: Training-Free Motion Control via Query-Warped Video Diffusion Transformers cites this paper.

QWERTY: Training-Free Motion Control via Query-Warped Video Diffusion Transformers AnyI2V: Animating Any Conditional Image with Motion Control

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:28:38.520077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T16:19:51.348169Z digest=sha256:3518ff6bcfd38a89495527c0444fd0ae7e044343dd68541ab1a6ecb30775fa30