Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:02.074963Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 4 inbound Pith citation observations for arXiv:2505.19901.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:02.074963Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T17:45:17.010883Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T08:53:15.680626Z
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 72aa2235-ed7d-4a1e-b956-d344de8f3c81 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Qwen Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ad673e4-4a94-4317-9f21-cd09aa1de307 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Frozen in time: A joint video and image encoder for end-to- end retrieval
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5625801f-3387-4d65-abba-e3727c721a87 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Stable video diffusion: Scaling latent video diffusion models to large datasets, 2023
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83c440f4-422f-42d8-8c32-4838e8fda54f · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dcc775db-a1f0-4b45-ac3d-874a0371499f · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Panda-70m: Captioning 70m videos with multiple cross-modality teachers
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1d864c9a-0500-4747-9984-74679bbc0d47 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Seine: Short-to-long video diffusion model for generative transition and prediction, 2023
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4e6ab43e-a61d-4b1b-9c5e-3edd53e09905 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks, 2024
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0576af17-93b7-454c-953d-79fa927e7b7a · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Glm: General language model pretraining with autoregressive blank infilling, 2022
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f74e934f-cd88-4374-9ec0-3a4106fdb662 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Guiding instruction-based image editing via multimodal large language models, 2024
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ec9e167b-c6ef-45e7-8b56-bd58cd3585f3 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Seed-x: Multi- modal models with unified multi-granularity comprehension and generation, 2025
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cfa1c5df-408b-4b15-a2ef-dff08987f7bd · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Denoising diffu- sion probabilistic models, 2020
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e3d4581-83d7-41de-835c-79bcdd1d5839 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Smartedit: Exploring complex instruction-based image editing with multimodal large language models, 2023
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d4496099-00c6-4eea-bced-54d3b2ae21e0 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Smartedit: Exploring complex instruction-based image editing with multimodal large lan- guage models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fc616fa2-5fd4-437c-a51d-1e71083f5844 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM VBench: Com- prehensive benchmark suite for video generative models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3cf50900-c011-46d9-99cb-c6d648a3569a · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Vbench++: Comprehensive and versatile benchmark suite for video generative models, 2024
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b6cd7b14-1bcf-45bd-b352-c42ba2864528 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM T2vbench: Benchmarking temporal dynamics for text-to- video generation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1ed9d684-8fa4-4a97-ade1-cc1f88d81b87 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Miradata: A large-scale video dataset with long durations and structured captions, 2024
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation aa339d00-f0ab-4f39-a214-1bc1aa13278b · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Co- tracker: It is better to track together, 2024
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6492a1b0-062b-4f98-bf61-4d9ca71a7a3a · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Hunyuanvideo: A systematic framework for large video generative models, 2025
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f02c080d-1631-4442-978d-a4ceffc245af · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Animateanything: Consistent and control- lable animation for video generation, 2024
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9894e930-f8e9-4f01-a7a3-45612583e2a6 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Evaluation of Text-to-Video Generation Models: A Dynamics Perspective
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 118e7302-b84f-4d47-a6c4-71f943984499 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Video-llava: Learning united visual representa- tion by alignment before projection, 2024
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 68ec7cf1-dcec-4e83-89dd-d13b1165fbdd · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Stiv: Scalable text and image conditioned video generation, 2024
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7c4adeb1-41e7-413a-92d7-7852e8ce260b · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Evalcrafter: Benchmarking and evalu- ating large video generation models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7bbe86ff-38a8-40e1-b972-e31d7e7a6155 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Qwen2vl-flux: Unifying image and text guidance for controllable image generation, 2024
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1806a057-43f3-4377-9110-5ac706fe34dd · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Openvid-1m: A large-scale high-quality dataset for text-to- video generation, 2024
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 996c4bfb-196b-4917-96b6-78e9ee86cc08 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Kosmos-G: Generating Images in Context with Multimodal Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39c33e2c-7768-43f3-9af6-eb5146e9763d · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Kosmos-g: Generating images in context with multimodal large language models, 2024
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d56617e2-834d-40d7-a800-f12d40feb9f5 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Askell Amanda, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 00f5ed96-84e4-4fe4-b58a-e6aff330096b · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bfe830b5-088e-4fdb-91af-422c53f1b5e6 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM ViBe: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f3e9a63-1144-4afd-9dd2-759a833fec09 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Consisti2v: Enhancing visual consistency for image-to-video generation, 2024
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1a46b82e-7a5d-458e-a224-fada33098c07 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM T2v-compbench: A comprehen- sive benchmark for compositional text-to-video generation,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation de6b89a0-6ec9-46a8-9d6e-1f89c5f7747a · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Generative multi- modal models are in-context learners, 2024
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6ea87874-43f1-4d34-a47b-ecc0262efe57 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Mimir: Improving video diffusion models for precise text understanding, 2024
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 45f36936-4ee2-4d7f-8ce8-6f79c40c8365 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Mige: A unified framework for multimodal instruction-based image generation and editing,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 730d09a2-221d-4bc3-8125-82d5864d176a · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Llama: Open and efficient foundation language mod- els, 2023
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a6edc8cb-6c1f-44b8-8a93-bc17d5bf146f · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Qwen2-vl: Enhancing vision-language model’s perception of the world at any resolution, 2024
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0b20784-0f36-48a6-addc-bdc49c96f5f1 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Koala-36m: A large-scale video dataset improving consistency between fine-grained conditions and video content, 2024
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 64ec8ce9-8681-4706-b58a-95507a8155e6 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Cogvlm: Visual expert for pretrained language models, 2024
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaa55a40-d8c5-4b28-8f62-263b46c10119 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Dynamicrafter: Animating open-domain im- ages with video diffusion priors, 2023
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c963fe9e-be22-4ef9-84d6-98c511a2faff · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM Easyanimate: A high-performance long video generation method based on transformer architecture, 2024
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 21470ee1-bf72-496c-b900-b03e1445f430 · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ee7e785-f87e-4cd5-b017-4364e2d1485d · outbound
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM I2vgen-xl: High-quality image-to-video synthesis via cascaded diffusion models, 2023
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e30efa21-81b8-40a6-877a-7caca3d92cd3 · inbound
Waver: Wave Your Way to Lifelike Video Generation Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e67b3fd2-7807-4c03-8bda-11054505215b · inbound
Image-to-Video Diffusion: From Foundations to Open Frontiers Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 080444ab-0976-4f05-9c09-eaea31c4d8c4 · inbound
SuperVoxelGPT: Adaptive and Ordered 3D Tokenization for Autoregressive Shape Generation Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e4f19425-5f3b-4bda-94f3-5cf54547d0c4 · inbound
EventOD: Event-Aware OD Flow Generation via LLM-Guided Semantic Modulation Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.