Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 58 inbound Pith citation observations for arXiv:2311.10709.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:57:28.806572Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T12:36:57.125908Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation ee8b18b6-b8ba-42f6-bc35-867b0c0204fc · inbound
Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 196
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3cb89a74-e556-407c-ab97-f64cabe95f79 · inbound
RoboDreamer: Learning Compositional World Models for Robot Imagination Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cb5caf97-71c3-4f11-abcf-dfb6933fb220 · inbound
CAT3D: Create Anything in 3D with Multi-View Diffusion Models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3dae8dae-6b11-4043-a03c-219153506474 · inbound
Diffusion Models Are Real-Time Game Engines Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 12b802e2-f91b-4393-b4ce-c66e0fd20f86 · inbound
Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c66e20ca-dc1f-4a25-88e9-b04316bbde4c · inbound
VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0eb5f3c8-f23e-4dfe-b862-5d4e6a6e63e7 · inbound
KFC-W: Generating 3D-Consistent Videos from Unposed Internet Photos Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e5049670-8073-4e81-a47b-7d23aaf1a366 · inbound
Generative Omnimatte: Learning to Decompose Video into Layers Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 754f10b5-a258-4902-ad67-7b47db1fafff · inbound
$\textit{Revelio}$: Interpreting and leveraging semantic information in diffusion models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0625430-8b40-479d-b21c-e00cb918bcaa · inbound
Towards Precise Scaling Laws for Video Diffusion Transformers Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8d0ba44-3b67-4b21-a455-a9981a39633c · inbound
Video-Guided Foley Sound Generation with Multimodal Controls Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2956df3-1384-4350-818a-87ae39263b43 · inbound
I2VControl: Disentangled and Unified Video Motion Synthesis Control Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcd1968c-67b0-4456-bd70-fd12bf6b028d · inbound
ROICtrl: Boosting Instance Control for Visual Generation Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0936cd5f-c0a4-4279-b6ab-c5135c9b3780 · inbound
ReconDreamer: Crafting World Models for Driving Scene Reconstruction via Online Restoration Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15fc769f-c302-4013-a2d4-a16049c5069e · inbound
MoTrans: Customized Motion Transfer with Text-driven Video Diffusion Models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b9ddeb9-6603-4097-b682-afc21f3c449a · inbound
World-consistent Video Diffusion with Explicit 3D Modeling Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71c850b4-401d-4bf5-9ee1-348ffdf0f099 · inbound
ShowHowTo: Generating Scene-Conditioned Step-by-Step Visual Instructions Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1886df4d-cb7a-4ed0-b01e-aea208aaa508 · inbound
Motion Prompting: Controlling Video Generation with Motion Trajectories Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 026324f5-ada2-4885-bae7-ef28f7f22124 · inbound
Navigation World Models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 681ba316-c6ff-45ab-9fc8-5b1b6a570203 · inbound
HunyuanVideo: A Systematic Framework For Large Video Generative Models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ec7def8e-cb2a-45f5-80aa-6db2cea31c5e · inbound
Factorized Video Autoencoders for Efficient Generative Modelling Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30e2ec5f-24e8-411b-aac1-02ccbc6817ed · inbound
SynCamMaster: Synchronizing Multi-Camera Video Generation from Diverse Viewpoints Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25442122-0eaa-4598-bfad-f4a1aed5da86 · inbound
DrivingRecon: Large 4D Gaussian Reconstruction Model For Autonomous Driving Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1c7108d-86e9-429d-a487-cdda0b005945 · inbound
LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c3d6d94-b945-458c-b211-343982de1c5e · inbound
Video Diffusion Transformers are In-Context Learners Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed0f69db-665f-483b-b7db-d425581e2d33 · inbound
GEM: A Generalizable Ego-Vision Multimodal World Model for Fine-Grained Ego-Motion, Object Dynamics, and Scene Composition Control Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48ecadf3-075f-4935-9102-38e3f58ebf4e · inbound
MotionBridge: Dynamic Video Inbetweening with Flexible Controls Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d08671be-7965-4847-96e7-52a5dd099746 · inbound
Flowing from Words to Pixels: A Noise-Free Framework for Cross-Modality Evolution Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b742eb0c-580f-4f5b-b34d-c306fa973399 · inbound
VideoMaker: Zero-shot Customized Video Generation with the Inherent Force of Video Diffusion Models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcb5023a-3e47-4494-b766-3ca7526411e2 · inbound
Ingredients: Blending Custom Photos with Video Diffusion Transformers Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1465803-95ba-4758-900e-5490b1021b54 · inbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fb7dced-ad04-48dd-8b2a-487eaa9b2b66 · inbound
Generative Physical AI in Vision: A Survey Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c0bdc77-daf9-4d5d-a6c7-979593719cff · inbound
Diffusion Autoencoders are Scalable Image Tokenizers Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7d2e53e-9376-4a61-a257-5282fcc8e6c7 · inbound
Goku: Flow Based Video Generative Foundation Models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73e4a5ff-ba75-45d7-9cc4-40c385a0e376 · inbound
Next Block Prediction: Video Generation via Semi-Autoregressive Modeling Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9411a428-63b1-4420-979f-a9456d8543e8 · inbound
Movie Weaver: Tuning-Free Multi-Concept Video Personalization with Anchored Prompts Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9622a6ce-725c-4a86-812f-9090e3a01a44 · inbound
RealCam-I2V: Real-World Image-to-Video Generation with Interactive Complex Camera Control Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae12a699-9f90-485d-98fc-2c21d24d3013 · inbound
Unified Video Action Model Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2e1445b9-2383-4218-86fc-de697303dfc7 · inbound
Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 19c71e63-3ee7-4435-aa98-11c0b565e9ce · inbound
T2S: High-resolution Time Series Generation with Text-to-Series Diffusion Models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e049380e-28a3-4eb2-b774-cb6c72371dfd · inbound
Character-Centered Dialogue Generation from Scene-Level Prompts Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e17a50bd-73a3-493e-989c-9ce2743ee4c2 · inbound
EPiC: Efficient Video Camera Control Learning with Precise Anchor-Video Guidance Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70454a36-adb2-49a2-80ba-a43db546785d · inbound
Medical World Model: Generative Simulation of Tumor Evolution for Treatment Planning Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f2b9577-c8c0-446a-9ca7-5e76c575eab1 · inbound
Edit360: 2D Image Edits to 3D Assets from Any Angle Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 761b04be-0f23-4ca1-be7d-3d9049da1c52 · inbound
From 2D to 3D Cognition: A Brief Survey of General World Models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 520d1e4f-28c8-4986-9d4f-61ed11c87aea · inbound
Consistent Zero-shot 3D Texture Synthesis Using Geometry-aware Diffusion and Temporal Video Models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f5dd659-901c-45e7-aeb2-0eab12855c8c · inbound
LiON-LoRA: Rethinking LoRA Fusion to Unify Controllable Spatial and Temporal Generation for Video Diffusion Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aca4e05c-c417-4e7a-8288-279394392d66 · inbound
Waver: Wave Your Way to Lifelike Video Generation Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb7500c6-0093-41dd-9a46-b97ba90ea362 · inbound
Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 397bf284-b29c-40b2-9a61-7039541f728b · inbound
Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b44a179a-fd78-4e80-866a-b537cbe73adc · inbound
Evolution of Video Generative Foundations Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8aae5da8-096e-45ae-86a1-e987c5ee7132 · inbound
CityRAG: Stepping Into a City via Spatially-Grounded Video Generation Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d2706235-b757-457b-8195-09d8ec60f07d · inbound
DCR: Counterfactual Attractor Guidance for Rare Compositional Generation Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation eb7a8750-1c8c-42cb-8804-28603fa479e0 · inbound
Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 150be9d5-dcc2-4da5-ab76-cf4c28c30415 · inbound
StreamingEffect: Real-Time Human-Centric Video Effect Generation Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f3495b01-7d04-45d5-9b88-553a08e7dddf · inbound
Image-to-Video Diffusion: From Foundations to Open Frontiers Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c51df560-1140-455f-b80a-4d35e944fb67 · inbound
Physics in 2-Steps: Locking Motion Priors Before Visual Refinement Erases Them Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 055ba8a3-e502-43ba-b1ce-7629263d1522 · inbound
TPD: Temporal Prior Decoupling for Text-to-Video Diffusion Models Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.