Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:30:33.055697Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 4 inbound Pith citation observations for arXiv:2506.07848.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:30:33.055697Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-11T20:11:31.576642Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T22:56:20.169424Z
48 of 48 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c4dcf0d4-b5af-472c-bf9a-4064cd2a576f · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Qwen2.5-VL Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65232c13-4a34-4297-b1a9-9fa062e9ebd3 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26b2b07c-bdf0-48a2-a14d-bf3ccb8b3f0a · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Carreira and A
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9e15de1b-3f68-4af7-9c4c-21e0d3498141 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Chefer, S
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation acd89131-1c62-4c1a-9edb-ce308a4e98e7 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 014d3878-8673-46d3-a57e-5a234eb77074 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e7ee9805-6d3c-4c63-8ac3-f8620a328848 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Multi-subject Open-set Personalization in Video Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b0db90e-4a28-4768-967f-fc6571f623e4 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9fd35efb-5bd5-4fc8-b149-aa5b49f6e6b2 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Esser, S
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d13d751c-5cc6-4351-86f9-1e9a92164b10 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement SkyReels-A2: Compose Anything in Video Diffusion Transformers
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 641b8e7e-ff02-4cd9-8491-e5529d7b9da5 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71e00960-50b1-449a-aae8-59de3dc5b93f · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Hailuo.https://hailuoai.video/, 2025
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 09860b72-42ef-44ba-9626-a86e246c6e5c · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement ID-Animator: Zero-Shot Identity-Preserving Human Video Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51327993-868a-4a4a-8765-b91f4ff8d639 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ee2d3f2-0ecd-49bb-9c5e-6daf89019b76 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement MotionMaster: Training-free Camera Motion Transfer For Video Generation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 788063b8-c336-4922-8764-25d8cbfc7f70 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2ebc5f8-5042-4246-882d-8f37098a098e · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 540271da-c943-4812-9335-e520231aefe1 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Huang, Y
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52ecb108-716a-4ba1-91b0-b5532fb7517b · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Jiang, T
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation adf7c961-802f-4aa7-b3aa-19e0f32162d4 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement VACE: All-in-One Video Creation and Editing
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f88c9a3f-7b87-4f0e-b7b8-ba0fa0638e13 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Keling.https://klingai.com/cn/, 2025
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7c5a358e-d494-44a7-bdc6-fdf727b93f40 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement YOLOv11: An Overview of the Key Architectural Enhancements
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 852fb42d-a9c3-4703-b06c-181505fcf0b5 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement HunyuanVideo: A Systematic Framework For Large Video Generative Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b117c48f-1a0b-4a48-bd43-d0cf9555ecd3 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 20c8f669-c750-4943-8128-6c68e642efac · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement OmniV2V: Versatile Video Generation and Editing via Dynamic Content Manipulation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ad2a0c6-d9a0-4088-b3d5-d113549528fa · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e62295d4-8fa2-4396-8deb-fac1b7d52bf6 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Phantom: Subject-consistent video generation via cross-modal alignment
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98a3f92e-9485-4ece-8bf1-d5fcfb7d2279 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation da3df5d8-0e4d-47c9-b56b-87c5fdc313fc · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a65e3a7-a8e5-431d-b192-dfaedde6aae6 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement DINOv2: Learning Robust Visual Features without Supervision
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20e6f371-1ccb-40d5-bbc1-4607a7dcaf99 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Peebles and S
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eba37099-d398-4880-b291-d32f32d5b804 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Pika.https://pika.art/, 2025
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe2140d3-0410-4633-b890-35f32edab8e9 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Movie Gen: A Cast of Media Foundation Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1193e3a8-e1ee-497a-bbb3-d7bfede70c0f · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Radford, J
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d58ba38d-d620-4cb4-b579-57496e270ca5 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement SAM 2: Segment Anything in Images and Videos
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bd4dd8c-3ca5-418b-9e70-6d987e5a7d91 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Rombach, A
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22903861-99de-4589-9c12-032ff825573f · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2170f9ea-fbf8-402f-9956-73cff330ae2a · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement OminiControl: Minimal and Universal Control for Diffusion Transformer
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ab8caa8-94fe-4a1f-bd69-fe8ea910bf47 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Vidu.https://www.vidu.cn/, 2025
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation aaa5cc40-d3d4-4670-90f3-c313426ea273 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Wan: Open and Advanced Large-Scale Video Generative Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b968ddfb-b492-40fb-9464-ff3a9d280d70 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Koala-36M: A Large-scale Video Dataset Improving Consistency between Fine-grained Conditions and Video Content
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05864c36-0f1c-4b0e-8657-9709df16deae · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71b94bfd-edf0-4df3-b9c6-891b5e00de7a · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 09a7381a-dea6-4daf-adfb-680a20eb4248 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f8307a88-ae46-4052-bf27-b46c3ef04f15 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97b49b59-a33d-416b-bff2-fb5d884efcde · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3761732-b039-47c0-a26f-1e59f3b149d6 · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Identity-Preserving Text-to-Video Generation by Frequency Decomposition
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef30a557-886f-49d0-8ac8-70202a503d7d · outbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Allegro: Open the Black Box of Commercial-Level Video Generation Model
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d494d21-a065-448a-af60-769d4e10c20c · inbound
Evolution of Video Generative Foundations PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement
Reference 206
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3d7b7d1c-4e97-49d1-9e0f-205daa78cebf · inbound
PresentAgent-2: Towards Generalist Multimodal Presentation Agents PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 49a8ca39-c798-42c2-9a7a-58f88f3ce637 · inbound
MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f1f519fa-6a6d-4a2d-baa5-688f1e0aaf91 · inbound
Aura: Consistent Multi-Subject Video Generation via VLM-Grounded Semantic Alignment PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.