Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T04:46:05.382405Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2508.03050.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T04:46:05.382405Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
60 of 60 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0cdfd198-ace5-49ca-938f-ec0ff35a64bd · outbound
Multi-human Interactive Talking Dataset GitHub repository
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3acb0df4-a24a-403b-9b3e-0e9e8fe8397f · outbound
Multi-human Interactive Talking Dataset wav2vec 2.0: A framework for self-supervised learning of speech representations
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2c37f9f6-713c-4fea-96d7-9182bcb6f0a7 · outbound
Multi-human Interactive Talking Dataset TalkNet: Fully-Convolutional Non-Autoregressive Speech Synthesis Model
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4bd0cc6d-2ddf-452a-8569-95a465a7cd43 · outbound
Multi-human Interactive Talking Dataset Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d291381-c3c7-4d22-a4da-7a1cf7ef60c9 · outbound
Multi-human Interactive Talking Dataset Magicdance: Realistic human dance video generation with motions & facial expressions transfer
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0a5c21c2-a261-4a67-b674-442421b2a455 · outbound
Multi-human Interactive Talking Dataset VideoCrafter1: Open Diffusion Models for High-Quality Video Generation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04945ce8-166c-4c18-b9a1-af83f84ca9a4 · outbound
Multi-human Interactive Talking Dataset EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4105ac0e-03fb-4bd3-b2b7-cd02a2c2c66a · outbound
Multi-human Interactive Talking Dataset Out of time: automated lip sync in the wild
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f8c61c2-97e5-4f21-9eab-7ef66fa10556 · outbound
Multi-human Interactive Talking Dataset VoxCeleb2: Deep Speaker Recognition
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e201cba-1307-4505-960c-56ce975e0a70 · outbound
Multi-human Interactive Talking Dataset Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dafcda37-a4f0-4a18-8f21-1a03e6c906f9 · outbound
Multi-human Interactive Talking Dataset DreaMoving: A Human Video Generation Framework based on Diffusion Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b10e9c2-488c-4edc-934e-7963998b7bd4 · outbound
Multi-human Interactive Talking Dataset Affective Faces for Goal-Driven Dyadic Communication
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a6ebae4-3608-4ccb-8724-bdf0b704c5ae · outbound
Multi-human Interactive Talking Dataset Learning individual styles of conversational gesture
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 525143e0-4471-4645-9138-0b81f1c55acd · outbound
Multi-human Interactive Talking Dataset AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30321dbd-b666-4811-9777-d4e665d03ccd · outbound
Multi-human Interactive Talking Dataset Co-speech gesture video generation via motion- decoupled diffusion model
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dc9fa1c7-c6c9-4b7e-b603-3e1e9081c766 · outbound
Multi-human Interactive Talking Dataset Animate anyone: Consistent and controllable image-to-video synthesis for character animation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c3cbb69b-bda5-4073-830d-9de0291b2d3c · outbound
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dd891624-2dfa-48d8-8afb-3f224a0dd716 · outbound
Multi-human Interactive Talking Dataset Perceptual conversational head generation with regularized driver and enhanced renderer
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8ac39e5e-8a09-4631-bb69-4d28ba9a7697 · outbound
Multi-human Interactive Talking Dataset Vbench: Comprehensive benchmark suite for video generative models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12f37f79-c744-4589-aa63-959fc0534037 · outbound
Multi-human Interactive Talking Dataset Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b8685a9-ffe7-4a01-9409-9ecc6342e2cf · outbound
Multi-human Interactive Talking Dataset Text2performer: Text-driven human video generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6b6eb7d4-1ed5-4b79-a431-773593f37b62 · outbound
Multi-human Interactive Talking Dataset Whole-body human pose estimation in the wild
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 34967ca4-0aae-41e9-8366-dd2a3c3760c0 · outbound
Multi-human Interactive Talking Dataset Sapiens: Foundation for human vision models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8873e075-bb5b-4ce9-8ca0-52ff8e6e0a6b · outbound
Multi-human Interactive Talking Dataset A Comprehensive Survey on Human Video Generation: Challenges, Methods, and Insights
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 782103c6-d6b4-4152-bf72-c0d101291091 · outbound
Multi-human Interactive Talking Dataset OpenHumanVid: A Large-Scale High-Quality Dataset for Enhancing Human-Centric Video Generation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9370856a-bb87-47d2-87dd-38fe730e3878 · outbound
Multi-human Interactive Talking Dataset TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fb68e29-6299-4801-804f-d9269e44cf0e · outbound
Multi-human Interactive Talking Dataset Customlistener: Text-guided responsive interaction for user-friendly listening head generation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 89b6971c-5938-4a0a-9e98-bee54b5bd831 · outbound
Multi-human Interactive Talking Dataset Learning hierarchical cross-modal association for co-speech gesture generation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4848fdaa-7499-4bce-8010-c945cdee0496 · outbound
Multi-human Interactive Talking Dataset Follow your pose: Pose-guided text-to-video generation using pose-free videos
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5610ee8e-98c2-4014-9d1b-1a3db1726d28 · outbound
Multi-human Interactive Talking Dataset Learning to listen: Modeling non-deterministic dyadic facial motion
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67282c5a-0dc5-4cc9-8554-8913de99d050 · outbound
Multi-human Interactive Talking Dataset Scalable diffusion models with transformers
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 882adab5-7d3a-4c8a-9520-a8f3cedb2778 · outbound
Multi-human Interactive Talking Dataset ControlNeXt: Powerful and Efficient Control for Image and Video Generation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44850ca4-9bf1-4f9c-9384-6555bf76b903 · outbound
Multi-human Interactive Talking Dataset A lip sync expert is all you need for speech to lip generation in the wild
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed8189e5-f776-49dd-993b-93e31b7833c8 · outbound
Multi-human Interactive Talking Dataset Speech drives templates: Co- speech gesture synthesis with learned templates
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 80a096b3-0296-45be-a6e4-6eb1d2eae47a · outbound
Multi-human Interactive Talking Dataset High- resolution image synthesis with latent diffusion models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0b63739-a436-4cef-9232-56b0f0fa4212 · outbound
Multi-human Interactive Talking Dataset MediConfusion: Can you trust your AI radiologist? Probing the reliability of multimodal medical foundation models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c89c97e4-3d0e-475b-add1-2a0b11bc546d · outbound
Multi-human Interactive Talking Dataset Make-A-Video: Text-to-Video Generation without Text-Video Data
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52ef254d-7061-4244-ac43-24e4e974b57d · outbound
Multi-human Interactive Talking Dataset Lip reading sentences in the wild
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f110161b-2b44-47a6-848e-45469ae2262d · outbound
Multi-human Interactive Talking Dataset Llama Learns to Direct: DirectorLLM for Human-Centric Video Generation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49d27cb2-3f02-41a5-a0a7-1e78d1932a71 · outbound
Multi-human Interactive Talking Dataset Diffused heads: Diffusion models beat gans on talking-face generation
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1c5507d2-af67-4c7a-9fe8-8d898218fdb2 · outbound
Multi-human Interactive Talking Dataset MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0c1f5e6-d51c-4daf-b9f5-2d5f6367ca36 · outbound
Multi-human Interactive Talking Dataset Edtalk: Efficient disentanglement for emotional talking head synthesis
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a3925a25-8fac-4919-8aa6-840e946351e4 · outbound
Multi-human Interactive Talking Dataset Dyadic Interaction Modeling for Social Behavior Generation
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ec413d04-8900-4808-9b44-97b1fa0e29ae · outbound
Multi-human Interactive Talking Dataset Realistic speech-driven facial animation with gans
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation accf5f45-1e4d-4eda-b2a4-6609cff35088 · outbound
Multi-human Interactive Talking Dataset AgentAvatar: Disentangling Planning, Driving and Rendering for Photorealistic Avatar Agents
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3894480e-4fe8-4be3-8cb4-4d99737df128 · outbound
Multi-human Interactive Talking Dataset EmotiveTalk: Expressive Talking Head Generation through Audio Information Decoupling and Emotional Video Diffusion
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c01e8b70-a59c-4002-baac-3ec27f6e87c4 · outbound
Multi-human Interactive Talking Dataset Mead: A large-scale audio-visual dataset for emotional talking-face generation
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c791ccfc-cd29-40f0-84ed-f9e2f1b706e4 · outbound
Multi-human Interactive Talking Dataset Draganything: Motion control for anything using entity representation
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cde52439-12ea-4f8b-9e43-9842e2f3b7b2 · outbound
Multi-human Interactive Talking Dataset Magicanimate: Temporally consistent human image animation using diffusion model
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 194abaf8-9235-4387-ae97-1d68579224ba · outbound
Multi-human Interactive Talking Dataset Cogvideox: Text-to-video diffusion models with an expert transformer
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 93cc9205-cab3-4e42-acf8-19831313e13f · outbound
Multi-human Interactive Talking Dataset Show-1: Marrying pixel and latent diffusion models for text-to-video generation
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8b1f1e3a-64e0-491e-8084-8de15dbb6c95 · outbound
Multi-human Interactive Talking Dataset Adding conditional control to text-to-image diffusion models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08bde3ec-3b0c-4783-a73b-0dd0982db4bb · outbound
Multi-human Interactive Talking Dataset Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4fb099b9-1173-4346-84d0-3e624918e8a1 · outbound
Multi-human Interactive Talking Dataset Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a8d6a4bf-0880-40a8-be54-6ae719648052 · outbound
Multi-human Interactive Talking Dataset Responsive listening head generation: a benchmark dataset and baseline
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a2503984-6b32-4526-a0ac-39f27d8b765f · outbound
Multi-human Interactive Talking Dataset Interactive Conversational Head Generation
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 714cd6f0-3ada-4664-ae7a-1e3308d5ebf0 · outbound
Multi-human Interactive Talking Dataset Audio-driven neural gesture reenactment with video motion graphs
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a22c984e-8a12-4baa-bdc1-e6c6364a5ffb · outbound
Multi-human Interactive Talking Dataset Celebv-hq: A large-scale video facial attributes dataset
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 94f3564d-f034-4b36-8adb-d63cb3e3a24d · outbound
Multi-human Interactive Talking Dataset Taming diffusion models for audio-driven co-speech gesture generation
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8569be9f-6540-496b-9667-ed88b5a6cfde · outbound
Multi-human Interactive Talking Dataset INFP: Audio-Driven Interactive Head Generation in Dyadic Conversations
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.