Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:18:27.810835Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 3 inbound Pith citation observations for arXiv:2506.05806.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:18:27.810835Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T06:45:49.611770Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-17T01:58:51.404766Z
40 of 40 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ca17bc80-f60d-4a04-98ef-1412b9481454 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models A lip sync expert is all you need for speech to lip generation in the wild
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 414366dc-3bdd-410c-bb28-36aa5f59796e · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Emo: Emote portrait alive generating expressive portrait videos with audio2video diffusion model under weak conditions
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f43dcb10-69d6-49e7-bdeb-7d8b0775adde · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Vasa-1: Lifelike audio-driven talking faces generated in real time.Advances in Neural Information Processing Systems, 37:660–684, 2024
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bff6e7e8-659e-4164-a4ff-5516565eb2fe · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 67a55ddd-daf9-4f89-b61c-93d999784cdc · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Denoising Diffusion Implicit Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c55c8c15-75b7-4eff-bc7f-ef6a04219bd3 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dd28e69-e6fd-4311-9756-8f0bc0ac0d65 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d26a6e7-3c58-4ad6-a627-9edd312330b6 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Megaportraits: One-shot megapixel neural head avatars
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bc068e31-25cc-4c12-bd41-c1c26721d8de · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models INFP: Audio-Driven Interactive Head Generation in Dyadic Conversations
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81f47936-4d82-4270-8276-134aa87531d0 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Echomimic: Lifelike audio-driven portrait animations through editable landmark conditions
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddb92de9-fd10-43c3-beb7-8a690711cdf2 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be5e88b4-c072-4605-8b1e-ab752b2d2f3e · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e92e6c34-72c1-457f-94e0-37c71a4b7bc5 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dce6ce4-191d-456a-94f1-149be2ffcbab · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a76ebc3-cbf2-4246-8763-529b4b596b24 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33133a8a-1ce5-4dcc-8d69-a991d324d416 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 697f31a6-831e-4965-8661-c50adda35c8e · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Consistency models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc55c5a8-c0ba-4577-ae14-afadc0db8298 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Animate anyone: Consistent and controllable image-to-video synthesis for character animation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5eaeb69-6529-43ee-972c-34fcf1bf1952 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39e9e265-0112-47c7-9195-558aaf3c522d · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models X-portrait: Expressive portrait animation with hierarchical motion attention
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f43c7580-7b67-4ca7-ace9-e864facf826f · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dc0c154-2846-459b-86c5-c2fdb5b91be2 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models First order motion model for image animation.Advances in neural information processing systems, 32, 2019
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c23bccc-b356-4b6b-bf42-54674df3b85a · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models OmniTalker: One-shot Real-time Text-Driven Talking Audio-Video Generation With Multimodal Style Mimicking
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adf63a13-6dec-4a02-bbb7-b95d647a2d93 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 717ece2f-dd5b-439a-aad0-e5076817d8f4 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Learning transferable visual models from natural language supervision
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0717dae3-1cfa-4ab1-995d-90e61e9b060a · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models GPT-4o System Card
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 968c5f37-a31d-4231-9f4e-a29f2ae3147a · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models wav2vec 2.0: A framework for self-supervised learning of speech representations.Advances in neural information processing systems, 33:12449– 12460, 2020
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbf8f96b-d270-49c2-93b7-92681762c5ff · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models wav2vec: Unsupervised Pre-training for Speech Recognition
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5943c728-876a-4f5d-874f-dffb4ad6e053 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models chinese speech pretrain, 2022
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2b9f9eef-7a4e-4e63-a5ac-b45a97da4f5b · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Hsemotion: High-speed emotion recognition library.Software Impacts, 14:100433, 2022
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a47c9991-4a64-4092-987f-714aa37a07b6 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models AnimateDiff-Lightning: Cross-Model Diffusion Distillation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50b578e2-7349-4c67-9925-0e2ecbd5ec06 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models SDXL-Lightning: Progressive Adversarial Diffusion Distillation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2e01aa1-ed59-4949-a6f3-8843792ae7a4 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Generative adversarial networks.Communications of the ACM, 63(11):139–144, 2020
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dacbc6b-1c76-4e54-a9b1-191da6f89350 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Animatelcm: Computation-efficient personalized style video generation without personalized video data
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e31a5c23-a23b-4ce4-9c00-b9aefb053ddf · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dac69101-cd27-46dd-8147-7781b68b1b9a · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Mead: A large-scale audio-visual dataset for emotional talking-face generation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6fab25fc-0325-4226-9fef-b30f63d9c48c · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models VoxCeleb2: Deep Speaker Recognition
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e97bcf03-adde-49d0-9043-b325031729a4 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models Celebv-hq: A large-scale video facial attributes dataset
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 548058e6-ad0b-41b6-8251-ab0ed6dd69a3 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models YOLOv6 v3.0: A Full-Scale Reloading
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2da14e8-311a-41ca-a8be-501cc43112c5 · outbound
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models MediaPipe: A Framework for Building Perception Pipelines
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d188d4fa-de9f-4e3e-9ba6-660af4482247 · inbound
Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fb09a11e-2136-4d79-978a-df66026497ea · inbound
Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72ceac6e-82b4-470d-b963-349472f310d9 · inbound
EchoCache: Energy-Guided Cross-Modal Caching for Efficient Audio-Driven Video Generation LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.