Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T21:24:08.721929Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2508.08891.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T21:24:08.721929Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
40 of 40 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation de2bf21f-4036-459e-8002-dcefadac5b4b · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f01b6fc9-2761-4af7-8c24-c0df7620a36c · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Emerg- ing properties in self-supervised vision transformers
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66a2cb35-9ef2-4aa1-990f-4f17cadf014f · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Dimitra: Audio-driven Diffusion model for Expressive Talking Head Generation
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d7248bf-9fae-4233-a7e6-de05a57562a0 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d160ab63-8d60-43af-b7dd-56f2ed06878c · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d42da7a8-dbf2-4411-9906-0afc46933a68 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Gemini 2.5 pro: Our most intelligent ai model, 2025
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dd4d3973-3b7d-4bba-98d6-54fc4a4c9f97 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Stylesync: High-fidelity generalized and personalized lip sync in style-based genera- tor
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 02cab4ec-693a-4221-b10c-8befa4cd10b2 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34596d63-1629-49b9-8b82-950be13908ba · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Video dif- fusion models.Advances in Neural Information Processing Systems, 35:8633–8646, 2022
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fa83f4b-3bcc-4a19-bffc-ba006d6dbe76 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Step-Video-TI2V Technical Report: A State-of-the-Art Text-Driven Image-to-Video Generation Model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6471fbcf-7a4d-4bf6-921e-b3a5270daf23 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos The accu- racy of psnr in predicting video quality for different video scenes and frame rates.Telecommunication systems, 49:35– 48, 2012
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e54293de-0b29-48d9-924f-b7446fb7133d · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Sonic: Shifting Focus to Global Audio Perception in Portrait Animation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80926fae-08ce-4f87-9008-0757494756b5 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 893ea3b1-3f2f-4b33-880b-bca847ba4a81 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Musiq: Multi-scale image quality transformer
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 994d9fab-d6d1-4b4d-8c98-01ac622ec4b9 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos HunyuanVideo: A Systematic Framework For Large Video Generative Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27d6c694-364a-4966-928b-b35d9c2e7396 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Luma ray 2 video model.https : / / lumalabs.ai/ray
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8bcc75af-195b-48d1-bd86-58844b0ad169 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Laion aesthetic predictor.https : / / github.com/LAION-AI/aesthetic-predictor,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 98cb3211-2ac2-4a1d-8998-c10475b07b4b · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57901d56-65ea-4a19-9d9f-23e6f2a6ce66 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57d2b5d0-0628-4f84-b681-3eafd24e774d · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Amt: All-pairs multi-field transforms for efficient frame interpolation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8685b4a7-33cc-4bf9-98ef-1e5cffc6b000 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Echomimicv2: Towards striking, simplified, and semi-body human animation.arXiv preprint arXiv:2411.10061, 2024
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c88e2b08-679f-473c-a951-557d39ae9b46 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos StyleTalker: One-shot Style-based Audio-driven Talking Head Video Generation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9da70060-4f55-40ab-a22f-0f3d457aee67 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Open-Sora 2.0: Training a Commercial-Level Video Generation Model in $200k
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7948f18-721e-4235-828d-3ad5a26b8791 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos A lip sync expert is all you need for speech to lip generation in the wild
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 67dfffa3-33ec-4b33-a529-05935706537a · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Versatile Multimodal Controls for Expressive Talking Human Animation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5101bf6c-1045-4805-b9a0-3fa55900749e · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos SkyReels-A1: Expressive Portrait Animation in Video Diffusion Transformers
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f735ff8-6cea-431c-bd36-ad67c6eb923c · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Learning transferable visual models from natural language supervi- sion
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aeca28fe-9986-42bb-a0a5-5681bdb3fc2b · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Make-A-Video: Text-to-Video Generation without Text-Video Data
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd24ca58-02cc-4f59-afc6-e10a8b08301d · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Diffused heads: Diffusion models beat gans on talking-face genera- tion
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5a0d1fb3-327d-4405-8d2e-8cd94feae4db · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Raft: Recurrent all-pairs field transforms for optical flow
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf2d0577-8566-4b21-bb82-73ba931ad1ae · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos StableAnimator: High-Quality Identity-Preserving Human Image Animation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42b3d4fe-6134-4f02-b366-e5a92e2e859a · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Towards Accurate Generative Models of Video: A New Metric & Challenges
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc65f202-0682-40f7-ace3-c363d5aed292 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Wan: Open and Advanced Large-Scale Video Generative Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3182e70d-7162-457e-a16b-6345e95beead · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Lavie: High-quality video generation with cascaded latent diffusion models.International Journal of Computer Vision, pages 1–20, 2024
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f3bcd8ed-ba27-4c22-85a8-309c428e58a6 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Image quality assessment: from error visibility to structural similarity.IEEE transactions on image processing, 13(4):600–612, 2004
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e2d9b1d-8293-4304-a815-28494b7920de · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Easyanimate: A high-performance long video generation method based on transformer architecture.arXiv preprint arXiv:2405.18991,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6d1aa7b-b701-498d-aa61-9fe116bdff55 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adcbc9e8-b530-4c38-8aea-e0c81a9ad333 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0689d79-1e84-4c1b-b2cf-a75876eef380 · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Pose-controllable talking face generation by implicitly modularized audio-visual rep- resentation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22c0ad9a-3c42-4827-90bf-7a15d8360f3d · outbound
Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos Unresolved cited work
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.