Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:29:41.247052Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 4 inbound Pith citation observations for arXiv:2506.19833.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:29:41.247052Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T20:06:46.703429Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T07:21:23.863541Z
56 of 56 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9ae6cfe5-ec19-4624-9fb5-bce06d374d66 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router wav2vec 2.0: A framework for self- supervised learning of speech representations.Advances in neural information processing systems, 2020
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 67cbf082-6afe-47f7-92ae-964a83bc28a0 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Qwen Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc12339c-6837-4ef5-b3dc-a834be500af1 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbce61ed-b268-474d-b389-49af8bed024c · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 071dee7c-97a3-44ba-a21b-8e192002fdca · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4303f6c1-5ee0-41ad-8b16-bdf48627e9e0 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Out of time: Automated lip sync in the wild
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 413fcc20-f2e7-4be2-896b-3263e2d8569d · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 894cea0e-1e9c-4818-b037-35899f37e355 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3cddfa5-64db-4e63-b678-8e547a6ddc66 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Insightface: An open-source 2d and 3d deep face analysis toolkit.https://github.com/ deepinsight / insightface, 2024
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e11e0319-0115-4f3c-872c-b6a6bf87de3a · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router MAGREF: Masked Guidance for Any-Reference Video Generation.arXiv preprint arXiv:2505.23742,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51c4967b-8fd6-4bf9-89cc-0a5730644613 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d2e6acc-5680-44a5-9f9d-1102e2ec2c34 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Ingredients: Blending custom photos with video diffusion transformers, 2025
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c8588426-c846-4c61-b75a-53637396cf4f · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router AD-NeRF: Audio Driven Neural Radiance Fields for Talking Head Synthesis
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 123a781b-c7b8-438f-a9fb-3a403dc6500a · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 2017
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0f2a3271-9c90-48cc-9389-d450073965bd · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Classifier-Free Diffusion Guidance
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9455f4c6-90f8-491f-8e1c-509e2c678394 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Depth-Aware Generative Adversarial Network for Talking Head Video Generation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0ac8b7b0-0015-45a8-816a-4c24221f424b · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e1cc1add-0c90-4bf9-8013-b43a171df0f9 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3185e00d-a467-465d-8335-82e1606546b4 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Sonic: Shifting Focus to Global Audio Perception in Portrait Animation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15a9ad3f-e860-469a-8b65-14ddb0865d37 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Ultralytics yolov8, 2023
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation aa4d0478-a48b-47e1-9c22-bf1c1b355111 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ae4bcb1-cfd6-436e-84f5-b1099d0c7f87 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 61c42d9a-1c42-4ce1-8918-8f2b5726e6f3 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a88c2ad-a372-417a-9364-9ad441156dad · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router EchoMimicV2: Towards Striking, Simpli- fied, and Semi-Body Human Animation.arXiv preprint arXiv:2411.10061, 2024
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 439bd3d7-e367-4777-9bec-566685bc2968 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Let's Go Real Talk: Spoken Dialogue Model for Face-to-Face Conversation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37fc7fac-fece-4695-a863-abbfb37dc2ae · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b0eb454a-880c-4762-8bb2-5e3c32d7d3e7 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1089d791-55eb-4a4b-8a70-e20c46eda12b · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Learning transferable visual models from natural lan- guage supervision
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ecd26ade-a319-47b2-a691-85ced96058b3 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Exploring the limits of transfer learn- ing with a unified text-to-text transformer.Journal of machine learning research, 2020
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 82b4e6de-7bc0-4979-bdc7-19f390b510a2 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router SAM 2: Segment Anything in Images and Videos
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 185b3d47-9e4b-4acd-948a-ea5257ad4149 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Facenet: A unified embedding for face recog- nition and clustering
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3bcd6596-ef3a-4c17-8633-ccbc8491f0b8 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Clearvoice: An open-source audio-visual speech processing toolkit
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3209e392-4c96-4471-94d8-462e36ccfac2 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Roformer: Enhanced trans- former with rotary position embedding.Neurocomput- ing, 2024
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ea381be4-5efc-4aaa-8261-4c3570f8d594 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Rethinking the inception architecture for computer vision
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1465142b-8e95-4a94-9831-939cb069a48c · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router EMO: Emote Portrait Alive Generating Expressive Por- trait Videos with Audio2Video Diffusion Model Under Weak Conditions.European Conference on Computer Vision, 2024
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation da3f1d5b-9ad4-4a4b-a287-15db42aa1353 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router EMO2: End-Effector Guided Audio-Driven Avatar Video Generation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f81d63b5-77a2-43fb-8adf-0d56fb2f5862 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Fvd: A new metric for video generation.Inter- national Conference on Learning Representations, 2019
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6635004c-1458-4206-9d79-89e730e8ec3a · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4977eb1-9df5-4fb5-afbb-6613645e2b27 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router JoyGen: Audio-Driven 3D Depth-Aware Talking-Face Video Editing
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94a023f8-43c8-4607-8568-0501ea41f3a1 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Friends-mmc: A dataset for multi-modal multi-party conversation under- standing
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation eb926631-9f96-4f14-965e-53cf9066cb53 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router MoCha: Towards Movie-Grade Talking Character Synthesis
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 526a7123-c880-447b-b8fc-f246009736d0 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Seeking the Shape of Sound: An Adaptive Framework for Learning V oice-Face Association.Computer Vision and Pattern Recognition, 2021
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c6eef73f-22e4-4800-92df-6f11d9dc3317 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Sophia Koepke, and Andrew Zisserman
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8d64fd63-8ca5-424f-bfd0-f54732f4fe4c · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Vfhq: A high-quality dataset and benchmark for video face super-resolution
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 46fe26e0-e612-4faf-914e-64ce8570f192 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router MegActor-$\Sigma$: Unlocking Flexible Mixed-Modal Control in Portrait Animation with Diffusion Transformer
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5490a590-c3e7-428c-96b6-2d729b5d2735 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6fefac0-7751-4e12-b0d0-766902a442ef · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router MagicInfinite: Generating Infinite Talking Videos with Your Words and Voice
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a60fa5e2-e5a0-4658-bb2f-0dffa95a4423 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Identity-preserving text-to-video generation by frequency decomposition
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 603caf57-e8fd-4926-947d-21f2bc562819 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9693cf19-839d-469d-9428-0eb9033de889 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ff7a4695-16b7-4eb7-a8de-d792b9ad4f75 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Teller: Real-Time Streaming Audio-Driven Portrait Animation with Autoregressive Motion Generation
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f35c5b6b-082a-4bdf-9cd7-1ee8743b90de · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Identity-Preserving Talking Face Generation with Landmark and Appearance Priors
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a55f6395-40da-4d8c-b826-c1f05efd0f3b · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Concat-ID: Towards Universal Identity-Preserving Video Synthesis
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff41877d-6a52-4214-82ca-ba08d7e8b4c0 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Responsive listening head gener- ation: a benchmark dataset and baseline
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 206ac9ac-1855-4936-931f-e1632cf09446 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router Makeittalk: Speaker-aware talking-head animation
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 65f06d46-f32a-4f29-bb42-726d23356af8 · outbound
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router INFP: Audio-Driven Interactive Head Generation in Dyadic Conversations
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6972d79b-1c29-4e4c-962c-3ad9928f8c59 · inbound
FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 637da00f-17f4-4838-b577-1697b3dc2884 · inbound
MIDAS: Multimodal Interactive Digital-humAn Synthesis via Real-time Autoregressive Video Generation Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1000af5a-71de-40fa-a0fb-7c935a561a92 · inbound
SocialDirector: Training-Free Social Interaction Control for Multi-Person Video Generation Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6b4962d2-9e61-441b-ad8b-528320b0f77e · inbound
EchoCache: Energy-Guided Cross-Modal Caching for Efficient Audio-Driven Video Generation Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.