Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T22:45:39.443440Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2605.31294.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T22:45:39.443440Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 238f65e8-cfd0-4a26-9957-7a4733319a1c · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62fac434-eb11-4576-8a5c-137dd0a46e75 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Distributed by Warner Bros
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbd75caa-c90c-48b8-bc4b-f1eb4b076a1b · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Gesturediffu- clip: Gesture diffusion model with clip latents.ACM Trans
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5779646-f04e-4321-b3a1-3406597830be · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens wav2vec 2.0: a framework for self-supervised learning of speech representations
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3532c69-855b-40a1-b1c8-d71ee4263583 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3a26bec-6bf7-4ffc-934d-f9004d3e69be · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Cohen and Dominic W
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e791cee-e7e1-48d0-a22e-f06f8b8e55d7 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Emotional speech-driven animation with content-emotion disentangle- ment
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4772aacd-727b-45ab-8376-35d1485ee341 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Moshi: a speech-text foundation model for real-time dialogue
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3572fb52-23a1-45dc-8585-94bfc1396daa · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens CosyV oice: A Scalable Multilin- gual Zero-shot Text-to-speech Synthesizer based on Super- vised Semantic Tokens
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab9c697c-b568-4e20-8332-32449bde583a · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens High Fidelity Neural Audio Compression
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54215726-8171-4e03-b9e1-15f1807f9e64 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens JALI: an animator-centric viseme model for expres- sive lip synchronization.ACM Transactions on Graphics, 35 (4):1–11, 2016
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c1e6fb2-4506-4646-b577-744c0bd4a6df · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Jali-driven expressive facial animation and multilin- gual speech in cyberpunk 2077
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d237f36a-daad-46c6-a01e-272f84860120 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Papka, Sanjif Shanmugavelu, Darshan Gandhi, Hengyu Zhao, Dun Ma, Kiran Ranganath, Rick Weisner, Jiunn-yeu Chen, Yuting Yang, Natalia Vas- silieva, Bin C
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cd5da28-e73a-4841-b2a2-47e90f21f24b · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Faceformer: Speech-driven 3d facial anima- tion with transformers
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a43b7391-d3e0-4f63-9793-c00a05313c7a · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens LLaMA-Omni: Seamless Speech Interaction with Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation abb55630-1a12-4b13-8d71-a9c68bac0533 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Tiny is not small enough: High quality, low- resource facial animation through hybrid knowledge distil- lation.ACM Trans
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f6a6edd-df17-4537-864c-70369f204574 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Classifier-free diffusion guidance, 2022
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05df52ea-6355-4b7f-907a-066a59850658 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens HuBERT: Self-Supervised Speech Representa- tion Learning by Masked Prediction of Hidden Units
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a41fa6a5-1451-481f-a4ac-bd38485d0055 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Speed- aware audio-driven speech animation using adaptive win- dows.ACM Transactions on Graphics, 44(1):1–14, 2024
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efda1815-324f-4f86-b2a5-ed21a5b461e8 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Audio-driven facial animation by joint end- to-end learning of pose and emotion.ACM Trans
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ca5f8fd-2dd8-44eb-8ea6-d08fca1f832f · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Audio Driven Real-Time Facial Animation for Social Telep- resence
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d9d66ad-16ec-46bf-a3b5-bd40822786e5 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Ditto: Motion-Space Diffusion for Control- lable Realtime Talking Head Synthesis
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79d377d4-9e4a-42c7-98b9-25a8f6d0c32e · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8aa55d09-f1d0-469c-b029-c07b142a810c · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3410bfb-594d-426b-bda8-2721f33ee2d6 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Medtalk: Multimodal controlled 3d facial animation with dynamic emotions by disentangled embed- ding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79612e0c-73df-4be6-bc9c-b3919089cbdd · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07dfb8f3-ff33-4cea-abdc-4f6a4d9a3e96 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Learning to Listen: Modeling Non-Deterministic Dyadic Facial Motion
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9be75481-cf9d-4e5e-8798-347fafcece5d · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens S3: Speech, Script and Scene driven Head and Eye Animation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d82dc12-564a-4c37-b2ed-70bb7f0fd9e2 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Model See Model Do: Speech-Driven Facial Animation with Style Control
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 291209c5-c984-4049-9279-e81a58391cd1 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens VOCAL: V owel and Consonant Layering for Expres- sive Animator-Centric Singing Animation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b52de19-93b2-47e4-8cf0-d8f8c09167df · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Emotalk: Speech-driven emotional disentanglement for 3d face anima- tion
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32263147-47ed-47dd-9f1b-642f2e349f99 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens MeshTalk: 3D Face Animation from Speech using Cross-Modality Disentanglement
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 967994d3-5c40-4bf5-842d-817651597e48 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Facediffuser: Speech-driven 3d facial animation synthesis using diffusion
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90065ea2-fb1a-4410-8037-b2c763e14fe5 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Diffposetalk: Speech-driven stylistic 3d facial animation and head pose generation via diffusion models.ACM Transactions on Graphics (TOG), 43(4):1–9, 2024
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52bde93e-af66-4805-b399-951e31defd30 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Turn-taking and Backchannel Pre- diction with Acoustic and Large Language Model Fusion
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44551fb8-e48a-4d2a-ae1d-92cf84f5b19e · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Mini-Omni2: Towards Open- source GPT-4o with Vision, Speech and Duplex Capabilities
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0104c4ab-a40e-4fce-9807-77c703ecd44d · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens CodeTalker: Speech- Driven 3D Facial Animation with Discrete Motion Prior
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3451b55e-0307-477a-9c11-d0d79c910593 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Qwen3-Omni Technical Report
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f6bdf2e-4e14-4bb1-b8fc-ead13aa08ea3 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens SoundStream: An End- to-End Neural Audio Codec
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81eba8c8-9cd8-481e-9cd2-2950c9081ecf · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens GLM-4-V oice: Towards Intelligent and Human-Like End-to- End Spoken Chatbot
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88007d3d-7bfe-4564-ac77-bf6a81eed46a · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens SpeechGPT: Empow- ering Large Language Models with Intrinsic Cross-Modal Conversational Abilities,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e21ec83-4b02-455c-bbde-4dc4cebada0c · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens MuseTalk: Real-Time High-Fidelity Video Dubbing via Spatio-Temporal Sampling,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bce2bbe5-805f-4ba4-b43a-2e806622f345 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6463b96a-c5fd-4f12-8615-7bbbc3e40e76 · outbound
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens Media2Face: Co-speech Facial Animation Generation With Multi-Modality Guidance
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
No inbound Pith citation observations are available.