Pith. sign in

Paper Citation Record · LEDGER

DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2312.09767.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.09767 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T18:23:42.096812Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T21:46:15.567696Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 59433fa3-7ab4-4237-965d-87ed000e5eda · inbound

JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation cites this paper.

JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:03:12.563841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T16:59:57.727035Z digest=sha256:ff2507040eaf46af33c1b1701d6da1449291f6e85e585580b49cc186bdd25600

Observation e4c76726-3be6-4ffe-b96a-2c3f4646ce14 · inbound

Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation cites this paper.

Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:23:15.345266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T17:19:51.411937Z digest=sha256:1fd37c155841cb8d36eaf49c182991406fc04d487fe41c806b11dbd9c993f375

Observation ad011661-eb30-4f38-b868-0fa2e5831a66 · inbound

Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark cites this paper.

Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T20:49:51.506385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T20:49:51.506385Z digest=sha256:e83b6afc894fa1df9da9c81e7e9bdda91e1641902177cd7c5518da9b37c4d661

Observation ac8f5de4-a582-4c54-9e15-4110675763a6 · inbound

Emotion Recognition and Generation: A Comprehensive Review of Face, Speech, and Text Modalities cites this paper.

Emotion Recognition and Generation: A Comprehensive Review of Face, Speech, and Text Modalities DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-09T18:23:42.096812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:23:42.096812Z digest=sha256:533aecd0f1ad807801f2cc9d3adb1b6a2a9d9a4d1c1a3b67e4b1afd39c2ece1e

Observation 5b524bad-e05e-4c5b-82cf-4f52c311a11b · inbound

SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers cites this paper.

SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:07.552280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:07.552280Z digest=sha256:5bd74ff882772aabdb384ce915985e358eb4338a72258d1dc6cd6f479abd574d

Observation eb2b6f1b-38bf-4588-8e9e-3f1e80c56545 · inbound

NTIRE 2025 XGC Quality Assessment Challenge: Methods and Results cites this paper.

NTIRE 2025 XGC Quality Assessment Challenge: Methods and Results DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:20:49.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:20:49.010442Z digest=sha256:721ea0238b66ca09f8db080442662497f017b31af77be9243a7585643c6d58d7

Observation 557aed13-f628-4c58-ba83-519d2bcb644d · inbound

Identity Deepfake Threats to Biometric Authentication Systems: Public and Expert Perspectives cites this paper.

Identity Deepfake Threats to Biometric Authentication Systems: Public and Expert Perspectives DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:53:39.932611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:53:39.932611Z digest=sha256:4ce24d9fa9c7c58f0fc9c21ed2b09ea3ab96a0533e98376c7f50dcedfbf0b199

Observation 95a99982-148d-4c3b-9696-7897732094f0 · inbound

SyncTalk++: High-Fidelity and Efficient Synchronized Talking Heads Synthesis Using Gaussian Splatting cites this paper.

SyncTalk++: High-Fidelity and Efficient Synchronized Talking Heads Synthesis Using Gaussian Splatting DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:43.559742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:43.559742Z digest=sha256:c32529413576ea4ae5a54287ea00401955aca246142c49d4764dbf27047d801f

Observation 914f27cf-23b6-424b-8de6-87d67fa033e7 · inbound

MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation cites this paper.

MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T22:15:08.419865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:15:08.419865Z digest=sha256:242e13b376cf9f291f1432af38923b8246ffa80975e6b037fc877a3e809d8432

Observation f26d6c43-0e8a-4147-8eca-c43287443d6f · inbound

JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching cites this paper.

JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:50:51.035807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T00:46:39.196042Z digest=sha256:30fb9fda411d10b68d0b4005aa20bcbf2b74a8aee5c5fbcd7727828fbd6a3d31

Observation 3d446572-f400-4b64-92f0-d6db6a9cd57a · inbound

MoDiT: Learning Highly Consistent 3D Motion Coefficients with Diffusion Transformer for Talking Head Generation cites this paper.

MoDiT: Learning Highly Consistent 3D Motion Coefficients with Diffusion Transformer for Talking Head Generation DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:38:20.504641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:38:20.504641Z digest=sha256:9cdbc93214aab33e8ca2dc6d2bdbb188dbb2f9d15a910327e5cc9a8e05318750

Observation 288b8791-b9ae-48dd-b171-d59d1f9503f7 · inbound

Mask-Free Audio-driven Talking Face Generation for Enhanced Visual Quality and Identity Preservation cites this paper.

Mask-Free Audio-driven Talking Face Generation for Enhanced Visual Quality and Identity Preservation DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T13:10:19.601406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:10:19.601406Z digest=sha256:b66de12f7864d307854eef2e58d6921fc7803428aa921f3b4bdfb0e039c81a53

Observation f647508e-8349-4959-be72-5874be2655b4 · inbound

Who is a Better Talker: Subjective and Objective Quality Assessment for AI-Generated Talking Heads cites this paper.

Who is a Better Talker: Subjective and Objective Quality Assessment for AI-Generated Talking Heads DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:19.020094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:19.020094Z digest=sha256:06bae44618f9506fbf6112b1c77c0ec1aa85c86850a7cb1dc939056dca461620

Observation e3cc0341-9468-470c-8ff9-aba2f48ad25d · inbound

FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation cites this paper.

FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T20:06:48.133326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:06:48.133326Z digest=sha256:e429f46624125b380c51a1b0446929e226e76c8b049c0935f77e9ed2e9a5160d

Observation 0f2707a4-8eb6-4427-833a-5c65644253e0 · inbound

Human Motion Video Generation: A Survey cites this paper.

Human Motion Video Generation: A Survey DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 158

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:57.394438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:57.394438Z digest=sha256:d83fc07305966b9071771c6a84b35812e58b9aa6f2f8e92c3fe982e6aa6b9eab

Observation 18ba8698-1e52-45ca-8930-12de44413092 · inbound

AV-EMO-Reasoning: Benchmarking Emotional Reasoning Capabilities in Omni-modal LLMS with Audio-visual Cues cites this paper.

AV-EMO-Reasoning: Benchmarking Emotional Reasoning Capabilities in Omni-modal LLMS with Audio-visual Cues DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T11:06:48.586938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:06:48.586938Z digest=sha256:d26f962b8d4694fb8fa549f5495b250909bcfde076180767a21a3ecbe0f91def

Observation 91e59207-8b3d-409f-9592-f19c63a335b9 · inbound

THEval. Evaluation Framework for Talking Head Video Generation cites this paper.

THEval. Evaluation Framework for Talking Head Video Generation DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-21T19:04:19.299013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T19:00:48.999456Z digest=sha256:4f1122e79b9e759bd08aa3c98f60a0e2e241161e59bed798965514679a56e9b4

Observation 9cc16b93-bc6a-42a8-9538-a897b52e120c · inbound

Beyond Monologue: Interactive Talking-Listening Avatar Generation with Conversational Audio Context-Aware Kernels cites this paper.

Beyond Monologue: Interactive Talking-Listening Avatar Generation with Conversational Audio Context-Aware Kernels DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:56:01.777015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:16:53.603971Z digest=sha256:0ab1d9f7411585af3f40aad0635d7384ef08030c6446a0e117a57a51d4beea52

Observation 0a37b97b-61f8-4dc0-b117-95322f754809 · inbound

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence cites this paper.

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:36:11.220707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T08:27:41.839123Z digest=sha256:b0e4baf59d6862b926cbd20b85f31b687faa814524e0dfdce2f96fabebaece98

Observation 008e79f9-6dda-4715-b264-bc208d659eeb · inbound

Hallo-Live: Real-Time Streaming Joint Audio-Video Avatar Generation with Asynchronous Dual-Stream and Human-Centric Preference Distillation cites this paper.

Hallo-Live: Real-Time Streaming Joint Audio-Video Avatar Generation with Asynchronous Dual-Stream and Human-Centric Preference Distillation DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:06:12.869767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T06:56:19.795651Z digest=sha256:946038ca21a7c31d9dd863fd81516700fe3e438ae995e3e6ec7ba95538281f27

Observation 2e44f674-51f3-4aea-8360-9a8502e88f09 · inbound

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation cites this paper.

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:44:01.740133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T22:36:53.138354Z digest=sha256:666b403673df395b4152193599b22e7ffa27d530f2a7b84039fd89c3a23562be

Observation 7c599578-2d61-45ab-8844-5890e33e7856 · inbound

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation cites this paper.

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T21:46:15.569027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T16:19:55.048831Z digest=sha256:b99c40c1e6ec7ce6c1a92d6abed7013818a96b5cad4e4570d9f5a14f2ebba0b6