Pith. sign in

Paper Citation Record · LEDGER

TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2304.00334.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.00334 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:50.626113Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T22:48:38.173879Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6c502d7a-fa97-41c4-8b13-25c3d36acaa5 · inbound

Exploring Timeline Control for Facial Motion Generation cites this paper.

Exploring Timeline Control for Facial Motion Generation TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:50.626113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:50.626113Z digest=sha256:dd149d1d6ce0fa3c35d88ee1f4bd0ee998a48bf77255d337ffc6755c81ce23d6

Observation 64cca734-c5d4-4cfc-b158-db51f42a451a · inbound

MEDTalk: Multimodal Controlled 3D Facial Animation with Dynamic Emotions by Disentangled Embedding cites this paper.

MEDTalk: Multimodal Controlled 3D Facial Animation with Dynamic Emotions by Disentangled Embedding TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:16:56.996444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:16:56.996444Z digest=sha256:88198cd4a72843e2299f2284042e117107c0575e6be37dc9f54c0cc90216f944

Observation 05d2aa18-834e-4629-82bc-df357b077bd1 · inbound

CEM-Net: Cross-Emotion Memory Network for Emotional Talking Face Generation cites this paper.

CEM-Net: Cross-Emotion Memory Network for Emotional Talking Face Generation TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T19:32:44.266588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T19:32:44.266588Z digest=sha256:be6e6de35fa155d33e80c0a7a6ef018bf45b72da474922742e5fe4f411c709f6

Observation d108b1ca-21c1-43c9-af87-5e2c40af0373 · inbound

EDTalk++: Full Disentanglement for Controllable Talking Head Synthesis cites this paper.

EDTalk++: Full Disentanglement for Controllable Talking Head Synthesis TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-05T19:07:40.503235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:07:40.503235Z digest=sha256:92833861877707d7c344cc2ad8b2aab8fa22e4fea4fc5aeb28223c31c797cf9a

Observation aafaaea3-8f20-4361-a88c-38dd72fe881e · inbound

Think2Sing: Orchestrating Structured Motion Subtitles for Singing-Driven 3D Head Animation cites this paper.

Think2Sing: Orchestrating Structured Motion Subtitles for Singing-Driven 3D Head Animation TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T11:47:50.527329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:47:50.527329Z digest=sha256:743827778af1a5afeecf085b4a5a71e901733ac7bf829497a4107d434afe46c4

Observation 3a40e739-f33a-4a78-9729-aed7e8e10021 · inbound

KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes cites this paper.

KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:48:38.175829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T22:46:51.954117Z digest=sha256:c06ec83993c16159edf20d668670e2577575ca0538656f3939dad139fb1e93d3

Observation 0dd9bd23-9571-421e-a853-2f245cd8f8f2 · inbound

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence cites this paper.

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:36:11.245024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T08:27:41.839123Z digest=sha256:68c288498fd120e3ff914b7fc9b058a7bff028a9335bce88c45d8e22ba6668f2

Observation 73213a83-b957-4825-bc89-a3b3c87e2f69 · inbound

EmoteGPT: 3D Human Facial Expressions from Natural Language Descriptions cites this paper.

EmoteGPT: 3D Human Facial Expressions from Natural Language Descriptions TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T07:49:23.168452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T07:49:23.168452Z digest=sha256:c0c4c5521fe3e6785243757b84da40411adf6dcbb83e6328c350602924fff19f