Pith. sign in

Paper Citation Record · LEDGER

FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2504.04842.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.04842 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:02:29.621663Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T04:47:38.002589Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 23562447-b0cf-4b97-8e8b-76addc290927 · inbound

HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters cites this paper.

HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:02:29.621663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:02:29.621663Z digest=sha256:167049a1d959a959d078a2cbfe37fa5f579fdd5dcaeb7c616fa50738c49c276d

Observation 259b03b1-ff70-43b9-b91d-4b75e3601cab · inbound

Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation cites this paper.

Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:07:32.607328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:07:32.607328Z digest=sha256:9e1fe5965cfb7c5d605235e85966ac1ca38852b49f0a30d70b302ebf1cffa46f

Observation 005dad01-d72c-4c8f-a3ec-0c0f2d47335a · inbound

SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers cites this paper.

SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.579080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.579080Z digest=sha256:310099485e464bfc6d245f6fabc00254ddd3d8f6e6818a655213f58a1b0f75a2

Observation ab4b4deb-7c2c-47d2-8177-452149814295 · inbound

AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation cites this paper.

AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:55:28.875127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:55:28.875127Z digest=sha256:ba8496855bc764e054ddc8a0b851b106467300e70df11b89a32d385132483eb8

Observation 736c1bca-251c-4914-b188-6a0b6334b529 · inbound

SpeakerVid-5M: A Large-Scale High-Quality Dataset for Audio-Visual Dyadic Interactive Human Generation cites this paper.

SpeakerVid-5M: A Large-Scale High-Quality Dataset for Audio-Visual Dyadic Interactive Human Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T17:49:46.738652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:49:46.738652Z digest=sha256:02841e405a71c932c6246c56cb9f75f6d81dec92849e7f71b9948fc9cd7f0c7c

Observation 114aa39a-d545-40bd-b713-bfd6ca3b771f · inbound

FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers cites this paper.

FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:38.613422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:39:38.613422Z digest=sha256:3271bf2529af7b1bbae6c7fe305d0a69882f3beb8dd950bde5b09a75303a4398

Observation bcaefb3f-3e96-453b-9af3-2d8aeef63864 · inbound

FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation cites this paper.

FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T20:06:49.148058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:06:49.148058Z digest=sha256:00e6466ffa717af164664069544b285ebd5d7a74fba9b223999c8902f1b49618

Observation ce572680-1171-40d8-af1b-acf777ebf76e · inbound

InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing cites this paper.

InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T18:50:16.285909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:50:16.285909Z digest=sha256:77d410a4d39c914c8854e46d5e73b490fb6e566ec2c4c0f276eebf85172b7508

Observation 54d8d8a5-a960-4da1-b302-d210086694cd · inbound

Wan-S2V: Audio-Driven Cinematic Video Generation cites this paper.

Wan-S2V: Audio-Driven Cinematic Video Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T16:24:59.254749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:24:59.254749Z digest=sha256:cb52914a8996841fb456dc9c1b50d58d1d7b032630663ca3c331742ef3ddde04

Observation 5cf5dd96-bd4c-4fee-82a1-fe661de31133 · inbound

InfinityHuman: Towards Long-Term Audio-Driven Human cites this paper.

InfinityHuman: Towards Long-Term Audio-Driven Human FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.239985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.239985Z digest=sha256:fa1ad1a6f61061bfd5927fba84f9fc8c49354a2a592ecd4849c53a60b9fc1727

Observation 20698b96-3e2e-44bb-8427-d1bfde1cc60e · inbound

UniVerse-1: Unified Audio-Video Generation via Stitching of Experts cites this paper.

UniVerse-1: Unified Audio-Video Generation via Stitching of Experts FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T00:05:47.032525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:05:47.032525Z digest=sha256:54233fb2352dedaa29e2ac6c0e3348eecfd46098740f20de8b82ff3075f74c51

Observation 1af6043e-009d-42fc-9d9a-2fc49a1ef10b · inbound

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation cites this paper.

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:20:22.475858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T22:20:16.320171Z digest=sha256:e4318663614b5c89f545c53466cddf525ed545ee7f95d9cc16d65b293e5666e7

Observation 7eb69972-41b0-4760-a2e1-582e119aa18a · inbound

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation cites this paper.

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:41:25.647082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:30:36.462567Z digest=sha256:97e9a6c42cfaa247432545f610d2b072e250cf1a6c766d64ce7d290f892a95a2

Observation b7e51fbd-0484-4d2f-91ea-84d1b8cd8a89 · inbound

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation cites this paper.

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T16:38:16.281646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:38:16.281646Z digest=sha256:2e7a4718a0c4eeafba8e87f66db764b68d09938e7909bc1ddb6b2f403eb0b3e6

Observation 84f34697-78db-4053-be6a-b312f29cfdf7 · inbound

IP-Adapter Is All You Need: Towards Fine-Tuning-Free Diffusion-Based Talking Face Generation cites this paper.

IP-Adapter Is All You Need: Towards Fine-Tuning-Free Diffusion-Based Talking Face Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.382287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T07:50:56.947671Z digest=sha256:aed1a3bbcf33b6872d5131f8305d33c012412576c77017238e7f9ce9e20a516c

Observation e4248c7d-043e-479c-8756-e5a9f7725f6f · inbound

Lip Forcing: Few-Step Autoregressive Diffusion for Real-time Lip Synchronization cites this paper.

Lip Forcing: Few-Step Autoregressive Diffusion for Real-time Lip Synchronization FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:47:38.004099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T13:40:14.506288Z digest=sha256:52e861a52a43802130ecdaae7db466401982d99fdad05889ca3c918d0440fb72

Observation ba0a8696-04c2-4358-970a-c6fa7f5bc396 · inbound

LeapTalk: Breaking the Latency-Quality Trade-off in Talking Head Generation cites this paper.

LeapTalk: Breaking the Latency-Quality Trade-off in Talking Head Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 125

Resolution
unresolved
no resolver link, observed 2026-08-04T01:32:13.185561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T01:32:13.185561Z digest=sha256:c6b94a5c217034c3246f4e7392e738fcdede2bca61a54ca43516f277add6a8da