Pith. sign in

Paper Citation Record · LEDGER

FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2504.04842.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.04842 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:02:29.621663Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T04:47:38.002589Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 23562447-b0cf-4b97-8e8b-76addc290927 · inbound

HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters cites this paper.

HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:02:29.621663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:02:29.621663Z digest=sha256:09911f9cd5a6c5fd2fc3a57c2beb701013dd240f1667dff35a4023b830992815

Observation 259b03b1-ff70-43b9-b91d-4b75e3601cab · inbound

Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation cites this paper.

Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:07:32.607328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:07:32.607328Z digest=sha256:0ecce95a76ee9b4db531637c18791c6af5d7b5eb81a0120b49f05a1970d0a843

Observation 005dad01-d72c-4c8f-a3ec-0c0f2d47335a · inbound

SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers cites this paper.

SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.579080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.579080Z digest=sha256:965afb7a7240b3601f6c8b9124047504df9aa31aafdb4dfb6eabab9af3ebb1e3

Observation ab4b4deb-7c2c-47d2-8177-452149814295 · inbound

AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation cites this paper.

AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:55:28.875127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:55:28.875127Z digest=sha256:bb7d309bbe461b011962222c09563e136e68093f5098d9ab412b04836028390f

Observation 736c1bca-251c-4914-b188-6a0b6334b529 · inbound

SpeakerVid-5M: A Large-Scale High-Quality Dataset for Audio-Visual Dyadic Interactive Human Generation cites this paper.

SpeakerVid-5M: A Large-Scale High-Quality Dataset for Audio-Visual Dyadic Interactive Human Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T17:49:46.738652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:49:46.738652Z digest=sha256:6024303dad417f8891bc703ff1adac8a8dd62fca267e30b839bd3572a69a138c

Observation 114aa39a-d545-40bd-b713-bfd6ca3b771f · inbound

FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers cites this paper.

FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:38.613422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:39:38.613422Z digest=sha256:58bdc2dfb7df9e78963e19241b6ef1f8a8655aed263081d82bd18b808d41a2c3

Observation bcaefb3f-3e96-453b-9af3-2d8aeef63864 · inbound

FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation cites this paper.

FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T20:06:49.148058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:06:49.148058Z digest=sha256:9a66c5ab55832594990c8c43fd75f84c86e05ab8842506431f105f75529f2922

Observation ce572680-1171-40d8-af1b-acf777ebf76e · inbound

InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing cites this paper.

InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T18:50:16.285909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:50:16.285909Z digest=sha256:8bd3f13a3de3317ed919bccf343baf78d0dcfac4ff387f6530ad71f49e3ac92b

Observation 54d8d8a5-a960-4da1-b302-d210086694cd · inbound

Wan-S2V: Audio-Driven Cinematic Video Generation cites this paper.

Wan-S2V: Audio-Driven Cinematic Video Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T16:24:59.254749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:24:59.254749Z digest=sha256:26332e4236c429cfd9cbb3f2d6bb74e1b38ab898d62573c87c1bb02ea4822e48

Observation 5cf5dd96-bd4c-4fee-82a1-fe661de31133 · inbound

InfinityHuman: Towards Long-Term Audio-Driven Human cites this paper.

InfinityHuman: Towards Long-Term Audio-Driven Human FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.239985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.239985Z digest=sha256:d3f20c6a2641254001f5fd065ba9559be83516ec3758f0ecf57b70d65b4a508b

Observation 20698b96-3e2e-44bb-8427-d1bfde1cc60e · inbound

UniVerse-1: Unified Audio-Video Generation via Stitching of Experts cites this paper.

UniVerse-1: Unified Audio-Video Generation via Stitching of Experts FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T00:05:47.032525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:05:47.032525Z digest=sha256:f54c360905b5463e38c161a1a61c04a274b3a3ec0857514126e6f90a4c06a00b

Observation 1af6043e-009d-42fc-9d9a-2fc49a1ef10b · inbound

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation cites this paper.

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:20:22.475858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T22:20:16.320171Z digest=sha256:e828956d585bb6f39f54c28fb92eb3100342c4c52e994e22a223bfddc4a1bda7

Observation 7eb69972-41b0-4760-a2e1-582e119aa18a · inbound

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation cites this paper.

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:41:25.647082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T17:30:36.462567Z digest=sha256:4e88d71db4b948660ebd0340c03fdda037c6e10b501568b39bf1e32c68885b01

Observation b7e51fbd-0484-4d2f-91ea-84d1b8cd8a89 · inbound

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation cites this paper.

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T16:38:16.281646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:38:16.281646Z digest=sha256:1d18c23098d7ef4cdd7b6e43985af993178e2cc7cda1522e43753b8080b30172

Observation 84f34697-78db-4053-be6a-b312f29cfdf7 · inbound

IP-Adapter Is All You Need: Towards Fine-Tuning-Free Diffusion-Based Talking Face Generation cites this paper.

IP-Adapter Is All You Need: Towards Fine-Tuning-Free Diffusion-Based Talking Face Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.382287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T07:50:56.947671Z digest=sha256:a31efd4afebb77b6bd53783f53d94a9d36607417c191526292e89018f937a710

Observation e4248c7d-043e-479c-8756-e5a9f7725f6f · inbound

Lip Forcing: Few-Step Autoregressive Diffusion for Real-time Lip Synchronization cites this paper.

Lip Forcing: Few-Step Autoregressive Diffusion for Real-time Lip Synchronization FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:47:38.004099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T13:40:14.506288Z digest=sha256:afd8014acd7b6ffdc89a2eaf86058d4c42eeee2c5d6721fa440626086c34d8ef

Observation ba0a8696-04c2-4358-970a-c6fa7f5bc396 · inbound

LeapTalk: Breaking the Latency-Quality Trade-off in Talking Head Generation cites this paper.

LeapTalk: Breaking the Latency-Quality Trade-off in Talking Head Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 125

Resolution
unresolved
no resolver link, observed 2026-08-04T01:32:13.185561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T01:32:13.185561Z digest=sha256:b0120a70be70d7f80fcc8ff34f44ecfe72cabd2515c363640385bf885358e5d9