Pith. sign in

Paper Citation Record · LEDGER

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation

As of 10 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2501.14646.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.14646 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:01:44.820555Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T05:45:39.198180Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T05:45:40.523888Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy26
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 82b61ee9-5571-4c77-98e1-1b32ebcd8420 · outbound

This paper cites wav2vec 2.0: a framework for self-supervised learning of speech repre- sentations.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation wav2vec 2.0: a framework for self-supervised learning of speech repre- sentations

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.485684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.638991Z digest=sha256:d8877923b544ba558f585f69f64ee41d3551ab74422c304e875e8fb486df45db

Observation 1b3d4ef7-995e-4f8e-8525-88cd9ff7ccd6 · outbound

This paper cites Ad-nerf: Audio driven neural radiance fields for talking head syn- thesis.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Ad-nerf: Audio driven neural radiance fields for talking head syn- thesis

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.392941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.672915Z digest=sha256:104a5d4571c960b2d1b6d1ea0ab481593490aa310aea1aa023b3d4d131aefd18

Observation 0b244bcd-9a75-47ec-883b-1d18e39bc291 · outbound

This paper cites Deep Speech: Scaling up end-to-end speech recognition.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Deep Speech: Scaling up end-to-end speech recognition

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T15:01:44.684552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:01:44.684552Z digest=sha256:08fcfaebbe2425a01f8a8f1b1105df2619e73844fee710212dd39014047347e7

Observation 8876f7d2-4eea-4251-bc43-06ab47a38873 · outbound

This paper cites Animate anyone: Consistent and control- lable image-to-video synthesis for character animation.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Animate anyone: Consistent and control- lable image-to-video synthesis for character animation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.293180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.709668Z digest=sha256:7f9b2401b43b3861fb4ce1b18a75127b61a5461b47c6708233b11773ea47cd90

Observation ddb2da2e-eb9b-4437-8a1c-7e39c7044af1 · outbound

This paper cites NeRFFaceSpeech: One-shot Audio-driven 3D Talking Head Synthesis via Generative Prior.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation NeRFFaceSpeech: One-shot Audio-driven 3D Talking Head Synthesis via Generative Prior

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T15:01:44.714560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:01:44.714560Z digest=sha256:00f4d5cad10eb706f01ae11e21de92cd75105fe6bd78c1c0e087f508828bbd9a

Observation ff18da6a-5a5c-4ec5-abcb-f4807646c074 · outbound

This paper cites Efficient region-aware neural radiance fields for high-fidelity talking portrait synthesis.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Efficient region-aware neural radiance fields for high-fidelity talking portrait synthesis

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.273157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.720266Z digest=sha256:6343b596c1fae30e3815aacdfc98277a3fd9ecabddacaf1bfe487bb15ce0a2a0

Observation 72423fe8-90d8-4319-8b47-781059443ec7 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view synthesis.Communications of the ACM, 65(1):99–106,.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Nerf: Representing scenes as neural radiance fields for view synthesis.Communications of the ACM, 65(1):99–106,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.250210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.726170Z digest=sha256:21f161a57f5c1c8175948026809d054259907ca3475223d0e8b9eb5205d9b795

Observation 6f8c8240-6915-4b56-9733-ab6e13cd68e2 · outbound

This paper cites Instant neural graphics primitives with a multiresolution hash encoding.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Instant neural graphics primitives with a multiresolution hash encoding

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.225761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.731357Z digest=sha256:47385a5e15b540dc527734e13cc6eedb494cd350cbcc1ed0de9d2c4ce6d7a8dd

Observation d64b36be-c4a5-46fc-9a96-91d1df3a5e29 · outbound

This paper cites Emotalk: Speech-driven emotional disentanglement for 3d face animation.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Emotalk: Speech-driven emotional disentanglement for 3d face animation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.208262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.736443Z digest=sha256:f43031dc9cd69bec8be7078b3c6971ff482fe77fc90ff9c49e1dd5bc5992d4ae

Observation 290ee58b-ddaa-4025-a0f1-55e6059dda3e · outbound

This paper cites Synctalk: The devil is in the syn- chronization for talking head synthesis.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Synctalk: The devil is in the syn- chronization for talking head synthesis

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.192312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.742475Z digest=sha256:d2b63671e617d47b03e0f54a429ecf8ea9679de56b6a901f30ccf987750b0728

Observation 9e4a2983-4531-40fe-8a7b-bd76d5ef4b0b · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation A lip sync expert is all you need for speech to lip generation in the wild

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.176659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.748060Z digest=sha256:c9206c88524a20c2d1a6bf07e4869fe2e307accbe405a26429b5a3a1042bdfec

Observation 37ffcea7-b7f6-4869-a3d7-7bb190c688a1 · outbound

This paper cites Speech drives templates: Co- speech gesture synthesis with learned templates.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Speech drives templates: Co- speech gesture synthesis with learned templates

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.158371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.752979Z digest=sha256:a5064f072cb5d31dc6687a334082713eb02f9f1508b38e465b2aeee57a31a2b2

Observation b3aea1e3-8801-4a53-b2a3-89dbf2192698 · outbound

This paper cites Fine-grained head pose estimation without key- points.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Fine-grained head pose estimation without key- points

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.138626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.757810Z digest=sha256:8fb1b1460c386bf50bae3eab165889dbddcb6b3d1ff279cc0d8d1267abc1c54b

Observation 125058ef-fae1-453f-8c51-c2a9a673b345 · outbound

This paper cites Eye blink detection using facial landmarks.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Eye blink detection using facial landmarks

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.119630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.762821Z digest=sha256:870a485933843faa5d0d24c7ea28806d12d66beead8d906c2c71e8e8160f1f74

Observation d9b0d04b-7c01-4334-8c23-4f5518a416f2 · outbound

This paper cites Image quality assessment: from error visibility to structural similarity.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Image quality assessment: from error visibility to structural similarity

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.056611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.776588Z digest=sha256:6abb340e760c537de5f147a64a85a79f713305d2fb2676d008448579d5fcac99

Observation b310905e-4a93-4c1a-9e19-0c630308f5a0 · outbound

This paper cites X-portrait: Expressive portrait animation with hierarchical motion attention.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation X-portrait: Expressive portrait animation with hierarchical motion attention

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.038257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.786737Z digest=sha256:bb12be4c210d2d61bb6855498abaa76ec3c79ed659f294e98df5646341cb7441

Observation 27476aa3-54fc-4eb9-a655-33d037732842 · outbound

This paper cites Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T15:01:44.791849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:01:44.791849Z digest=sha256:e51ea118795bb1c1fb7bc480e763cf6210aa4f44fc3c06aff2937eea6f13f07e

Observation b61ada0f-d010-4844-84ce-013b07cccbff · outbound

This paper cites GeneFace++: Generalized and Stable Real-Time Audio-Driven 3D Talking Face Generation.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation GeneFace++: Generalized and Stable Real-Time Audio-Driven 3D Talking Face Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T15:01:44.797529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:01:44.797529Z digest=sha256:07fa016b58e2fbd763ca044f75afbdcfbe2bcf8464acb4e6b407b40c085467d0

Observation 213d709c-08ba-4058-9605-46dfddb5b52f · outbound

This paper cites GeneFace: Generalized and High-Fidelity Audio-Driven 3D Talking Face Synthesis.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation GeneFace: Generalized and High-Fidelity Audio-Driven 3D Talking Face Synthesis

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T15:01:44.804369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:01:44.804369Z digest=sha256:623e3d625c6c343a92bfaccbab3c6c52e1fa0ee3c4f4ba8c9728b7c138ffa5e9

Observation c23a30c4-394c-481c-ba03-a7ce1a0594ab · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation The unreasonable effectiveness of deep features as a perceptual metric

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T15:01:44.810166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:01:44.810166Z digest=sha256:279fd221921fed5e25a54a8a73d9bac104379261f658fb9a5fca0e00de283174

Observation 60e6a430-d338-465c-af20-4ce76a87d647 · outbound

This paper cites Flow-guided one-shot talking face gen- eration with a high-resolution audio-visual dataset.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Flow-guided one-shot talking face gen- eration with a high-resolution audio-visual dataset

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.008320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.815835Z digest=sha256:bf1ec024e64dabd37cfbd004f8be278656517aed4fcac7de83581622dade4df4

Observation 7ed10855-80bd-4cd9-92aa-03286d07c420 · outbound

This paper cites Identity-preserving talking face generation with landmark and appearance priors.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Identity-preserving talking face generation with landmark and appearance priors

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:44.989395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.820555Z digest=sha256:a041e869959f20d61eb7272ca3e661ac60ad7b1d555f26fe46e6c0aaff272ec8

Observation e2336307-1481-4d90-b181-f7a4f5e6671e · outbound

This paper cites V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 2004

Resolution
unresolved
no resolver link, observed 2026-08-10T15:01:44.781180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:01:44.781180Z digest=sha256:010b91092d201c3750ede9df58db5bf1d796ba02b921868e4b12abb81ebf3d69

Observation 675e242e-86da-4b1a-b91e-ecc917e9c119 · outbound

This paper cites Hubert: Self- supervised speech representation learning by masked pre- diction of hidden units.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Hubert: Self- supervised speech representation learning by masked pre- diction of hidden units

Reference 2010

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.309724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.704545Z digest=sha256:6612e8cce8dd59c010b7374faad05d94dd9ad9ce710502e0f7293379b8751d1b

Observation 9d103422-4f22-4743-b736-bcc7c0be6c93 · outbound

This paper cites Facexhubert: Text-less speech-driven e(x)pressive 3d facial animation synthesis using self- supervised speech representation learning.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Facexhubert: Text-less speech-driven e(x)pressive 3d facial animation synthesis using self- supervised speech representation learning

Reference 2014

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.357344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.689254Z digest=sha256:bc783d459027d83f9fe0a4a2dbb5b4a6661db438a8d784e6e6bdd6eb1a4cef89

Observation 9fcfa7e7-e062-4aa7-8f70-3a1168c98366 · outbound

This paper cites Lip movements gener- ation at a glance.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Lip movements gener- ation at a glance

Reference 2015

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.451347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.649994Z digest=sha256:cfea1f0e864f6e83727f9b3cc106706ce3ffba9f723362d0dbd793171b2d5e52

Observation c077df2f-fd86-42c8-8c0b-625042188144 · outbound

This paper cites Edtalk: Efficient disentanglement for emotional talking head synthesis.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Edtalk: Efficient disentanglement for emotional talking head synthesis

Reference 2016

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.099215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.767512Z digest=sha256:1154a12af5f9a4ae5738ef634775ffb8bfd3cb793839eb2aae54e1319d2d1772

Observation 9bb256ad-a0a2-44fb-821f-663b99f3e5b8 · outbound

This paper cites Im- age quality metrics: Psnr vs.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Im- age quality metrics: Psnr vs

Reference 2017

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.327269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.699486Z digest=sha256:a9d769a412cba2e5d3710a0e49a940dde3a1c84062572d38425c15ef3bbf921d

Observation 6af8f5a6-d25d-4da8-9f7c-5c8098542f5a · outbound

This paper cites EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-10T15:01:44.655704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:01:44.655704Z digest=sha256:084eb6644144b6ddee100996172d0123b10bc714cc61d65dfceb8713fcb53850

Observation 492d0417-9073-4e53-8eee-cc268f11baa7 · outbound

This paper cites [Baltruˇsaitis et al., 2015] Tadas Baltru ˇsaitis, Marwa Mah- moud, and Peter Robinson.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation [Baltruˇsaitis et al., 2015] Tadas Baltru ˇsaitis, Marwa Mah- moud, and Peter Robinson

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.467954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.644755Z digest=sha256:b950888c8f0d0ac28e0f2529a947eb0424929ad93df66ed6cdda4017868547f4

Observation 3b5fc4b9-a62b-44df-99c4-4ce6b3e73516 · outbound

This paper cites I2v-adapter: A general image-to-video adapter for diffusion models.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation I2v-adapter: A general image-to-video adapter for diffusion models

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.374176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.679871Z digest=sha256:70ef1f0c025d84547f4891e41d91fee0bb55a15ce6c072165c8a11f06e7f851e

Observation 4c55d2a0-0de3-4f49-a801-0831b36a721a · outbound

This paper cites Shortcut learn- ing in deep neural networks.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Shortcut learn- ing in deep neural networks

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.409990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.666146Z digest=sha256:fa03df389f4dbbdf3835b5ed4de42d05d042c5c24060a1f1b798b2b7753c6933

Observation 37cfb4ae-c34a-4bd8-9dc4-76443abc1c59 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T15:01:44.694568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:01:44.694568Z digest=sha256:b772d1e1e0a4e4b71480245a01efcf2b312aa07d3c7359f133c327704d18dedc

Observation a9489c02-1ea7-4193-af32-f2ca6f3356e1 · outbound

This paper cites Videoretalking: Audio- based lip synchronization for talking head video editing in the wild.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Videoretalking: Audio- based lip synchronization for talking head video editing in the wild

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.430957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.661034Z digest=sha256:5020912dae37cdddcc1a1872f151f754ecdca2123ab4aa00e7c867b9e96556f0

Observation 902f7944-0724-4932-90a6-8ea927816a17 · outbound

This paper cites Emo: Emote portrait alive generating expres- sive portrait videos with audio2video diffusion model un- der weak conditions.

SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation Emo: Emote portrait alive generating expres- sive portrait videos with audio2video diffusion model un- der weak conditions

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:01:45.077352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:01:44.772269Z digest=sha256:ec0b933bbb7cd078264fe8bddb0e56586eb99a2616cd2c91bb75cf9c5a4aede9

Pith citing papers

Observation 8070059e-f4b3-43ea-9760-6e0722ca8919 · inbound

MoGaFace: Momentum-Guided and Texture-Aware Gaussian Avatars for Consistent Facial Geometry cites this paper.

MoGaFace: Momentum-Guided and Texture-Aware Gaussian Avatars for Consistent Facial Geometry SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-06T05:45:40.529287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T05:45:39.198180Z digest=sha256:bc03d8a890c31266cb22f8ed9ec7feb2829f6700ae4be9cdcbd8641ad2ee900c