Pith. sign in

Paper Citation Record · LEDGER

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation

As of 10 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2509.20128.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.20128 v2

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-18T14:18:02.076576Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T14:18:02.076576Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-05-18T14:21:28.416903Z

Reference resolution

31 of 31 outbound references displayed

  • verified exact1
  • verified fuzzy25
  • unresolved3
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eb3ed498-689b-41fc-b7e1-efa123530228 · outbound

This paper cites KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T14:21:28.419042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:44db7a3d97bcd4bec71c9afcd01ab8edee943b47846d6fdce48fe778207c8a02

Observation 524236a2-c7a1-4ce4-893b-6c9f7f13bd94 · outbound

This paper cites an unresolved cited work.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-05-18T14:22:40.870577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:1149ab569b01a3a61b73091e4b5316784dd81ff58cd48d3f48867c8331522999

Observation 14298c15-b4c2-4b14-ae3c-182f77e4ca19 · outbound

This paper cites Dataset We train and evaluate our model on two benchmarks.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Dataset We train and evaluate our model on two benchmarks

Reference 3

Resolution
malformed identifier
raw_fallback, observed 2026-05-18T14:22:40.862010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:f1fa355cc4b72bb2321381f317a7fa1e7067069e1149bdfdb6ad8e665965e953

Observation d6425d6b-ceee-4d6f-8f3c-03fd823e1784 · outbound

This paper cites an unresolved cited work.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-05-18T14:22:40.865213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:1104b0ad20dc02bb87c1342bb2c2d14cf8713ee43f150ac5506edd16ee9735d7

Observation 1c9ed7f2-7038-438f-9ee1-122cb4a34e54 · outbound

This paper cites an unresolved cited work.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-05-18T14:22:40.823338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:f860f8fb433e81f3ba932bf8404fee5cf5d9874bda15f466c0fc864a8a9fbe71

Observation fa464dfa-eaa9-437d-adbc-9de5b15d5c7f · outbound

This paper cites Facediffuser: Speech- driven 3d facial animation synthesis using diffusion.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Facediffuser: Speech- driven 3d facial animation synthesis using diffusion

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.837800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:f6fe2394b5c735c321a2871d43aeb23cf2e0298742907680c943f4b87e696ac1

Observation 1223e6a9-01d4-4471-8fd8-0e2ba71ed2f5 · outbound

This paper cites Difftalk: Crafting diffusion models for generalized audio- driven portraits animation.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Difftalk: Crafting diffusion models for generalized audio- driven portraits animation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.809585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:535c23f12826b3c1516c236fddf9953f0ec2687f84d5290d6c09f0c06ede639a

Observation 56562c9d-a4e3-488b-9d83-2ceeb5410941 · outbound

This paper cites DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:21:28.415049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:add026c6dc58ecfea9d6f87465812b91aee94610387af24bf260870a8e9816a5

Observation 7d31c9a5-3cca-4099-9c66-8b439395e30b · outbound

This paper cites Emotivetalk: Ex- pressive talking head generation through audio information de- coupling and emotional video diffusion.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Emotivetalk: Ex- pressive talking head generation through audio information de- coupling and emotional video diffusion

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.805335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:821062ec946ebe79ce270f5fef2f6c3b9b5c29d805cb52cbb73e5a9005d20f8d

Observation eefb76c7-2aeb-43b9-8a5b-05ef2b929d1f · outbound

This paper cites Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.799341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:bb054d793c187b0a8986764e51b2cfc8f8fdab5e7978ad4b2fc8cc5719c50b49

Observation 4c23ca70-4119-46f6-b3f5-e1dbc91596bf · outbound

This paper cites Synctalk: The devil is in the synchro- nization for talking head synthesis.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Synctalk: The devil is in the synchro- nization for talking head synthesis

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.924511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:387b1923b05b1136802541a966f08444212adf18bf658c3e49ccec13678a5c2d

Observation 4075752b-5b85-4d09-bc8e-aba6af630474 · outbound

This paper cites Prosodytalker: 3d visual speech animation via prosody de- composition.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Prosodytalker: 3d visual speech animation via prosody de- composition

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.920557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:17b305ad2c62e0bcc7ae8dc7620acbc223ee5a79eb2709a3393ac7dbc51916d8

Observation 41225913-770f-4596-b402-a9bcc0d97999 · outbound

This paper cites Keyface: Expressive audio-driven facial animation for long sequences via keyframe interpolation.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Keyface: Expressive audio-driven facial animation for long sequences via keyframe interpolation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.916558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:cf39cbbc3feccb3c5e85eacd53988f80c15e383a5ef5903978caddafb415a43d

Observation a7a3b5e2-9c9c-4751-a504-daf40d7dae21 · outbound

This paper cites Speak: Speech-driven pose and emotion- adjustable talking head generation.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Speak: Speech-driven pose and emotion- adjustable talking head generation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.913070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:09553e38a61291d6859afe0c7a3dd15f6219855b6b0e8dd74d2e197c80e4324b

Observation 29bd2907-6bf2-4142-8af7-a179eb30449d · outbound

This paper cites Fd2talk: Towards gener- alized talking head generation with facial decoupled diffusion model.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Fd2talk: Towards gener- alized talking head generation with facial decoupled diffusion model

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.909664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:0380607d4466d73912ba92cb726aefd2bb49165f8fa00dcf952e5da5a9797121

Observation e66ee7f3-c55a-4a14-8053-eb40964e37bd · outbound

This paper cites Learning an animatable detailed 3d face model from in-the-wild images.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Learning an animatable detailed 3d face model from in-the-wild images

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.905080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:f2c7a0ccce368d2e437abbff548c5a6bd12625222bd94fcce5a88383b728ecf9

Observation da8c9033-b841-4664-b959-90a603d48439 · outbound

This paper cites Disco- head: audio-and-video-driven talking head generation by dis- entangled control of head pose and facial expressions.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Disco- head: audio-and-video-driven talking head generation by dis- entangled control of head pose and facial expressions

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.900229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:d3f487e8ccc3facf0ec2998229be95e712b596e67f23c1a3c3877954fb854bb4

Observation 8f62fb29-327d-45b3-98ca-ec67218cc8bd · outbound

This paper cites Nerf-3dtalker: Neural radiance field with 3d prior aided audio disentanglement for talking head syn- thesis.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Nerf-3dtalker: Neural radiance field with 3d prior aided audio disentanglement for talking head syn- thesis

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.896595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:91a8d7ce77dfc15b623e945159fea2bf896a921603ffa5ef1d6f767dfcbdcba8

Observation 8ab9707f-82e6-449a-82c1-25603f0328b2 · outbound

This paper cites wav2vec: Unsupervised pre-training for speech recognition.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation wav2vec: Unsupervised pre-training for speech recognition

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.892830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:5a435a924fcdad2d886ad916a7eb4193a037e83e6dc3f44e068bf844d162be21

Observation dc6a7aa7-8566-4c2b-9638-1c258428a924 · outbound

This paper cites Spsinger: Multi-singer singing voice synthesis with short reference prompt.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Spsinger: Multi-singer singing voice synthesis with short reference prompt

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.887342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:a681bac154a475917b9287d0f6af24cbff6fd2a66646029c0b609870c1cc53f2

Observation c2560861-549d-47fe-b38d-d0a98e250f28 · outbound

This paper cites Prosody-Adaptable Audio Codecs for Zero-Shot V oice Conversion via In-Context Learn- ing.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Prosody-Adaptable Audio Codecs for Zero-Shot V oice Conversion via In-Context Learn- ing

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.883688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:e03ccf7fb4822e652d0efdbde61ed29b9efafe82c0c1d2cf69c0debe7e11d860

Observation 26b28b96-2fd3-445b-bf14-db9b433fe9bc · outbound

This paper cites Film: Visual reasoning with a general conditioning layer.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Film: Visual reasoning with a general conditioning layer

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.879932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:3e67568527e02335edfe954278f22057c3547d1d2efb14863c64e1b5c5bd588d

Observation 164b0662-aa5c-47ac-9d5a-6e6e0e64ecd4 · outbound

This paper cites Improved parallel wavegan vocoder with per- ceptually weighted spectrogram loss.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Improved parallel wavegan vocoder with per- ceptually weighted spectrogram loss

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.876106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:771b31c39febf07938095d9181d096064ae72fee37382b53fb2d0030f3e3c384

Observation 6dc7dfbd-83c9-44a3-80c8-a88b3a5dfee1 · outbound

This paper cites Flow-guided one- shot talking face generation with a high-resolution audio-visual dataset.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Flow-guided one- shot talking face generation with a high-resolution audio-visual dataset

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.871754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:a7e0b45cc3cef9e94652d9114458948b00238f5ef0b9865b8356e61496ba2f2a

Observation 19b246d7-0020-49fd-8cf5-9eb74c6546ec · outbound

This paper cites V oxceleb: A large-scale speaker identification dataset.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation V oxceleb: A large-scale speaker identification dataset

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.864861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:1ecb96b44bd7ae34d66863f52c7d29a2d9f8d77233955c214f711b0726e797d9

Observation 4baf7e6b-bcc6-4e7e-b8d3-eabe5609d273 · outbound

This paper cites Hallo2: Long-duration and high- resolution audio-driven portrait image animation.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Hallo2: Long-duration and high- resolution audio-driven portrait image animation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.859854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:57c0d44eb5c41ed8ca8fbe775ac86f1c7d45ad1615cf090988c485e288b8176e

Observation 026f25cf-ed6c-4621-b2cc-8d4f843ab51b · outbound

This paper cites Meshtalk: 3d face animation from speech using cross- modality disentanglement.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Meshtalk: 3d face animation from speech using cross- modality disentanglement

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.852453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:59369c7187e8b78cd97f560e36b81afeada25c0ae709766e40ac00377b291c7d

Observation daca2d65-86e7-47af-9fb2-7760c3dbc0d6 · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation A lip sync expert is all you need for speech to lip generation in the wild

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.848920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:07f534b2801c35f71f5c8e1d8a950b4598ee5486eb581040142f555b75dafe73

Observation ec6bad9b-95f9-479c-afbb-ac19f0b4c697 · outbound

This paper cites Fine-grained head pose estimation without keypoints.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Fine-grained head pose estimation without keypoints

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.844160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:a33549405afb7d1913b23a0ced74fd053e4fd183013370654afcc8b88feba522

Observation 21fcfb77-80ee-4ee8-93a2-68dfb07fc2cf · outbound

This paper cites Bailando: 3d dance generation by actor- critic gpt with choreographic memory.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Bailando: 3d dance generation by actor- critic gpt with choreographic memory

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.839699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:15b5cc9eca27832c786201679fab225ba64dc5f3ccccdddd8f40416ea07ee8bc

Observation 0241366b-e474-4b36-be55-2c3d835f8e63 · outbound

This paper cites Robust speech recognition via large-scale weak supervision.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Robust speech recognition via large-scale weak supervision

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.836070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:89fc77c457f1e5682e1d213c98257b9f0aea66e8d34d7dd086d62a0550ce0f54

Pith citing papers

Observation eb3ed498-689b-41fc-b7e1-efa123530228 · inbound

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation cites this paper.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T14:21:28.419042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:44db7a3d97bcd4bec71c9afcd01ab8edee943b47846d6fdce48fe778207c8a02