Pith. sign in

Paper Citation Record · LEDGER

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations

As of 18 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2505.18096.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18096 v2

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:39:13.746751Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact1
  • verified fuzzy39
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d3b796bc-c835-4589-85c3-27c5b74dad10 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations.Advances in neural infor- mation processing systems, 33:12449–12460, 2020.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations wav2vec 2.0: A framework for self-supervised learning of speech representations.Advances in neural infor- mation processing systems, 33:12449–12460, 2020

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:24.059388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:06.664873Z digest=sha256:968a8b096bd0d4f085def52ab3fe657d07a6755a6121ec0e30ef17cd8c21d84a

Observation 06a19676-da99-4faa-b7bf-e69fd9de279f · outbound

This paper cites High-fidelity fa- cial avatar reconstruction from monocular video with gen- erative priors.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations High-fidelity fa- cial avatar reconstruction from monocular video with gen- erative priors

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:23.759935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:06.727929Z digest=sha256:52fac203865747beedd0c68f4c97ee4d1b77407cef9a1d8a557b89015dbc55bd

Observation ff12b923-45aa-4306-a924-592a3dcc8f5a · outbound

This paper cites Pyannote.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Pyannote

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:23.359982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:06.854167Z digest=sha256:2950b520373ab92cebb14b512ae9e3687e551ada288cc0198d30a7be070a495f

Observation 84ff27c2-43a8-4706-a46d-1e6bacefe58d · outbound

This paper cites Expressive speech-driven facial animation.ACM Transactions on Graphics (TOG), 24(4):1283–1302, 2005.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Expressive speech-driven facial animation.ACM Transactions on Graphics (TOG), 24(4):1283–1302, 2005

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:23.050517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:06.984124Z digest=sha256:cd8790ed157fb704d0c34ee3947498ee9b3c9c0a8c4eb7be021e2cdb4b3e2f99

Observation 3f2b911b-466e-4456-ba0a-cab1070360c2 · outbound

This paper cites Human conversation as a system framework: Designing embodied conversational agents.Em- bodied conversational agents, pages 29–63, 2000.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Human conversation as a system framework: Designing embodied conversational agents.Em- bodied conversational agents, pages 29–63, 2000

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:22.655777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:07.100300Z digest=sha256:2864c3d5da07fc35ecf40b6c2aa5355977eada41c2756c42d92f5cb5b000d54b

Observation c4643fe1-6037-4362-b040-2837c21949fe · outbound

This paper cites Cafe-talk: Generating 3d talking face animation with mul- timodal coarse-and fine-grained control.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Cafe-talk: Generating 3d talking face animation with mul- timodal coarse-and fine-grained control

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:22.372229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:07.224197Z digest=sha256:3384b36ac1fc18f7aba8f78baf7390422635e1fa2e90a07937df81ec47a6835e

Observation fd2b9128-deaf-4bc0-a755-0d6898416b8a · outbound

This paper cites Capture, learning, and synthe- sis of 3d speaking styles.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Capture, learning, and synthe- sis of 3d speaking styles

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:22.019144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:07.362385Z digest=sha256:0ad01305e79582b29163169d4e4ab25f600f8790595fef64db00fd43e1508d7b

Observation 8d7d3b56-896f-4f2f-8a5a-3977234923ad · outbound

This paper cites an unresolved cited work.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:39:21.672545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:07.466214Z digest=sha256:bd378d09d1128a901677de174b25c63462b9e36bba499e29c564c9252ad65245

Observation dfd6207e-f5b7-464a-90d4-e2c7fbae2d00 · outbound

This paper cites UniTalker: Scaling up Audio-Driven 3D Facial Animation through A Unified Model.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations UniTalker: Scaling up Audio-Driven 3D Facial Animation through A Unified Model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:07.573603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:07.573603Z digest=sha256:dd4aefc01fef120447898777272dca73a3cb8879705d7d2723ded8a9db42b3f4

Observation 4827c36e-ac0c-4dd1-970f-f3ce33a5c38a · outbound

This paper cites Faceformer: Speech-driven 3d facial anima- tion with transformers.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Faceformer: Speech-driven 3d facial anima- tion with transformers

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:21.535978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:07.722156Z digest=sha256:05d4cc3c33f7fcdd607b6951f72f5ec67a00fcbb7beb3c66c6ee7b8dbd59640e

Observation 7cc3a6c8-ad41-4a75-8f5a-08ec483ea04b · outbound

This paper cites A 3-d audio-visual corpus of af- fective communication.IEEE Transactions on Multimedia, 12(6):591–598, 2010.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations A 3-d audio-visual corpus of af- fective communication.IEEE Transactions on Multimedia, 12(6):591–598, 2010

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:21.366870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:07.853427Z digest=sha256:48d6bbd958c9e0b0e4ce4ac52dbe7834d2a883667cb4dc8960d2fdd5eb873e92

Observation 55544b14-4360-47ec-b385-6c9b16d8e246 · outbound

This paper cites Affective Faces for Goal-Driven Dyadic Communication.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Affective Faces for Goal-Driven Dyadic Communication

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:07.961032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:07.961032Z digest=sha256:04e93dd10218bfe809a8cf9e6701e1d152f9cb7943b008a581486e8cea09630e

Observation eddc3004-6431-47f9-8f8d-d31ca1d0e223 · outbound

This paper cites From Pixels to Portraits: A Comprehensive Survey of Talking Head Generation Techniques and Applications.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations From Pixels to Portraits: A Comprehensive Survey of Talking Head Generation Techniques and Applications

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:08.118474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:08.118474Z digest=sha256:16989cfd6d224b68867d5b5998fa31f901be457fec8f2421ae2909453642322b

Observation ca303e15-0c16-4bb4-be8e-e2eac657cb9a · outbound

This paper cites Long short-term memory.Neural Computation MIT-Press, 1997.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Long short-term memory.Neural Computation MIT-Press, 1997

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:21.231329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:08.246824Z digest=sha256:d7f65f110e6d87d7aead3ec44b092f5ed702b36abe5b96ee8345319d984d2efd

Observation 3f0ddd31-29ba-466f-b13e-f4d3ebfcf144 · outbound

This paper cites Toward rnn based micro non-verbal behavior generation for virtual listener agents.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Toward rnn based micro non-verbal behavior generation for virtual listener agents

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:20.979997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:08.357873Z digest=sha256:a4a4ab06f8b262e1f52b29325a32f3806c009210f9254e0d5ba21c4bd1762d93

Observation 6309504b-f066-41d6-953b-f129ecc8e25b · outbound

This paper cites Audio-driven facial animation by joint end- to-end learning of pose and emotion.ACM Transactions on Graphics (TOG), 36(4):1–12, 2017.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Audio-driven facial animation by joint end- to-end learning of pose and emotion.ACM Transactions on Graphics (TOG), 36(4):1–12, 2017

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:08.474671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:08.474671Z digest=sha256:1b5c930cf44e279526517cc1d250764236ddb339a0e2f39401458a1beb8eb0ea

Observation 7ef5a1c9-c907-4aac-9f80-de3e4a631326 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Adam: A Method for Stochastic Optimization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:08.615584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:08.615584Z digest=sha256:1d5cccb82333152dd3be8815018351cec1b2d39e6a5c1df6d7f2683765973834

Observation a180a451-e273-4a16-9e5e-ae34426f373f · outbound

This paper cites Iianet: An intra-and inter-modality attention network for audio- visual speech separation.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Iianet: An intra-and inter-modality attention network for audio- visual speech separation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:20.711792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:08.753831Z digest=sha256:064012555557627e6a5645e4c3ee2e9d655b6182b4cdda6e180345e869b9eda1

Observation 29ff24c6-5222-46b6-941d-14e1c6653dd4 · outbound

This paper cites Learning a model of facial shape and expression from 4d scans.ACM Trans.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Learning a model of facial shape and expression from 4d scans.ACM Trans

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:20.498754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:08.849891Z digest=sha256:73ff78460887542b3628f5544935b7ba40c55918b524d69b605dcb8551f43a8d

Observation 4d1a5a81-b6b6-4ed5-894c-c1e9831dc938 · outbound

This paper cites One-shot high-fidelity talking- head synthesis with deformable neural radiance field.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations One-shot high-fidelity talking- head synthesis with deformable neural radiance field

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:20.205414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:08.936420Z digest=sha256:784bf3cfaddeee512a33df4e6138cf1fcb85594f014289f2dce5a59a6cc87a4b

Observation 7b4f4828-477d-4500-9c81-0c5ef5543cc2 · outbound

This paper cites Proactive con- versational agents in the post-chatgpt world.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Proactive con- versational agents in the post-chatgpt world

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:19.990434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:09.113832Z digest=sha256:157f9bc41c29caf009d77c4a8dce549c8047b6b402c9bb991c080c7359e1aff6

Observation d54a3abc-70b4-4dbc-b34d-9f4fd95a92ca · outbound

This paper cites Mfr-net: Multi-faceted responsive listening head generation via denoising diffusion model.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Mfr-net: Multi-faceted responsive listening head generation via denoising diffusion model

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:19.774491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:09.273424Z digest=sha256:02df64a2d58a7e7cd40bdef17ea637253c3b38c9956bd5d0b4478a8971a6327f

Observation 7eaf8bea-ed03-437c-af30-c8df55ba20ab · outbound

This paper cites Customlistener: Text-guided responsive inter- action for user-friendly listening head generation.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Customlistener: Text-guided responsive inter- action for user-friendly listening head generation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:19.530943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:09.391277Z digest=sha256:188dd75a5ad0341d40477ed01edf308588068e7beccacd7e19c0966582893136

Observation 5644a605-6430-4064-a6ec-56219208bce4 · outbound

This paper cites MediaPipe: A Framework for Building Perception Pipelines.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations MediaPipe: A Framework for Building Perception Pipelines

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:09.516177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:09.516177Z digest=sha256:d190e8e686a5519c0c79c23b289948154df5eaa0633775206f04773daad26974

Observation 4abc0771-f79b-42ec-9f2a-145e8594ce56 · outbound

This paper cites ReactFace: Online Multiple Appropriate Facial Reaction Generation in Dyadic Interactions.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations ReactFace: Online Multiple Appropriate Facial Reaction Generation in Dyadic Interactions

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:39:14.078463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:09.681648Z digest=sha256:191619004f3e2b031462763599a519a2127663c124ac0faf83678a2679e6ca02

Observation 474495de-3784-412f-b3f0-820278826618 · outbound

This paper cites DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:09.795467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:09.795467Z digest=sha256:9a983f29743e2ea117b260b8f08e10572d6083f62d7824cf126a4642e915f8e2

Observation 323b7c63-c06b-4cf7-af42-8cf5fb3c9973 · outbound

This paper cites Learning to listen: Modeling non-deterministic dyadic facial motion.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Learning to listen: Modeling non-deterministic dyadic facial motion

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:19.277225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:09.900396Z digest=sha256:274b24355642f740c49c1d0ef0f5a7ac1c7b75dcaff8f37c599519b85c0e843e

Observation 3edcf984-b1f2-4670-b137-3871b492d646 · outbound

This paper cites Can language models learn to listen? InProceedings of the IEEE/CVF In- ternational Conference on Computer Vision, pages 10083– 10093, 2023.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Can language models learn to listen? InProceedings of the IEEE/CVF In- ternational Conference on Computer Vision, pages 10083– 10093, 2023

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:19.076045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:10.012426Z digest=sha256:e5ad294835658a9de20779703f191d2f79a4068e866b357c7f855b07862a110b

Observation 6149e777-ba56-4bfb-ab1c-234408708823 · outbound

This paper cites From audio to photoreal embodiment: Synthesizing humans in conversations.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations From audio to photoreal embodiment: Synthesizing humans in conversations

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:18.821438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:10.117617Z digest=sha256:9df27bea6d72582f9fbe98d5424d8303fce14ae6d54aaf32ca5b6ae722266b0f

Observation 2d384c6f-2910-4821-bb5d-474f55be4d64 · outbound

This paper cites Real-time 3d talking head from a synthetic viseme dataset.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Real-time 3d talking head from a synthetic viseme dataset

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:18.577419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:10.259878Z digest=sha256:4303599eb10eedb0876e24c969b9acfbe60885cb30f2511aee5d37580603379b

Observation 298d4bb5-b1bd-42a8-bfd0-31f5d2d268f1 · outbound

This paper cites ScanTalk: 3D Talking Heads from Unregistered Scans.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations ScanTalk: 3D Talking Heads from Unregistered Scans

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:10.356811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:10.356811Z digest=sha256:1650cfa979696747afd72f3dfc340f4eae2d76204ff795b257768c87bf072c6f

Observation 22857179-683d-4e48-a559-6bea681d0ce2 · outbound

This paper cites Dpe: Dis- entanglement of pose and expression for general video por- trait editing.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Dpe: Dis- entanglement of pose and expression for general video por- trait editing

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:10.470371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:10.470371Z digest=sha256:bf60b3879831e503c19c60fd72c8b16f59dff656641f8d25edf2c74ff88f6919

Observation b777e672-7a0f-4494-b473-a645593ed547 · outbound

This paper cites Selftalk: A self- supervised commutative training diagram to comprehend 3d talking faces.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Selftalk: A self- supervised commutative training diagram to comprehend 3d talking faces

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:18.342972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:10.596544Z digest=sha256:3ec12276e13b9b8c330638d4291c17401b443f9ac42b267dfa13135c47af064b

Observation cf1bd416-72da-4c94-bef3-b8f3d1f3b6c9 · outbound

This paper cites Emotalk: Speech-driven emotional disentanglement for 3d face anima- tion.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Emotalk: Speech-driven emotional disentanglement for 3d face anima- tion

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:18.136046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:10.717041Z digest=sha256:b432222b2c0a48e175c3ef29b48a48f4a6fc55b70a84c84946f5428685e6649e

Observation 3b82443b-c7d1-49ac-a94f-fbc492e180dd · outbound

This paper cites Synctalk: The devil is in the synchronization for talking head synthesis.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Synctalk: The devil is in the synchronization for talking head synthesis

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:17.923327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:10.847276Z digest=sha256:4b7695c06a16122f9bfe663db615f6f2d7d06e9193acc69d361e2da914ae909c

Observation ca293a36-ff56-41e0-bb22-1ebf757d3c72 · outbound

This paper cites Meshtalk: 3d face an- imation from speech using cross-modality disentanglement.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Meshtalk: 3d face an- imation from speech using cross-modality disentanglement

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:17.685594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:11.024346Z digest=sha256:4597a32cde79c440d8ae2a1d6a9b4beafe515c732958560ac1086d3ca92992eb

Observation 744d021a-c70d-45b0-8101-8011d94f188f · outbound

This paper cites Emotional listener portrait: Neural lis- tener head generation with emotion.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Emotional listener portrait: Neural lis- tener head generation with emotion

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:17.459200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:11.164248Z digest=sha256:7bb25b6ff018e07e49aef83ea5b709d72d3ae02429972270c758c37ae47e822a

Observation 422e804e-b1db-41f4-84f9-84e6df6d5335 · outbound

This paper cites React2023: The first multiple appropriate facial reaction generation chal- lenge.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations React2023: The first multiple appropriate facial reaction generation chal- lenge

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:17.268369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:11.336635Z digest=sha256:960dc071d53e384d62366e78732664f5a004a0a28c3ffc66f5310efb9099d669

Observation 6cd5d85d-b83a-4509-8cbc-15c47c3a9567 · outbound

This paper cites TransNet V2: An effective deep network architecture for fast shot transition detection.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations TransNet V2: An effective deep network architecture for fast shot transition detection

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:11.450440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:11.450440Z digest=sha256:6161f9c5b2a919d836a941684118f50511813d86b8062740bff2bf898be330a0

Observation 8c9f52ef-28d9-45f5-9f83-a0ceb0336a92 · outbound

This paper cites Laughtalk: Expressive 3d talking head generation with laughter.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Laughtalk: Expressive 3d talking head generation with laughter

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:17.031219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:11.606460Z digest=sha256:765ee89772d6b127e423574ca60fed5657a62abc15880e0281b136651475b66d

Observation 03f9e27e-2e52-44d7-b817-9560636c1dfe · outbound

This paper cites Imitator: Personalized speech-driven 3d facial animation.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Imitator: Personalized speech-driven 3d facial animation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:16.845207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:11.704385Z digest=sha256:12aa05bd974017595586b6409c36cdb4b3322cb8ae272b4bfb01e87e75e8132d

Observation c45ad741-e225-4f69-90c9-ea95eaa12a2b · outbound

This paper cites Dyadic Interaction Modeling for Social Behavior Generation.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Dyadic Interaction Modeling for Social Behavior Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:11.891223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:11.891223Z digest=sha256:b62211a0256bad44d33856f1e3ade237f42aa4d1b8940b32be289b8217740fa0

Observation 5cc2cdf7-7a8a-4c89-a0db-6bd81d8d7aae · outbound

This paper cites Attention is all you need.Advances in Neural Information Processing Systems, 2017.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Attention is all you need.Advances in Neural Information Processing Systems, 2017

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:16.620401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:12.029845Z digest=sha256:d05c9ee0972fc069c98f1471e8a19fb245b3ee959758e7629c175ccd9c3b8b4a

Observation 91285c40-8cba-4b34-8140-50857e423498 · outbound

This paper cites VGG-Tex: A Vivid Geometry-Guided Facial Texture Estimation Model for High Fidelity Monocular 3D Face Reconstruction.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations VGG-Tex: A Vivid Geometry-Guided Facial Texture Estimation Model for High Fidelity Monocular 3D Face Reconstruction

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:12.156492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:12.156492Z digest=sha256:3351cfd3dc59bd9caf2b6cba627db218b6f35424bfce24f9f413e479b1ed3ac0

Observation f0df8f66-6cb5-4eb3-8895-e9c2949a2219 · outbound

This paper cites AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:12.290540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:12.290540Z digest=sha256:c66e246a7a2f4d2a976352ea52bfce875a2c2fa90504ab4a14604d702cf6c756

Observation 563c1a0b-a1bb-428e-946d-541f3395724e · outbound

This paper cites Codetalker: Speech-driven 3d facial animation with discrete motion prior.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Codetalker: Speech-driven 3d facial animation with discrete motion prior

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:16.436078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:12.406648Z digest=sha256:9b76045482c1a49b12a16340dbeda55e6280be542770f9c152694360ee9c4729

Observation c93a91c0-ced5-4d0f-b9c0-7e464d4e2201 · outbound

This paper cites Nofa: Nerf-based one-shot facial avatar recon- struction.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Nofa: Nerf-based one-shot facial avatar recon- struction

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:16.283323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:12.535992Z digest=sha256:b02e6b3489487ef149f347b05cfb0df7fa5ef628cbe2cc3547280709f8285c44

Observation 2e539a54-fe2a-433d-aafa-2864bcae76bc · outbound

This paper cites Human-computer interaction system: A survey of talking-head generation.Electronics, 12(1):218, 2023.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Human-computer interaction system: A survey of talking-head generation.Electronics, 12(1):218, 2023

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:16.057653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:12.656839Z digest=sha256:feb45d93c2da0e328c0f9498716cfae3707116208cafc54a6bbf650a4a3adecc

Observation bfb7cbfd-d1cd-4672-8cd5-d9f5676e63dd · outbound

This paper cites Responsive listening head generation: a benchmark dataset and baseline.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Responsive listening head generation: a benchmark dataset and baseline

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:15.836708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:12.761642Z digest=sha256:43e328fb872ec1e086781e36e0d5d7ab7bda3ad1e76b275c74aa788de218a3b9

Observation e6428501-3063-4c70-be21-f5a5aa78b528 · outbound

This paper cites Meta-Learning Empowered Meta-Face: Personalized Speaking Style Adaptation for Audio-Driven 3D Talking Face Animation.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Meta-Learning Empowered Meta-Face: Personalized Speaking Style Adaptation for Audio-Driven 3D Talking Face Animation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:12.873023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:12.873023Z digest=sha256:1849785ab26077c9a78a5d6d23b08b4484b0e32f6880903badfb154b487d9639

Observation 22127174-c4f6-4c6c-8004-ca10ebe3e35d · outbound

This paper cites Visemenet: Audio- driven animator-centric speech animation.ACM Transac- tions on Graphics (TOG), 37(4):1–10, 2018.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Visemenet: Audio- driven animator-centric speech animation.ACM Transac- tions on Graphics (TOG), 37(4):1–10, 2018

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:15.631924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:12.982202Z digest=sha256:e3efa7863383bac3041da1d8b43439ae3f2df0c0217046c1ce40333aa442aa6e

Observation e4c46f00-d496-4697-8ba5-f8be58c03463 · outbound

This paper cites Network Architecture In this section, we provide comprehensive implementation details of our DualTalk framework.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Network Architecture In this section, we provide comprehensive implementation details of our DualTalk framework

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:15.443570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:13.092431Z digest=sha256:aab4ec474f2a9afaf2345bb265dacdb1ab8d627ed9b2eba20cd67b96b765edde

Observation 9660e38e-2e40-4514-b2d1-b2349330c556 · outbound

This paper cites Here, we provide detailed in- formation about our data collection, processing procedures, and dataset statistics.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Here, we provide detailed in- formation about our data collection, processing procedures, and dataset statistics

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:14.955563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:13.359279Z digest=sha256:44d06a54dafccf51f5c574d3227418b6c5d1e524852f42d87e3a36abae5f1a71

Observation 0b79ff21-5b4d-459d-8fac-6cb437bf2369 · outbound

This paper cites an unresolved cited work.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:39:14.715905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:13.474746Z digest=sha256:53856dc50fae054d8454f4af22f71390d23a64478b4505fea12471c33ffc07b4

Observation 0fa3e84a-56c3-4942-a4a8-81a4a20ccb57 · outbound

This paper cites an unresolved cited work.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:39:14.567364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:13.629376Z digest=sha256:c127a8ae0c221328110066f3dc05a1a32f8b26863f7c448f448b2207d756756d

Observation 0bef61ec-2aa8-4f8d-9050-f704ba02b849 · outbound

This paper cites While DualTalk ex- cels in creating synchronized and natural two-speaker con- versations, it cannot yet handle multi-party interactions, which are common in real-world applications.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations While DualTalk ex- cels in creating synchronized and natural two-speaker con- versations, it cannot yet handle multi-party interactions, which are common in real-world applications

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:14.318967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:13.746751Z digest=sha256:219117930d952bf7fd8c55276df196a3f0fbc94dcd5f0cf5b30e122348f8c2f4

Observation 483660ca-40fb-420d-be30-2d4a6a2692a7 · outbound

This paper cites The decoder follows a similar structure but includes additional cross- attention layers to integrate information from both speakers.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations The decoder follows a similar structure but includes additional cross- attention layers to integrate information from both speakers

Reference 512

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:15.192741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:39:13.227145Z digest=sha256:32a443a0b07a1ad57d2119befe702164bc6b86d8e31634c03ecb2b377e562693

Pith citing papers

No inbound Pith citation observations are available.