Pith. sign in

Paper Citation Record · LEDGER

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync

As of 18 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 0 inbound Pith citation observations for arXiv:2507.20452.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.20452 v1

Coverage vector

measured 86 of 86 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:39:36.267401Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

86 of 86 outbound references displayed

  • verified exact3
  • verified fuzzy46
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b5c61b92-42de-440b-8864-c2c13d694652 · outbound

This paper cites Facewarehouse: A 3d facial expression database for visual computing.IEEE Transactions on Visualization and Computer Graphics, 20(3):413–425, 2013.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Facewarehouse: A 3d facial expression database for visual computing.IEEE Transactions on Visualization and Computer Graphics, 20(3):413–425, 2013

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:27.844754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:27.844754Z digest=sha256:5e1e5e6fdb4aa1e267b98b59674c86abb4b07b5c3a47091dae5c38b57b42beff

Observation e53ddc25-af8d-4e39-a107-3f167b852de0 · outbound

This paper cites Hiface: High-fidelity 3d face reconstruction by learning static and dynamic details.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Hiface: High-fidelity 3d face reconstruction by learning static and dynamic details

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:27.988641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:27.988641Z digest=sha256:7e72e756cbd94ecad9368cdbbb005d553f97ae2260dcfb39085157ba5b71f81b

Observation 9ccc7da9-a064-4f53-b334-f8df148f8e21 · outbound

This paper cites IQA-PyTorch: Pytorch toolbox for image qual- ity assessment.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync IQA-PyTorch: Pytorch toolbox for image qual- ity assessment

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:28.098360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:28.098360Z digest=sha256:9a30eb260935d06eb38b9b91b8e3c204cf58a31e5d195335883036bf027c9b70

Observation 036fa333-8e8a-48f0-b837-092eeda73bf4 · outbound

This paper cites Topiq: A top-down approach from semantics to distortions for image quality assessment.IEEE Transactions on Image Processing, 2024.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Topiq: A top-down approach from semantics to distortions for image quality assessment.IEEE Transactions on Image Processing, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:28.199572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:28.199572Z digest=sha256:8b41d3942f695a174ff2f96bc2a30eaf044a1e8a631b160a4bfa358302b33ac9

Observation 2b8a60f2-fb4c-42f6-9e76-3f61b8d69fcc · outbound

This paper cites VoxCeleb2: Deep Speaker Recognition.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync VoxCeleb2: Deep Speaker Recognition

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:28.325314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:28.325314Z digest=sha256:1f7b35bd1720f57bae97f6238246a390eb51814f48fa656ab9de4276f0d3b777

Observation 64300e98-df5e-45d8-a1ff-9ed4a51d992d · outbound

This paper cites Emoca: Emotion driven monoc- ular face capture and animation.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Emoca: Emotion driven monoc- ular face capture and animation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:28.419724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:28.419724Z digest=sha256:0ad503a3c8422dce842f44bb274e90b656815c94be1e7fcb49a11c6563629257

Observation 5e8311fb-0280-47f4-8c02-596ffa3c3ea7 · outbound

This paper cites Arcface: Additive angular margin loss for deep face recognition.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Arcface: Additive angular margin loss for deep face recognition

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:28.554745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:28.554745Z digest=sha256:838cb8b76839ce5260663c542092c68d4d19dcf17256e81c99d1544f913f34c7

Observation a8416907-96b5-4a95-ad01-e07d2f657584 · outbound

This paper cites Ac- curate 3d face reconstruction with weakly-supervised learning: From single image to image set.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Ac- curate 3d face reconstruction with weakly-supervised learning: From single image to image set

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:52.154750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:28.634750Z digest=sha256:c7466769b334a7a2f95099b5d76ad0aaadd1284b2785d2786493e3c4769d204f

Observation b6826680-cb5a-4022-be23-c926fbddc557 · outbound

This paper cites Headgan: One-shot neural head synthesis and editing.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Headgan: One-shot neural head synthesis and editing

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:51.704889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:28.779920Z digest=sha256:896f07b124e7f68580a3a57409e538e2c622384fc3b35c4f7881f838e65e5b9b

Observation 9594e41b-880b-42f0-a799-2e44e584cdbe · outbound

This paper cites Free-headgan: Neural talking head synthesis with explicit gaze control.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Free-headgan: Neural talking head synthesis with explicit gaze control

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:51.294070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:28.914742Z digest=sha256:18e20474115d09eb50457318528a42ff9de53a1e10eeb9617bf9fecced3413b9

Observation 40b383c9-873e-45b6-b18e-906817f839bd · outbound

This paper cites Megaportraits: One-shot megapixel neural head avatars.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Megaportraits: One-shot megapixel neural head avatars

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:51.035436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:29.098001Z digest=sha256:69e077af22ad1a4515ef2507c3b187b4a0a16908e8d6e9111eaa621b07153bca

Observation 35fb1ffb-7c72-4aa9-95bd-33bc82142019 · outbound

This paper cites Emoportraits: Emotion-enhanced multimodal one-shot head avatars.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Emoportraits: Emotion-enhanced multimodal one-shot head avatars

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:50.747678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:29.252262Z digest=sha256:def5a6899aaf8aa9e4792d23e2f424b8066cfa1efc3225224cdcfdfd714fb7b2

Observation aa420262-44c8-4865-bc32-8fc217c28963 · outbound

This paper cites 3d morphable face models—past, present, and future.ACM Transactions on Graphics (ToG), 39(5):1–38, 2020.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync 3d morphable face models—past, present, and future.ACM Transactions on Graphics (ToG), 39(5):1–38, 2020

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:50.434842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:29.340935Z digest=sha256:ce3c4c25c3e4f1ee87b59e82a043b47d93a12bd024457eb6c82a7b41d5fba416

Observation 2f75bd06-e7b8-4aa1-9fc2-b4d209e1fabb · outbound

This paper cites Facial action coding system.Environmental Psy- chology & Nonverbal Behavior, 1978.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Facial action coding system.Environmental Psy- chology & Nonverbal Behavior, 1978

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:50.028325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:29.438106Z digest=sha256:08fd1acc805c1918e1f4f51c9609dabc8130fa7b277a51bfeb322c4b188ec323

Observation e23e40ce-d579-4825-974c-0d672e3df384 · outbound

This paper cites Looking to Listen at the Cocktail Party: A Speaker-Independent Audio-Visual Model for Speech Separation.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Looking to Listen at the Cocktail Party: A Speaker-Independent Audio-Visual Model for Speech Separation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:29.510807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:29.510807Z digest=sha256:de870f7feed3eb2c1ef3bcb68153f82c5a33feee058a85eeb33d45bedcfcfa80

Observation 5c86ad74-8e25-477c-aeab-aef5108e5339 · outbound

This paper cites Learning an animatable detailed 3d face model from in-the-wild images.ACM Transactions on Graphics (ToG), 40(4):1–13, 2021.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Learning an animatable detailed 3d face model from in-the-wild images.ACM Transactions on Graphics (ToG), 40(4):1–13, 2021

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:29.590290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:29.590290Z digest=sha256:0daf3d15778bba5df823f277f2f9244452f276741025ee87bd4d8728bb4f4b2c

Observation 833d90c2-3484-425f-a387-38cb244c7af6 · outbound

This paper cites Surface simplification using quadric error met- rics.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Surface simplification using quadric error met- rics

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:49.634750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:29.646577Z digest=sha256:5a57f3719bdcce18f25bf4dd31c57cad3b85cddc9c3f433765f94ccb32dffc73

Observation 2d03d31e-d535-4d28-a03a-7c05d0886086 · outbound

This paper cites Morphable face models-an open frame- work.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Morphable face models-an open frame- work

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:49.325837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:29.710651Z digest=sha256:044526deea4743138aefcfcdd4f21bf0877e99fb7ea5cfa9cd151666a68d60ab

Observation 23d37766-fe89-4821-89f4-32a86b935009 · outbound

This paper cites Attention Mesh: High-fidelity Face Mesh Prediction in Real-time.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Attention Mesh: High-fidelity Face Mesh Prediction in Real-time

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:29.799219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:29.799219Z digest=sha256:478ec01683b4e730d386829aee7b68eee6e1a18e1a032e108497c113c5aab7e8

Observation 5ceb2151-0041-4694-aa51-e82987f7f473 · outbound

This paper cites LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:29.907882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:29.907882Z digest=sha256:a1b7870a3137eb674387f929b7376a81760b61e5a0e3b2be731968c5e1d32220

Observation 5aaf2052-996d-4143-a032-daeb9c554c38 · outbound

This paper cites Deep residual learning for image recognition.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Deep residual learning for image recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:29.997607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:29.997607Z digest=sha256:213e020d49e011d60c22fad74dc53676e302e1c2b5c044577a9f04a511265c50

Observation 10ebb122-024f-4579-a734-e317f77588ec · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:30.129703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:30.129703Z digest=sha256:a9377f200390e16c201dbff665f680679d129cc554a7514955c0eaa927fb056d

Observation 2f9c5b4b-a39c-4015-8bad-19eb3343bda5 · outbound

This paper cites Classifier-Free Diffusion Guidance.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Classifier-Free Diffusion Guidance

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:30.197962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:30.197962Z digest=sha256:ab55abfd72b484a925fe7691a949c8c429fa0353e6dfec50483c72c5b8a5775c

Observation dc607002-be97-4dde-8835-1fdd06b0a10f · outbound

This paper cites Denoising diffusion probabilistic models.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Denoising diffusion probabilistic models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:30.313052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:30.313052Z digest=sha256:c3e3bde3ad138036f5d880a5bf75fced9eab8bbec23d14cf06bf14752030a0d4

Observation aa1f875f-e376-4ce1-a937-d31e3f6860e3 · outbound

This paper cites Image-to-image trans- lation with conditional adversarial networks.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Image-to-image trans- lation with conditional adversarial networks

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:48.769236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:30.376358Z digest=sha256:245f7fe3441a85ce8aed270c6d9fda5b16c5037a8adb95cf10d819129d98d159

Observation 91a3f7c3-d4f3-4483-990a-bb19a66d338a · outbound

This paper cites RealTalk: Real-time and Realistic Audio-driven Face Generation with 3D Facial Prior-guided Identity Alignment Network.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync RealTalk: Real-time and Realistic Audio-driven Face Generation with 3D Facial Prior-guided Identity Alignment Network

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:30.454868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:30.454868Z digest=sha256:523bb18131867cae24ba1db646624a16070c407bd808d80a5495a411eb984f94

Observation d28459ac-b837-43cc-bd0a-d8b7408469c6 · outbound

This paper cites Eamm: One-shot emotional talking face via audio-based emotion-aware motion model.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Eamm: One-shot emotional talking face via audio-based emotion-aware motion model

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:48.466735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:30.530150Z digest=sha256:42ed7ec0642bf466d88c57450cfe6c7fa6eaf3033bd793ed2a5ac8dde4061f61

Observation 8c28182c-38c5-4419-ae04-ebc6fc929316 · outbound

This paper cites Loopy: Taming audio-driven portrait avatar with long-term motion dependency.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Loopy: Taming audio-driven portrait avatar with long-term motion dependency

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:48.172525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:30.643230Z digest=sha256:bf3a2f9beab8dbcd3cf5441bef978e4c7d0112df934abd576295423851d15fca

Observation 56aeb3b4-4077-4ecc-9755-50f082108855 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Adam: A Method for Stochastic Optimization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:30.716088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:30.716088Z digest=sha256:672e8450bada8b975253eeec8b8b1674d7fb1f2bfd9ac19807cfa16d5182347a

Observation 97bd352f-9815-4354-8069-726945ac47f2 · outbound

This paper cites Photo-realistic single image super-resolution using a generative adversarial network.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Photo-realistic single image super-resolution using a generative adversarial network

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:47.884757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:30.809494Z digest=sha256:a5355bf53a955f2db135d5bbc7556acdfa1b98a6159c76596b20969c89a10236

Observation 3299ad7a-593c-4076-8a1e-68af829e8e84 · outbound

This paper cites LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:30.957571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:30.957571Z digest=sha256:6528dc0504b60d892dbfde3a690a53750e09b1b55392e2e0d1b366c1cce1e473

Observation 09d231e1-49f5-4228-b14c-a788cc1481c1 · outbound

This paper cites Learning formation of physically-based face attributes.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Learning formation of physically-based face attributes

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:47.487963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:31.057220Z digest=sha256:d56a768eac27d796a09c1993553fb5113d7e2d4bec7f7cdee36aa4a7a29b65ea

Observation ed54530d-224a-4b51-8051-7b4468aa05e2 · outbound

This paper cites Geometric GAN.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Geometric GAN

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:31.123805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:31.123805Z digest=sha256:7af613fe53634191b8366a905bc2f50cc6a3114c6eb60c9692b77e135ea1b43e

Observation 5abfb1b5-46c5-4e93-be8e-d653e54737b1 · outbound

This paper cites Robust high- resolution video matting with temporal guidance.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Robust high- resolution video matting with temporal guidance

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:47.275043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:31.204493Z digest=sha256:b733c1f0cbc763ad0728a4419fa86387ea0f6a4d716e33797dfbd0dd6b54a76b

Observation 447359e1-4e18-4db1-b8b2-17496f8b76da · outbound

This paper cites Anitalker: animate vivid and diverse talking faces through identity-decoupled facial motion encoding.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Anitalker: animate vivid and diverse talking faces through identity-decoupled facial motion encoding

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:46.944757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:31.283100Z digest=sha256:c12a65179b25dbf92a33b1dc511072965f536a8aba4201b9f43588a584605fde

Observation 916803fc-2b96-480e-86f2-02362b0caadb · outbound

This paper cites Decoupled Weight Decay Regularization.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Decoupled Weight Decay Regularization

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:31.347974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:31.347974Z digest=sha256:40d95cc6d7531e15d7cdad5871aa9b51dc0f3315b38433e211c8f967f8923fad

Observation 6d010939-c2c8-4da8-be34-77c14a988298 · outbound

This paper cites Repaint: Inpainting using denoising diffusion probabilistic models.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Repaint: Inpainting using denoising diffusion probabilistic models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:31.456523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:31.456523Z digest=sha256:fc4416b20799723a1574fe08d4462f2aca2ad85ae7f0bd4a8071148cc9f69f58

Observation 59ece410-6f95-4a26-8001-f29fb2306383 · outbound

This paper cites Implicit warping for animation with image sets.Advances in Neural Information Processing Systems, 35:22438–22450, 2022.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Implicit warping for animation with image sets.Advances in Neural Information Processing Systems, 35:22438–22450, 2022

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:46.564761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:31.554628Z digest=sha256:2a89a893c5a815d0d2009e1efeb77342628f74dc41d327b2baa003e896683a5f

Observation 629c79a5-ad66-474c-8235-c0b5df06a66b · outbound

This paper cites Sidgan: High-resolution dubbed video generation via shift-invariant learning.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Sidgan: High-resolution dubbed video generation via shift-invariant learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:46.209299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:31.616512Z digest=sha256:24e62435b3d88ba504ab25acddc71852e1109c402bf061c4b0743e07448311eb

Observation 78b3d442-d2a7-46ec-8f74-827b8593fadd · outbound

This paper cites SAiD: Speech-driven Blendshape Facial Animation with Diffusion.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync SAiD: Speech-driven Blendshape Facial Animation with Diffusion

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:31.735563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:31.735563Z digest=sha256:534ca70d0e50fe813e8072c0a7ca8a511f69064e76acb993c34d5dbeaf0f0d4b

Observation e12a04d5-9e4c-4f95-bbd4-fa6df4d89496 · outbound

This paper cites Synctalk- face: Talking face generation with precise lip-syncing via audio-lip memory.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Synctalk- face: Talking face generation with precise lip-syncing via audio-lip memory

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:45.930393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:31.837154Z digest=sha256:57e0a1c420efafe527db95d7c41f12abd69c6e2110f861888fe10b8333c1efe5

Observation 18b4a7fe-d2cd-440b-9152-21e9f6e1907d · outbound

This paper cites Interpretable Convolutional SyncNet.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Interpretable Convolutional SyncNet

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-06T13:39:37.325379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:31.975485Z digest=sha256:3ae299262d6a9824354a9c4f24f848c657a6ca4292ac9fbfc4934b25e6cb09c2

Observation 9b9c38e4-9715-443e-b3cd-74df5d3d15ed · outbound

This paper cites Semantic image synthesis with spatially-adaptive normalization.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Semantic image synthesis with spatially-adaptive normalization

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:45.534830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:32.077299Z digest=sha256:99c3af9801ca1198e04d8ae497450aeb74d1566c5c00f96b1ac8f082d092108a

Observation 7a8cfbf9-506e-4a66-a661-6413bbe2ba9e · outbound

This paper cites A 3d face model for pose and illumination invariant face recognition.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync A 3d face model for pose and illumination invariant face recognition

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:45.197617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:32.253438Z digest=sha256:1f405b3925d94652cd53a8ac442612eebcac3c3b770a183a16786f9eae43bef9

Observation 43871328-a0f3-4e5d-a8e3-3299a07bf0da · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync A lip sync expert is all you need for speech to lip generation in the wild

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:32.384741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:32.384741Z digest=sha256:19ddb69a139109853b8e7cdeea11cdf83ff789f5332bb1e3311933ed1579ba8e

Observation 13cc7b11-b53d-476c-adfe-efcb38840494 · outbound

This paper cites Accelerating 3D Deep Learning with PyTorch3D.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Accelerating 3D Deep Learning with PyTorch3D

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:32.431687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:32.431687Z digest=sha256:fe998be583170320c803a1d39f3ffcf4cc9d33d33d5ea0b677e9afa23e8b9814

Observation 04e3ab69-bc01-496f-805c-c7c2f6fad83c · outbound

This paper cites Pirenderer: Control- lable portrait image generation via semantic neural rendering.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Pirenderer: Control- lable portrait image generation via semantic neural rendering

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:44.816217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:32.553045Z digest=sha256:bc14832a74c87d5ed9ea9e2ecb624373d9a7f5d40df5ac9059126fbc375c276e

Observation f03f55cd-d0b9-40a6-9506-02867c514aad · outbound

This paper cites U-net: Convolutional net- works for biomedical image segmentation.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync U-net: Convolutional net- works for biomedical image segmentation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:32.603221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:32.603221Z digest=sha256:b1a6bd9af36ac5c5da81f41332d43269533966b08ef708ed82f0afcb97bc3014

Observation 15b40bc3-5270-416b-8f8f-306d813fd4f6 · outbound

This paper cites Palette: Image-to-image diffusion mod- els.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Palette: Image-to-image diffusion mod- els

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:44.479606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:32.656474Z digest=sha256:d6c2d61fb6f58edafd9b4ce193e664467144a685d92044af2f03d86963494769

Observation a5e4fb18-61d8-4aa4-9951-19e4c71d618d · outbound

This paper cites Improved techniques for training gans.Advances in neural information processing systems, 29, 2016.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Improved techniques for training gans.Advances in neural information processing systems, 29, 2016

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:32.802566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:32.802566Z digest=sha256:107c15712e0ca46e1e4914aa483d3dd5933f5f83c4d28581b742bf54e9aa584b

Observation d96f5c64-a23c-48f1-b7d9-a8dcc8fa3ca1 · outbound

This paper cites pytorch-fid: FID Score for PyTorch.https://github.com/ mseitzer/pytorch-fid, August 2020.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync pytorch-fid: FID Score for PyTorch.https://github.com/ mseitzer/pytorch-fid, August 2020

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:44.054773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:32.952752Z digest=sha256:acf6f00af7ff7c0e2afeef4faf02db2f8fbabe9882d3bb43042c37bad2bf7ccc

Observation 207ed7f0-e8c4-4a9a-a9ab-e46f2171d204 · outbound

This paper cites First order motion model for image animation.Advances in neural informa- tion processing systems, 32, 2019.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync First order motion model for image animation.Advances in neural informa- tion processing systems, 32, 2019

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:43.721816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:33.025935Z digest=sha256:50e47bc9bfbcbebadce4cc633d822139ced7da3d981aa9bc53ce63f28f0e0887

Observation 634d6d80-3e21-434b-87c3-1a12f1065992 · outbound

This paper cites Very Deep Convolutional Networks for Large-Scale Image Recognition.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Very Deep Convolutional Networks for Large-Scale Image Recognition

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:33.146640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:33.146640Z digest=sha256:d28348bc7804748e252faa498e894e524df742c5f3928e69357539b004c5e8e7

Observation f6285648-cc33-4b02-8ec7-0d7bc345be58 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Deep unsupervised learning using nonequilibrium thermodynamics

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:33.258476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:33.258476Z digest=sha256:9b401280ff40c7aaf0abc73599be65a2bcb51cfd68569d96f5ed7a8692bf253c

Observation 8a61772a-7f49-4d2c-b904-cb72350e8f88 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Score-Based Generative Modeling through Stochastic Differential Equations

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:33.344753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:33.344753Z digest=sha256:b83bfe15d309f1a57006371141805bd434b57af5f5b7dd6cb19510ab28d45abb

Observation 4d5a9aed-d76c-4689-9b42-cb70207d4d25 · outbound

This paper cites UniAvatar: Taming Lifelike Audio-Driven Talking Head Generation with Comprehensive Motion and Lighting Control.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync UniAvatar: Taming Lifelike Audio-Driven Talking Head Generation with Comprehensive Motion and Lighting Control

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:33.440194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:33.440194Z digest=sha256:b767b81b1592db868f42b69eedc7cd4888d942917822bdfbba1f700601d5c6f8

Observation e030ceec-abb6-462c-bf73-a8f6a3641bd0 · outbound

This paper cites Diffposetalk: Speech-driven stylistic 3d facial animation and head pose generation via diffusion models.ACM Transactions on Graphics (TOG), 43(4): 1–9, 2024.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Diffposetalk: Speech-driven stylistic 3d facial animation and head pose generation via diffusion models.ACM Transactions on Graphics (TOG), 43(4): 1–9, 2024

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:43.329796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:33.563374Z digest=sha256:5a03537f8b98823cbc9f207c054be9bf827278081eb78ef80a9f49ecaa6310f6

Observation c74ae494-b109-4b0d-ada8-38ea9c7732ed · outbound

This paper cites Mofa: Model-based deep convolutional face autoencoder for unsupervised monocular reconstruction.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Mofa: Model-based deep convolutional face autoencoder for unsupervised monocular reconstruction

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:43.026476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:33.671556Z digest=sha256:a12450e1681542394bc159fab5a54aafc4587d2b7b3e52917c612a13352668d1

Observation 4c3ffe1a-c906-4fa6-b52a-69d3472ea9fd · outbound

This paper cites Instance Normalization: The Missing Ingredient for Fast Stylization.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Instance Normalization: The Missing Ingredient for Fast Stylization

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:33.789332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:33.789332Z digest=sha256:a0a1b23230963fc62c2e8f60c206c3a074a4ef16d85c418b53bc6047fe252b8b

Observation e95bcd4b-54e0-4c99-8704-59a1955c660f · outbound

This paper cites Seeing what you said: Talking face generation guided by a lip reading expert.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Seeing what you said: Talking face generation guided by a lip reading expert

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:42.762236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:33.915452Z digest=sha256:f77b54a2e368c3da53ddd2905fbbdf1bc56791fd8e690dbe5100f74fae808893

Observation cca0aa83-f00d-42b8-a2ec-60bf1bb2b0d0 · outbound

This paper cites JoyGen: Audio-Driven 3D Depth-Aware Talking-Face Video Editing.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync JoyGen: Audio-Driven 3D Depth-Aware Talking-Face Video Editing

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-06T13:39:36.934732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:34.035222Z digest=sha256:fb42563478458a5975f3edc1e94ec91ca2e3dccbed2e743ca594f78d6d3cd65d

Observation b0910bb9-7e8e-4faf-8ca2-8bc1a6c9c04c · outbound

This paper cites One-shot free-view neural talking- head synthesis for video conferencing.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync One-shot free-view neural talking- head synthesis for video conferencing

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:42.524760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:34.095342Z digest=sha256:c1125c43a945022f41245a76fae5d00161e799508ea9693bb951acdc416c5f00

Observation 800e569a-8dc8-44f6-a940-de396405586b · outbound

This paper cites Im- itating arbitrary talking style for realistic audio-driven talking face synthesis.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Im- itating arbitrary talking style for realistic audio-driven talking face synthesis

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:42.177100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:34.191233Z digest=sha256:8dd31e048df2bce968dc9efbb05d3bdf46e95ef82389809c52d288a2740f1b0f

Observation e1cd2faa-24b2-4e10-b08b-b60d51c27c4d · outbound

This paper cites Group normalization.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Group normalization

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:34.321999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:34.321999Z digest=sha256:3c442fcdd94e0772798704bdc197aebf94d37a233ed037aae8bf8afa4b5dccd3

Observation 3a0fb7a7-b975-4c4c-82b1-a21a80bac6a8 · outbound

This paper cites Vfhq: A high-quality dataset and benchmark for video face super-resolution.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Vfhq: A high-quality dataset and benchmark for video face super-resolution

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:41.956369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:34.448317Z digest=sha256:457b1ca0271421b346f848170ace76223e73c8878c88bc77bc415fceb6d46da4

Observation 5c7f96f4-0f53-472a-9c71-5ff53db2c77b · outbound

This paper cites High-fidelity generalized emotional talking face generation with multi-modal emotion space learning.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync High-fidelity generalized emotional talking face generation with multi-modal emotion space learning

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:41.706446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:34.547704Z digest=sha256:95bf228f7b32d4dc37890fa621b8d851fa95140c37db925885f43bd52dcc272f

Observation 776c542c-0ebf-4ce1-8283-319b259dc0f0 · outbound

This paper cites Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:34.642522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:34.642522Z digest=sha256:fbc09f83362849cafc0fe0a98c638baa80656dfb0413fdf0e3bf7b8108c7934e

Observation 664cfe33-f7d2-48b7-8b67-823303c094aa · outbound

This paper cites Vasa-1: Lifelike audio-driven talking faces generated in real time.Advances in Neural Information Processing Systems, 37: 660–684, 2025.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Vasa-1: Lifelike audio-driven talking faces generated in real time.Advances in Neural Information Processing Systems, 37: 660–684, 2025

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:41.478297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:34.736629Z digest=sha256:72c37b1785105c5a0651249c7fff87ae57eda5fb6d5cba552f8d2635c67b6d89

Observation a621fd21-2307-4fea-94fe-f4f451b6878f · outbound

This paper cites Real3D-Portrait: One-shot Realistic 3D Talking Portrait Synthesis.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Real3D-Portrait: One-shot Realistic 3D Talking Portrait Synthesis

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:34.794532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:34.794532Z digest=sha256:0bca9ee73d7086f23e0ab6375b1c49604961a96098d9b171cb939f3b239726ca

Observation d52bddd0-e699-4fc2-a245-258d573e19b4 · outbound

This paper cites Dynamic Neural Textures: Generating Talking-Face Videos with Continuously Controllable Expressions.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Dynamic Neural Textures: Generating Talking-Face Videos with Continuously Controllable Expressions

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-06T13:39:36.594730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:34.839399Z digest=sha256:036ea4412053fa66034fd9ca97cf214872103663ff8397a9699f8bce0f8736a9

Observation 63ba9994-1cfe-4f93-b806-b72b859aa270 · outbound

This paper cites Audio-driven Talking Face Video Generation with Learning-based Personalized Head Pose.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Audio-driven Talking Face Video Generation with Learning-based Personalized Head Pose

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:34.940981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:34.940981Z digest=sha256:15061d3d42d285c60e0a7f74fe36dc0320e1b3f8139b2fc6871f5f2bd1415eec

Observation 4157b59f-bfa4-4588-9252-5a8b327f1495 · outbound

This paper cites Face animation with an attribute-guided diffusion model.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Face animation with an attribute-guided diffusion model

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:41.262334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:35.033457Z digest=sha256:ea49604c2bb4a353459eca270679e62ef4f5968cffb3ae6ea0689fba7f940670

Observation 9bac0f53-df88-4098-a062-d3d80a038363 · outbound

This paper cites Facial: Synthesizing dynamic talking face with implicit attribute learning.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Facial: Synthesizing dynamic talking face with implicit attribute learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:41.002242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:35.108427Z digest=sha256:50e79bbbb69373f5996a2c0effa74cb37ee670613757bcbd64f7dcd5a73bb7d5

Observation 0422318c-2edd-4c9b-a99b-19dd9e1bdd53 · outbound

This paper cites Refa: Real-time egocentric facial animations for virtual reality.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Refa: Real-time egocentric facial animations for virtual reality

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:40.674745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:35.202295Z digest=sha256:794ada044b881522f2bf840f706e3cea14d388f13d9fd20d1bc02eb56197352c

Observation 879dae3c-ae67-4888-b272-598b973b81a8 · outbound

This paper cites Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:40.308738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:35.285388Z digest=sha256:deb2aa63e8b77b61b58490194606cc02b05b03b397713ce4c3cbc31acc574d42

Observation 9811c709-e290-485d-8a75-120a73b6c8ff · outbound

This paper cites MuseTalk: Real-Time High-Fidelity Video Dubbing via Spatio-Temporal Sampling.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync MuseTalk: Real-Time High-Fidelity Video Dubbing via Spatio-Temporal Sampling

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:35.346930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:35.346930Z digest=sha256:622d04821e4cf400e16833efc71223f94fc29e88193be078ee9fff0bce43a4a2

Observation 8b9d950b-6414-4087-a251-91625d824470 · outbound

This paper cites Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:40.054748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:35.424448Z digest=sha256:4405ccac6aac63668f41cd2226ced6da55265f888c9bea5a9be4b881d9f8dcf0

Observation c020f487-236e-43b8-a64f-a0949742b08c · outbound

This paper cites Dinet: Deformation inpainting network for realistic face visually dubbing on high res- olution video.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Dinet: Deformation inpainting network for realistic face visually dubbing on high res- olution video

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:39.824124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:35.481510Z digest=sha256:a1be34fe08e7c8a543f3f338982b35f49a3b2fa367e8fd04faf980fa74125503

Observation b6829569-aedf-4c87-bd57-1b7fd6d8b006 · outbound

This paper cites Avid: Any-length video inpainting with diffusion model.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Avid: Any-length video inpainting with diffusion model

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:39.579572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:35.584749Z digest=sha256:0751e7a4f2fc42efb60c90f25bae5dc211dad2ee6aefd9746fd6b3325c15484d

Observation b4b5f0c6-1f7b-416b-a740-3c14836a54bc · outbound

This paper cites General facial representation learn- ing in a visual-linguistic manner.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync General facial representation learn- ing in a visual-linguistic manner

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:39.366507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:35.672562Z digest=sha256:75c34ba30d8b6be09fdc131593cfaf17144175ef8aeb4d2d1185283e4fe9db53

Observation db3a4d5b-bec0-4ead-afda-b9904fdd7efc · outbound

This paper cites Identity-preserving talking face generation with landmark and appearance priors.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Identity-preserving talking face generation with landmark and appearance priors

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:39.125796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:35.695632Z digest=sha256:f575f7f74a3c32dc9554735299f011bcc50a91672f8fcecc629fcf2a5a3435fa

Observation 30928c61-73dd-4168-b5f1-80bb44a3e58a · outbound

This paper cites On the continuity of rotation representations in neural networks.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync On the continuity of rotation representations in neural networks

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:38.914745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:35.804750Z digest=sha256:c0f2f462d2d2f908ceb73ebaac10862c211e42a21b7ddff907ebca06b1947003

Observation 16350022-bbee-41f7-8a26-41640f31ec27 · outbound

This paper cites Celebv-hq: A large-scale video facial attributes dataset.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Celebv-hq: A large-scale video facial attributes dataset

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:35.884744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:35.884744Z digest=sha256:2fd64a5852f35b5707c7dbc98499314a8b0e92f9063ae54244eba27924c1e5c6

Observation 6ddcae72-4a54-4602-addd-1572a4f83b33 · outbound

This paper cites Face alignment across large poses: A 3d solution.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Face alignment across large poses: A 3d solution

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:38.679177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:35.995606Z digest=sha256:241da05ed458154804dc37fc8dc838173c80c1b7484638e557f613eb30b014f5

Observation fa695015-4b98-413e-adea-91038e6223b8 · outbound

This paper cites We also include identity consistency lossL id =∥α 1 −α 2∥2 2, where 1 and 2 indicate identity param- eters extracted from the same video but from different frames.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync We also include identity consistency lossL id =∥α 1 −α 2∥2 2, where 1 and 2 indicate identity param- eters extracted from the same video but from different frames

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:38.329313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:36.109562Z digest=sha256:4a31bc84a83e84b006bf7d24a60f10979f0d93a8afca2b95640f4df97755ecf1

Observation 3b55c148-1b18-4996-b193-375f349ce86a · outbound

This paper cites Recall from Sec.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Recall from Sec

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:38.046610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T13:39:36.267401Z digest=sha256:cfce2ec2b6ed607969d5352b600fc0e24c0f0bfe7bf08bf86821e3eabe1d0a40

Pith citing papers

No inbound Pith citation observations are available.