Pith. sign in

Paper Citation Record · LEDGER

V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2406.02511.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.02511 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T21:11:25.331488Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T21:46:15.589296Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 097ae7fe-22e1-4922-9e50-080a7471ffe1 · inbound

Long-Term TalkingFace Generation via Motion-Prior Conditional Diffusion Model cites this paper.

Long-Term TalkingFace Generation via Motion-Prior Conditional Diffusion Model V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T21:11:25.331488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:11:25.331488Z digest=sha256:445016f4ab6f66e14d773528cea49ec617927d8c976dd66f1b63144610601a8a

Observation b23639c4-e012-4a26-938e-9e2c4ca63775 · inbound

HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters cites this paper.

HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:02:29.417057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:02:29.417057Z digest=sha256:01486c993c53479963ee0dd0177f5d91053ac6df60265e621fc857191500c14c

Observation 74043573-0dc9-4a75-ae0e-87a21f8a5153 · inbound

Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation cites this paper.

Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:07:33.665330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:07:33.665330Z digest=sha256:ef14193f95a095a8acf5417651a132589f6d5b4e89dbf068fc7d2097b496bd9f

Observation d072b7de-f82f-4ae4-b929-62e7ced42bc0 · inbound

SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers cites this paper.

SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.482037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.482037Z digest=sha256:e5db955795856c81195432101396b380ba8b6d1bf77b065ecd19b56442cffa08

Observation ee190d78-90bf-4f79-ad74-d0c88e692da8 · inbound

FixTalk: Taming Identity Leakage for High-Quality Talking Head Generation in Extreme Cases cites this paper.

FixTalk: Taming Identity Leakage for High-Quality Talking Head Generation in Extreme Cases V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T20:58:29.324951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:58:29.324951Z digest=sha256:a5563e3361a3d3b9d4d7659f075c5f6accee8a4c643da76abd6129984d5a42fa

Observation 500c702c-1c46-4edf-a9cb-f9d47c11dd2e · inbound

FashionPose: Unified Text-Driven Fashion Synthesis with Joint Geometric and Photometric Control cites this paper.

FashionPose: Unified Text-Driven Fashion Synthesis with Joint Geometric and Photometric Control V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T16:32:42.926266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:32:42.926266Z digest=sha256:669723900594472b807ea009f0058779e4fb69a988144402a621db8e8842484b

Observation d3878481-826b-487f-ba06-d17730e39e9b · inbound

X-NeMo: Expressive Neural Motion Reenactment via Disentangled Latent Attention cites this paper.

X-NeMo: Expressive Neural Motion Reenactment via Disentangled Latent Attention V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T11:03:44.881560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:03:44.881560Z digest=sha256:15adafcc773f4870fe6867a3296880dcf780f0373237956840a70114eeb18cb6

Observation 7ce530bd-40aa-469e-adc9-d68d084b7a8a · inbound

Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering cites this paper.

Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T05:05:48.759306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:05:48.759306Z digest=sha256:a330248d67586f66378210e766f5be36361ada7c7bd47eb0cff140a88235bb11

Observation 272fa9aa-e990-4612-add2-f17b4de8a0d7 · inbound

InfinityHuman: Towards Long-Term Audio-Driven Human cites this paper.

InfinityHuman: Towards Long-Term Audio-Driven Human V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.223978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.223978Z digest=sha256:da16fe9eb817150900279135a954a6e99a61a8b5e379f4c464310e3c4aeb2f96

Observation 26f65964-2bed-4c94-90f5-35971cda52c0 · inbound

Human Motion Video Generation: A Survey cites this paper.

Human Motion Video Generation: A Survey V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 159

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:57.399032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:57.399032Z digest=sha256:f1c2dc9a4bbd2503fab9afefb63ce278ce938be19374b497053140899a4e67e3

Observation 4ee12298-942c-463a-a180-5e511c6bfef6 · inbound

FluentAvatar: Flicker-Free Talking-Head Animation via Phoneme-Guided Autoregressive Modeling cites this paper.

FluentAvatar: Flicker-Free Talking-Head Animation via Phoneme-Guided Autoregressive Modeling V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:42:43.773689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T16:42:25.803856Z digest=sha256:d02f1ded6c43edd817811c481beede88ef76a6629b1979746bb3533af3d59745

Observation 61b7f1b7-40a3-48d5-9cff-3319c4fa43aa · inbound

STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits cites this paper.

STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-03T16:30:38.891584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:30:38.891584Z digest=sha256:62ed6adc5419bcc13a00cf62080801d415f5944b4aa2f786caed9c5d774491fd

Observation 91c439cf-9570-4587-a110-7225295da789 · inbound

Instant Expressive Gaussian Head Avatars at Over 100 FPS cites this paper.

Instant Expressive Gaussian Head Avatars at Over 100 FPS V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-03T15:28:58.184824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:28:58.184824Z digest=sha256:aa0f2b06d1ab34225a5737349b816a727ae4397290867d44467e9de21145e6ff

Observation 2edf10ef-7fcf-4e7e-a7bf-b6a5bdedf3e2 · inbound

VDCook:DIY video data cook your MLLMs cites this paper.

VDCook:DIY video data cook your MLLMs V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:56:19.200428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T16:52:36.882218Z digest=sha256:5a83476952caf8d9ef37863349f31a256d73f4ed52cc02a7a005a8ebf7818c6b

Observation 5ebed5a6-56b1-4d79-a7ba-9720bf97db99 · inbound

Image-to-Video Diffusion: From Foundations to Open Frontiers cites this paper.

Image-to-Video Diffusion: From Foundations to Open Frontiers V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 166

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:08:25.093913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T15:06:02.084336Z digest=sha256:8b5dfa51ab68348d664c0c5baf7d47c3804c705d7bc51e5d0129301db6769469

Observation 8eee60c2-68fa-4ddf-9180-6240558ab7d4 · inbound

Loki: Representation over Architecture for Diffusion-Based Portrait Animation cites this paper.

Loki: Representation over Architecture for Diffusion-Based Portrait Animation V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:24:49.833461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T15:23:48.899214Z digest=sha256:cd727d04e5a785b1da1cc62fd940979a988088af5dc4b85657417aa3c0dc7577

Observation ed753058-b32e-42a7-8723-eae14907f57e · inbound

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation cites this paper.

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T21:46:15.590782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T16:19:55.048831Z digest=sha256:7db147ad1869612c991e4556d678a2dfe13beb0bd43ed2effd2eb174a8c4c3e1

Observation fbdaa9bc-95a3-44cf-8aab-da822f545909 · inbound

Geometry-guided Emotion Modulation for Controllable and Photorealistic Emotional Talking Face Generation cites this paper.

Geometry-guided Emotion Modulation for Controllable and Photorealistic Emotional Talking Face Generation V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T01:03:11.743959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:03:11.743959Z digest=sha256:6e755ca51f4f08e2a7f5354cce5f0d083250366aec547c5df660c9874d496dd8