Pith. sign in

Paper Citation Record · LEDGER

Human Motion Video Generation: A Survey

As of 17 August 2026, this Paper Citation Record lists 100 of 228 outbound references and 0 inbound Pith citation observations for arXiv:2509.03883.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03883 v1

Coverage vector

measured 100 of 228 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:36:55.691493Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 228 outbound references displayed

  • verified exact9
  • verified fuzzy0
  • unresolved91
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f95da5f0-5166-4cbc-91b5-4003ab12495c · outbound

This paper cites Deep video portraits,.

Human Motion Video Generation: A Survey Deep video portraits,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:48.887967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:48.887967Z digest=sha256:800a24522dc8dfd7839ade69b366c40cce39e02ae5e51de393e9ad824ba4cace

Observation a24baf39-178c-4360-adda-a78ea16f0c88 · outbound

This paper cites Faceformer:Speech- driven 3d facial animation with transformers,.

Human Motion Video Generation: A Survey Faceformer:Speech- driven 3d facial animation with transformers,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:48.970520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:48.970520Z digest=sha256:f6489acada3c1a53b15b94403881ad20bdd5654650da5f525fd1616d7f0c3c34

Observation cd7f49ad-3ec6-4700-bd27-747d8cefa3a8 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Human Motion Video Generation: A Survey Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.092801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.092801Z digest=sha256:e6e0a80f649085535a0f8712cb4718b35152289c0aa9b6875754621d24a678de

Observation 26813512-657e-4181-b184-6d7aa6204fa2 · outbound

This paper cites Animatediff: Animate your personalized text-to- image diffusion models without specific tuning,.

Human Motion Video Generation: A Survey Animatediff: Animate your personalized text-to- image diffusion models without specific tuning,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.159879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.159879Z digest=sha256:d09c621b372ce4c2ed71aef07be38d0301376dc6a6998088320a5dac3623dec5

Observation 94db31dc-7003-4a00-8710-6af5c5cab0aa · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Human Motion Video Generation: A Survey Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.239210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.239210Z digest=sha256:d59fde46a49c22f7f12098f5364cc6b532015608d939c9133b396df3198370b2

Observation 7b169ad2-017c-4d24-8270-8e3568cb2bfa · outbound

This paper cites Faces that speak: Jointly synthesising talking face and speech from text,.

Human Motion Video Generation: A Survey Faces that speak: Jointly synthesising talking face and speech from text,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.318463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.318463Z digest=sha256:538f123c27e93f80e4dee5fd91d2ad6df6a788b78017e1487d6d0e1bea838be3

Observation 1cef78bd-e8c1-4cf2-9800-3447126b7a23 · outbound

This paper cites MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion.

Human Motion Video Generation: A Survey MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.374447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.374447Z digest=sha256:6237255a51b8f1329e9671093a37f7e0fe13e757afd6af3a9bd89746571a1f2a

Observation d40138a0-a809-478d-b4d4-c46d4ed8ac14 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view synthesis,.

Human Motion Video Generation: A Survey Nerf: Representing scenes as neural radiance fields for view synthesis,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.451505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.451505Z digest=sha256:9bcf52cddd05db97633920967d5aa6478ff68d826718a8ab23192d8052ed17b2

Observation 7b7fd77b-ed8b-4cef-9f87-1a6789fa5f00 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering,.

Human Motion Video Generation: A Survey 3d gaussian splatting for real-time radiance field rendering,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.513818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.513818Z digest=sha256:dcf38f921db8075a3619f4037b772ea00e0e3e137523b332106bc606fc3a1f5e

Observation 7eeb9952-bb2b-4d3b-b5c7-7f1c41679862 · outbound

This paper cites Deep person generation: A survey from the perspective of face, pose, and cloth synthesis,.

Human Motion Video Generation: A Survey Deep person generation: A survey from the perspective of face, pose, and cloth synthesis,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.591482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.591482Z digest=sha256:9d14747ed05196136a324f0168684063865487be49d8b767bf990b23219d1bf1

Observation 294a4644-7e29-4b56-846e-aeb8949c0813 · outbound

This paper cites A Comprehensive Survey on Human Video Generation: Challenges, Methods, and Insights.

Human Motion Video Generation: A Survey A Comprehensive Survey on Human Video Generation: Challenges, Methods, and Insights

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.645540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.645540Z digest=sha256:e5aae3037a69c8982634f72e86018f1678ae1d28c25ace47485c59cb2166a2d4

Observation b01d6abe-af0b-4308-b541-213e8c2377e1 · outbound

This paper cites Image-based virtual try-on: A survey,.

Human Motion Video Generation: A Survey Image-based virtual try-on: A survey,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.742383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.742383Z digest=sha256:ccc504e400fadf651a11481f811e9bc4a7654d28a48963795e3f830a8906a43e

Observation 52958468-5877-4275-ae6e-7b9dcd5dec0d · outbound

This paper cites A Comprehensive Taxonomy and Analysis of Talking Head Synthesis: Techniques for Portrait Generation, Driving Mechanisms, and Editing.

Human Motion Video Generation: A Survey A Comprehensive Taxonomy and Analysis of Talking Head Synthesis: Techniques for Portrait Generation, Driving Mechanisms, and Editing

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.954166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T10:36:49.794534Z digest=sha256:572b1fbe403f91e563bd862b5f2ce0eec9a5de75e40face55cb42381e199e09e

Observation f829e0b9-2e0e-4499-bada-7ba5bb7aadbc · outbound

This paper cites Multilingual video dubbing—a technology review and current challenges,.

Human Motion Video Generation: A Survey Multilingual video dubbing—a technology review and current challenges,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.872986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.872986Z digest=sha256:04e36dd40cb77f992cedf1284c92760f5042742a284bca521414512cb8a8dddd

Observation b3705933-00bd-4c0b-8dea-af99e0cac63f · outbound

This paper cites Difftalk: Crafting diffusion models for generalized audio-driven portraits anima- tion,.

Human Motion Video Generation: A Survey Difftalk: Crafting diffusion models for generalized audio-driven portraits anima- tion,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.948773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.948773Z digest=sha256:7f01a1ce19457ea5685e050a32c31493e47f2119edaf8fe78bd217d1e2c84edd

Observation 20c57b9c-547e-4543-ab8f-cb9dbbbb4769 · outbound

This paper cites Identity-preserving talking face generation with landmark and appear- ance priors,.

Human Motion Video Generation: A Survey Identity-preserving talking face generation with landmark and appear- ance priors,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.025023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.025023Z digest=sha256:2bd8a43b9f5e4afd9329b7961670b8b671034b6ef0f9b9a3baeddcae4a5efc01

Observation 1b2b5664-056c-4e89-9ed6-3158cdefa1af · outbound

This paper cites Affective Faces for Goal-Driven Dyadic Communication.

Human Motion Video Generation: A Survey Affective Faces for Goal-Driven Dyadic Communication

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.075167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.075167Z digest=sha256:9329f0689b8d75d1eae981d2ac6917df8bc1db5176053e9f213e9f2d5a3411d2

Observation 772294c5-ddd2-4528-8f9f-a8cd4e122bd3 · outbound

This paper cites AgentAvatar: Disentangling Planning, Driving and Rendering for Photorealistic Avatar Agents.

Human Motion Video Generation: A Survey AgentAvatar: Disentangling Planning, Driving and Rendering for Photorealistic Avatar Agents

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.913967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T10:36:50.144987Z digest=sha256:28ba4f0f2e7d89bbd1018fda6c2ac6f24df5804cf1cb04d3b80fe03d3de3c525

Observation 9a49104e-7b28-46f4-b392-fc81ff03714d · outbound

This paper cites Instructavatar: Text-guided emotion and motion control for avatar generation,.

Human Motion Video Generation: A Survey Instructavatar: Text-guided emotion and motion control for avatar generation,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.226236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.226236Z digest=sha256:24fb70d7609599af93f91773112e4d993c3cecb72753204984859789d3359a4d

Observation 8b556212-d124-413a-80fe-02527b13bd9c · outbound

This paper cites Human motion generation: A survey,.

Human Motion Video Generation: A Survey Human motion generation: A survey,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.287757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.287757Z digest=sha256:1879c8fe8f7c7b600f859d31a72f2573362e71b35e18b425d4d9afc53a9c93d9

Observation 4f36bf34-f59a-48cb-9bbc-a03cf0aceac9 · outbound

This paper cites A survey of talking-head generation technology and its applications,.

Human Motion Video Generation: A Survey A survey of talking-head generation technology and its applications,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.350825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.350825Z digest=sha256:a1affb9871462981192276bef604a6e6bc748ad069fbb9fc7986af79a318e39b

Observation fe0bf50d-710e-4c45-bd0a-8ecb171f4c50 · outbound

This paper cites Unsupervised high-resolution portrait gaze correction and animation,.

Human Motion Video Generation: A Survey Unsupervised high-resolution portrait gaze correction and animation,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.427366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.427366Z digest=sha256:ad00f96bae7a612b0d9e060b67764798e93511197a6965c4ffdd41f43221c48c

Observation 96978db2-617b-45c6-a88c-2b9dcd7adace · outbound

This paper cites Expression domain translation network for cross-domain head reenactment,.

Human Motion Video Generation: A Survey Expression domain translation network for cross-domain head reenactment,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.506521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.506521Z digest=sha256:b88c78aa10585ca4f7d9ac0c7872d88dedca883c18d6ada398d6941d0a67a863

Observation 15cb19b0-2feb-468d-917d-09e9e1f2921c · outbound

This paper cites Otavatar: One-shot talking face avatar with controllable tri-plane rendering,.

Human Motion Video Generation: A Survey Otavatar: One-shot talking face avatar with controllable tri-plane rendering,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.593467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.593467Z digest=sha256:5a11e2265f8c6c104eb59b56968bce7a39b0d7045793d4e644a0d5575bf5eab9

Observation 61e15e95-e4b3-4a44-9078-783033ffbf80 · outbound

This paper cites Follow-your-emoji: Fine-controllable and expressive freestyle portrait animation,.

Human Motion Video Generation: A Survey Follow-your-emoji: Fine-controllable and expressive freestyle portrait animation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.663654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.663654Z digest=sha256:1061adbdea3070768f5218d785abafd80fa6639d874190616bb04e5508aff328

Observation 32360012-a949-4738-80f6-3f8eadb82ceb · outbound

This paper cites LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control.

Human Motion Video Generation: A Survey LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.724953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.724953Z digest=sha256:ba2d8a7961e3125a3749b9f830d1f5c077957637c06b9ebe1b129fcd0f0a67b9

Observation 56a2bb31-334d-4406-9232-280ad738af38 · outbound

This paper cites X-portrait: Expressive portrait animation with hierarchical motion attention,.

Human Motion Video Generation: A Survey X-portrait: Expressive portrait animation with hierarchical motion attention,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.800840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.800840Z digest=sha256:af030279b99ac19b18bc361a99347a795284bb2358ba49e12a4a74641f949ae2

Observation 83530f0c-6aad-4473-b9ad-45305848293f · outbound

This paper cites MobilePortrait: Real-Time One-Shot Neural Head Avatars on Mobile Devices.

Human Motion Video Generation: A Survey MobilePortrait: Real-Time One-Shot Neural Head Avatars on Mobile Devices

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.872770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T10:36:50.885066Z digest=sha256:aa6589833a28063e3a30823b0eea01499ce06955a9fed000c5255a5c8962aaf7

Observation c0ff62f4-6f49-463f-8945-49c096b8a056 · outbound

This paper cites Everybody dance now,.

Human Motion Video Generation: A Survey Everybody dance now,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.958740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.958740Z digest=sha256:7ec6fe36b072e7ce701b42023a71ded730e4fa963b2b941f6efe699f9da37d43

Observation c22c8b7c-2ad8-42fb-9342-4e3d98fb9a9c · outbound

This paper cites Human motionformer: Transferring human motions with vision transformers,.

Human Motion Video Generation: A Survey Human motionformer: Transferring human motions with vision transformers,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.046502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.046502Z digest=sha256:d6d51539289b2e29c39634a5be1b302714c1ee0c66aacde728be5c666d239bb3

Observation 36ec9f0e-890f-45dd-9bf3-6d155fb9f32c · outbound

This paper cites Bidirectional temporal diffusion model for temporally consistent human animation,.

Human Motion Video Generation: A Survey Bidirectional temporal diffusion model for temporally consistent human animation,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.122153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.122153Z digest=sha256:eb6a8736e054d67334bf787b2504558fa7718ee00f41a0a83e92123c79e65122

Observation 376dcc8e-92eb-48a6-8726-8ea2a0052aa4 · outbound

This paper cites Disco: Disentangled control for realistic human dance generation,.

Human Motion Video Generation: A Survey Disco: Disentangled control for realistic human dance generation,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.184832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.184832Z digest=sha256:8371d5e623a1675b62747363313d926bccb6fdabda3d6f1cedb37058041546f2

Observation 0e943829-ccb7-47e3-93fa-a4f5586821f7 · outbound

This paper cites Animate anyone: Consistent and controllable image-to-video synthesis for character animation,.

Human Motion Video Generation: A Survey Animate anyone: Consistent and controllable image-to-video synthesis for character animation,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.227878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.227878Z digest=sha256:2e56203e87639aa159a37826d6a6814d224583b563fe2d13946fe9ee87357e4e

Observation 3362df08-e56c-40f9-82cb-c6b2badd3684 · outbound

This paper cites Towards Multiple Character Image Animation Through Enhancing Implicit Decoupling.

Human Motion Video Generation: A Survey Towards Multiple Character Image Animation Through Enhancing Implicit Decoupling

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.276195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.276195Z digest=sha256:fcf7013501be5ed82767aeca6912ab55a5fcba436c9e824ae78cc8446318964d

Observation fd85064b-470f-4b74-a4e5-6886a46b3cf6 · outbound

This paper cites Human4DiT: 360-degree Human Video Generation with 4D Diffusion Transformer.

Human Motion Video Generation: A Survey Human4DiT: 360-degree Human Video Generation with 4D Diffusion Transformer

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.336374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.336374Z digest=sha256:eaf3c138fd6b3128f1559f1515d41105b28d066fe48ebe31de568f17d7b3753b

Observation 2b9aff82-d393-46fd-87cf-80804041afcd · outbound

This paper cites MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance.

Human Motion Video Generation: A Survey MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.408510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.408510Z digest=sha256:1649e81e92f413961b512fac057c66be4b795c4feb71f04fc530aae1ffc33674

Observation c63e07aa-2336-4688-a64f-0b50089fa4a4 · outbound

This paper cites I2v-adapter: A general image-to-video adapter for diffusion models,.

Human Motion Video Generation: A Survey I2v-adapter: A general image-to-video adapter for diffusion models,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.491529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.491529Z digest=sha256:54d14bd552d7d0c9ab75c4d2ebb71b15597ff79bb1cceffbc681afbcc38baaa8

Observation fa9be2f3-e5e9-410a-b530-6a312ebe37a2 · outbound

This paper cites ViViD: Video Virtual Try-on using Diffusion Models.

Human Motion Video Generation: A Survey ViViD: Video Virtual Try-on using Diffusion Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.546340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.546340Z digest=sha256:4ced80bb5d083b72b87266d2231e4ccb05b6fe65114e1ddc061ebda64729aa9c

Observation d2cc0be8-39bb-45d9-9c37-b4ecd5fd78aa · outbound

This paper cites Dreampose: Fashion image-to-video synthesis via stable diffusion,.

Human Motion Video Generation: A Survey Dreampose: Fashion image-to-video synthesis via stable diffusion,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.631647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.631647Z digest=sha256:4f07b0c10c560c6d5722de5e186eb57f8d886d09f42715562480c541bc1e72a3

Observation 9d8ee456-6b82-45e0-ad5d-8ea8e26f1ef1 · outbound

This paper cites Make-your-anchor: A diffusion-based 2d avatar generation frame- work,.

Human Motion Video Generation: A Survey Make-your-anchor: A diffusion-based 2d avatar generation frame- work,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.697233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.697233Z digest=sha256:2848916e7314adcfee9d0201a9dd70671be8f7374cda125d4202d02316dc7756

Observation 427b93f9-4609-4dbb-9155-50f757a47e9d · outbound

This paper cites Write-a-speaker: Text-based emotional and rhythmic talking-head gen- eration,.

Human Motion Video Generation: A Survey Write-a-speaker: Text-based emotional and rhythmic talking-head gen- eration,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.768275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.768275Z digest=sha256:d306675cad8371243b593e7230994cc7d352af585be273758fb8a32a43484176

Observation 5039d190-f301-4c72-9d74-ef59776b9279 · outbound

This paper cites ID-Animator: Zero-Shot Identity-Preserving Human Video Generation.

Human Motion Video Generation: A Survey ID-Animator: Zero-Shot Identity-Preserving Human Video Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.833596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.833596Z digest=sha256:8781779b6c059c42f23700619da607a25eef5d00fa09e498d59e10e8e46f6d7c

Observation 26db761e-aec1-4493-a7a7-cda9aa260e5c · outbound

This paper cites Edit-Your-Motion: Space-Time Diffusion Decoupling Learning for Video Motion Editing.

Human Motion Video Generation: A Survey Edit-Your-Motion: Space-Time Diffusion Decoupling Learning for Video Motion Editing

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.888269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.888269Z digest=sha256:d00b135f32cf795d419a87905a23bb01e22c6fb70bd4b309a29dd98c4c4d748d

Observation d555322f-e779-41da-9a18-576579a9c741 · outbound

This paper cites Follow your pose: Pose-guided text-to-video generation using pose- free videos,.

Human Motion Video Generation: A Survey Follow your pose: Pose-guided text-to-video generation using pose- free videos,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.996370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.996370Z digest=sha256:cf3badd6adc7e7c2848d6b55d78319d30a932dfe72b2969fc80347e1b7644286

Observation 9cc1b22d-2e2c-43b6-b574-ba33f0834aa2 · outbound

This paper cites Text2performer: Text-driven human video generation,.

Human Motion Video Generation: A Survey Text2performer: Text-driven human video generation,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.049764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.049764Z digest=sha256:87faaae63489a6a150a5ac4112a6e3f6b85205aa83a383a095e1aef5869356b2

Observation bbe3cd7c-9b7f-48d4-8566-34a95489be60 · outbound

This paper cites Styleheat: One-shot high-resolution editable talking face generation via pre-trained stylegan,.

Human Motion Video Generation: A Survey Styleheat: One-shot high-resolution editable talking face generation via pre-trained stylegan,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.111709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.111709Z digest=sha256:e2302e911cba3a66260e92ab3bdda10950f15f6dba77fe095c01cbd5ad2de9ec

Observation 6993dca4-834c-4cac-bfaf-32c063d5f9b6 · outbound

This paper cites Pose- controllable talking face generation by implicitly modularized audio- visual representation,.

Human Motion Video Generation: A Survey Pose- controllable talking face generation by implicitly modularized audio- visual representation,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.290746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.290746Z digest=sha256:f8433ebf7004b6a02a5a320daa495d3e5180a9e5ffe14dae5e548a724f52f2f6

Observation 09be9cb8-1621-43ec-9083-065b215e6636 · outbound

This paper cites Edtalk: Efficient disentanglement for emotional talking head synthesis,.

Human Motion Video Generation: A Survey Edtalk: Efficient disentanglement for emotional talking head synthesis,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.375519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.375519Z digest=sha256:b94b986ec9c1482e247374d30efd2f708539d963933ed2931fe28be889401745

Observation e47c006d-73d1-47e7-8e90-e6c2ecf9bc31 · outbound

This paper cites EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions.

Human Motion Video Generation: A Survey EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.470098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.470098Z digest=sha256:8a2d9f4adecf7852540aa39627026d09d2e0d7b41ab5fc88f094e064f6fd6ac0

Observation 908e1770-efa3-4e48-8a70-5ecf9d94803f · outbound

This paper cites Emo: Emote portrait alive generating expressive portrait videos with audio2video diffusion model under weak conditions,.

Human Motion Video Generation: A Survey Emo: Emote portrait alive generating expressive portrait videos with audio2video diffusion model under weak conditions,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.561850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.561850Z digest=sha256:63899d76a14202b23b326d6c8bfb04fffac6a0538ae9b334e3b98a46c63f2dc9

Observation e0f06b7c-dffa-4f88-a39b-534522cb11ef · outbound

This paper cites Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation.

Human Motion Video Generation: A Survey Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.656632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.656632Z digest=sha256:2c5abfec015c98bd5824c98ff21b291250368a35bc3cc2650bfe5d392475b220

Observation 1840df13-0112-470c-926c-ea26659a34e1 · outbound

This paper cites Emotional Conversation: Empowering Talking Faces with Cohesive Expression, Gaze and Pose Generation.

Human Motion Video Generation: A Survey Emotional Conversation: Empowering Talking Faces with Cohesive Expression, Gaze and Pose Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.721044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.721044Z digest=sha256:015eb07d7a7daa220cccc0005a406a66960faa59674486bbbcedce46518c46e8

Observation db688d56-7413-4c21-a9cb-29c65e141393 · outbound

This paper cites Makeittalk: Speaker-aware talking-head animation,.

Human Motion Video Generation: A Survey Makeittalk: Speaker-aware talking-head animation,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.772995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.772995Z digest=sha256:86fa1016b5281465f88a47b02571d9176c348f406191db6f528cc106c650282b

Observation b815adc8-8abc-4055-8865-f5e35b3ae217 · outbound

This paper cites Live speech portraits: Real-time photore- alistic talking-head animation,.

Human Motion Video Generation: A Survey Live speech portraits: Real-time photore- alistic talking-head animation,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.824523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.824523Z digest=sha256:4fbb0451be23efeeff57c2b9acb59828aa761087f7ad797bdb17cc0b0312f667

Observation 75aff3a9-89ae-4ee7-82fe-024be2d77f6e · outbound

This paper cites VLOGGER: Multimodal Diffusion for Embodied Avatar Synthesis.

Human Motion Video Generation: A Survey VLOGGER: Multimodal Diffusion for Embodied Avatar Synthesis

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.868747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.868747Z digest=sha256:ac148f9c40c2da2762b4153bd94181390f86766a764a819db07c6dc5f8e7a728

Observation e4ed9920-2db5-4442-853b-631547e6c554 · outbound

This paper cites Dance Any Beat: Blending Beats with Visuals in Dance Video Generation.

Human Motion Video Generation: A Survey Dance Any Beat: Blending Beats with Visuals in Dance Video Generation

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.680484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T10:36:52.930420Z digest=sha256:dba2bc08ed10c31952544f629e8e64bba5477a748370d01001d40e755b22e1ce

Observation 4dd2bf06-54a9-41ab-85b5-337b555a0042 · outbound

This paper cites Auto-Encoding Variational Bayes.

Human Motion Video Generation: A Survey Auto-Encoding Variational Bayes

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.986651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.986651Z digest=sha256:79766ab77cdf34e818be8bd5f4336141f6445a7c5fc9551a99be37b7f794b1cf

Observation 3083795b-dd6a-49a6-b9bb-51d2e49b28e8 · outbound

This paper cites Geneface: Generalized and high-fidelity audio-driven 3d talking face synthesis,.

Human Motion Video Generation: A Survey Geneface: Generalized and high-fidelity audio-driven 3d talking face synthesis,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.070888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.070888Z digest=sha256:90c81f8cf7ba08041d02e081a80dd2cae7146aba9ead7449e8d3e3e16d977944

Observation d79c4897-2ea4-4059-bd1f-e962667fb198 · outbound

This paper cites GeneFace++: Generalized and Stable Real-Time Audio-Driven 3D Talking Face Generation.

Human Motion Video Generation: A Survey GeneFace++: Generalized and Stable Real-Time Audio-Driven 3D Talking Face Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.123535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.123535Z digest=sha256:866c35e27c9dc0ea199cb5eeba4c9373fb9571a23587f9f40979636a2953c3d0

Observation cb5922ee-b2fb-48c4-8552-736c338f3a81 · outbound

This paper cites Neural discrete representation learning,.

Human Motion Video Generation: A Survey Neural discrete representation learning,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.184941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.184941Z digest=sha256:604575c4b1ea2984fe829e49caf1c6313e031d6e76a8f5dc74b20f0b763ae4f5

Observation b3d7e218-04ec-42f8-a14d-a7c24694e47b · outbound

This paper cites Generative adversarial nets,.

Human Motion Video Generation: A Survey Generative adversarial nets,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.226243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.226243Z digest=sha256:95faba48057f0bbca2c0f765b4ae160b7e5b5697082b811a583c959be721f504

Observation 8033ef9e-f150-46a2-af76-6b2d726a9db5 · outbound

This paper cites A style-based generator architecture for generative adversarial networks,.

Human Motion Video Generation: A Survey A style-based generator architecture for generative adversarial networks,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.303004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.303004Z digest=sha256:7898e717562d745787168aee5e15daf6d2d67248b36b41828295bd0a780c3cf3

Observation c9f97cf7-0614-4fa7-989d-6b4aee2fb06d · outbound

This paper cites Analyzing and improving the image quality of stylegan,.

Human Motion Video Generation: A Survey Analyzing and improving the image quality of stylegan,

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.394240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.394240Z digest=sha256:57066b29f78045b016208bb2423b4d089630250fee88a6e1e130c45078e446f0

Observation 85c06782-9c16-45a3-a75a-1911663f5447 · outbound

This paper cites An identity-preserved framework for human motion transfer,.

Human Motion Video Generation: A Survey An identity-preserved framework for human motion transfer,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.443233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.443233Z digest=sha256:158074c01cce068b97712f1a97519081407994eb395251c0364ce5b7ee17d90a

Observation 39d5a18e-48cc-4112-b3e6-17f8765e922c · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics,.

Human Motion Video Generation: A Survey Deep unsupervised learning using nonequilibrium thermodynamics,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.506844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.506844Z digest=sha256:2041500bb0e7ca593cb2272c9b96372fac0e4ed65ab19fc932035fb19e4c95a3

Observation 48c3a0b4-1986-412c-bde6-5acd5cac82a3 · outbound

This paper cites Improved techniques for training score-based generative models,.

Human Motion Video Generation: A Survey Improved techniques for training score-based generative models,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.572864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.572864Z digest=sha256:199f6f189c26c630fdb146ec03da672a3d26e48957dc425de48f4b80ec589983

Observation e63960fd-1b14-47fe-b658-9c84020277ed · outbound

This paper cites Improved denoising diffusion proba- bilistic models,.

Human Motion Video Generation: A Survey Improved denoising diffusion proba- bilistic models,

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.628440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.628440Z digest=sha256:696394989ee72b74a482ed3c032d71ffec705d974fba80a557f49b9309c649cf

Observation 41063b16-eb8b-410f-943a-b7554e3697c2 · outbound

This paper cites Denoising diffusion implicit mod- els,.

Human Motion Video Generation: A Survey Denoising diffusion implicit mod- els,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.693869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.693869Z digest=sha256:77a708b4859af042c48b0594d8da945e446a09ef89387fba9db499f7240422ef

Observation 23363c7a-46bb-418b-9d09-6ed5900097d3 · outbound

This paper cites Diffusion models beat gans on image synthesis,.

Human Motion Video Generation: A Survey Diffusion models beat gans on image synthesis,

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.755074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.755074Z digest=sha256:0a3106a85f3635560908e3f2e300ab146b9b72399689c20a3c58049d60808f0f

Observation cff4e806-623b-46ca-b538-4a87f6e4e325 · outbound

This paper cites A survey on generative diffusion models,.

Human Motion Video Generation: A Survey A survey on generative diffusion models,

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.789291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.789291Z digest=sha256:40bfedea08f805c8198fc777be0baf3cee49a8acba018eb4ad0aa5017f1ac50f

Observation 902e7970-5ad5-45b6-ac03-5b2b8410360d · outbound

This paper cites Denoising diffusion probabilistic models,.

Human Motion Video Generation: A Survey Denoising diffusion probabilistic models,

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.822594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.822594Z digest=sha256:15c922a82d74c009c45acf29bbf8b3a5ade795f78a4b090746284d99e4524fb1

Observation 722e2ae5-ab12-4b38-97fe-1aa5f7795f8d · outbound

This paper cites Dance Your Latents: Consistent Dance Generation through Spatial-temporal Subspace Attention Guided by Motion Flow.

Human Motion Video Generation: A Survey Dance Your Latents: Consistent Dance Generation through Spatial-temporal Subspace Attention Guided by Motion Flow

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.622748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T10:36:53.851100Z digest=sha256:88c19e7715b20b1c333cff4dff3b2b4e5c620f660358f8e16a9b3405b7ee6527

Observation 4ee471ec-042e-4a69-8fd4-a15731b82897 · outbound

This paper cites Human Modelling and Pose Estimation Overview.

Human Motion Video Generation: A Survey Human Modelling and Pose Estimation Overview

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.598466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T10:36:53.890018Z digest=sha256:ba4f020ada4bb11598da4308fc93f813316fc62dfc6c5b9e2c9c2015de7ec07e

Observation 1c81d839-0969-4e81-aed5-9c32b5da1054 · outbound

This paper cites Champ: Controllable and consistent human image animation with 3d parametric guidance,.

Human Motion Video Generation: A Survey Champ: Controllable and consistent human image animation with 3d parametric guidance,

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.932596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.932596Z digest=sha256:2be24aa7afefbcb22ce43863cc41f41a9b25dd5ea637ce22e19b91a650d6f960

Observation 3d3ca1b2-7a7e-41b2-a397-b9d0a406c75e · outbound

This paper cites Openpose: Realtime multi-person 2d pose estimation using part affinity fields,.

Human Motion Video Generation: A Survey Openpose: Realtime multi-person 2d pose estimation using part affinity fields,

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.005829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.005829Z digest=sha256:8ac6367df5005f87cff5448f20398cc9f212b7fd35d5d2b6a3372a0ec4ee8be6

Observation 00303c56-3dd2-43b6-a9c8-d02c764d3a83 · outbound

This paper cites Effective whole-body pose estimation with two-stages distillation,.

Human Motion Video Generation: A Survey Effective whole-body pose estimation with two-stages distillation,

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.073205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.073205Z digest=sha256:b192857c8f491e62225f53083a3827df80f424c3aa19c564ce3e25747d30579d

Observation fca9ccc0-861b-41cd-a53e-4046ced5ca35 · outbound

This paper cites VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation.

Human Motion Video Generation: A Survey VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.118365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.118365Z digest=sha256:c285ef05336dc69f815fccdf5a8f0a7c9ca6addd06a329822e129600fe5904cd

Observation 8cef5567-ea4f-42e2-8552-9e11eea741bb · outbound

This paper cites Magicanimate: Temporally consistent human image animation using diffusion model,.

Human Motion Video Generation: A Survey Magicanimate: Temporally consistent human image animation using diffusion model,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.172725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.172725Z digest=sha256:d7e1b4a024f94bfd0d6b409327032d34b9dc51fd11fbf5a7177965558b8326bc

Observation c0371e81-1ba4-4557-89c8-b7d49adb864e · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Human Motion Video Generation: A Survey Learning transferable visual models from natural language supervision,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.269247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.269247Z digest=sha256:b14ff2ef2a13da48fc549f930f6c60e8da94d3bd42396465e7aed7330dddfa40

Observation 53c38695-d68d-4e77-95e5-73b17a9f919f · outbound

This paper cites Conformer: Convolution- augmented transformer for speech recognition,.

Human Motion Video Generation: A Survey Conformer: Convolution- augmented transformer for speech recognition,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.305057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.305057Z digest=sha256:2f27c8c7e1660ba40f056b8fb2c99268bab1652512099fe033ec96fd72933f22

Observation 00de6092-d0b5-4c1e-b388-ef81c2fb428b · outbound

This paper cites Omniavatar: Geometry-guided controllable 3d head synthesis,.

Human Motion Video Generation: A Survey Omniavatar: Geometry-guided controllable 3d head synthesis,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.380434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.380434Z digest=sha256:b82d72289ea8f6e3c6b1a75b6c6d36b621ad02b880295eef92b5e4733f7c0a38

Observation f3e72e65-8d87-413e-a9b3-442f8a2327e3 · outbound

This paper cites MegActor: Harness the Power of Raw Video for Vivid Portrait Animation.

Human Motion Video Generation: A Survey MegActor: Harness the Power of Raw Video for Vivid Portrait Animation

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.430237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.430237Z digest=sha256:b559932d4e8fa5542612fe6e6fb5e1147b1dd031ef1e337776168df2eb89a5ba

Observation bb96ac18-cfcb-4dd1-addf-5fc27ca2a887 · outbound

This paper cites Faceoff: A video-to-video face swapping system,.

Human Motion Video Generation: A Survey Faceoff: A video-to-video face swapping system,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.497161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.497161Z digest=sha256:a5ede00c9317c47b9e87268b3ce36910b55b0fa36a4ee0274225ace5995fae4c

Observation 1031deee-83e3-44be-ae43-5bebd3d0aa85 · outbound

This paper cites Finemogen:Fine- grained spatio-temporal motion generation and editing,.

Human Motion Video Generation: A Survey Finemogen:Fine- grained spatio-temporal motion generation and editing,

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.633347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.633347Z digest=sha256:ac9c6e00bd2f13ded832c444bd0ba40a8dcdb8f55e6de09c278ab5868dad7f1b

Observation adf6c3d4-e998-4b9e-b1a5-84e6f8448b0f · outbound

This paper cites Plan, Posture and Go: Towards Open-World Text-to-Motion Generation.

Human Motion Video Generation: A Survey Plan, Posture and Go: Towards Open-World Text-to-Motion Generation

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.701074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.701074Z digest=sha256:b9d36f7af5ce2fa02f88d5b4828ef8460e59b129106399c4387aa0065d67b7c8

Observation 4d30bd36-f327-41d8-b0e1-3e1bc4a9f982 · outbound

This paper cites Avatargpt: All-in-one framework for motion understanding planning generation and beyond,.

Human Motion Video Generation: A Survey Avatargpt: All-in-one framework for motion understanding planning generation and beyond,

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.779686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.779686Z digest=sha256:4c1d32da5c17eafcdaacbc7e23ec76d05e3cd04b783899fb8cf04586dfe099e5

Observation 1c3d735b-8c4f-4ff5-996e-a0d9bad80c5c · outbound

This paper cites Motiongpt:Finetunedllmsaregeneral-purpose motion generators,.

Human Motion Video Generation: A Survey Motiongpt:Finetunedllmsaregeneral-purpose motion generators,

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.842912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.842912Z digest=sha256:7ba8c9cdead38a0bcb8f8d92917d7deef90647905ba6756db8accd1101e7efda

Observation 1f47c7cd-23b3-45ad-89d5-e377bb05ad77 · outbound

This paper cites Motionscript: Natural language descriptions for expressive 3d human motions,.

Human Motion Video Generation: A Survey Motionscript: Natural language descriptions for expressive 3d human motions,

Reference 88

Resolution
verified exact
raw_fallback, observed 2026-08-05T10:36:58.521924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T10:36:54.889171Z digest=sha256:c22b801711428070db08e640e692b355ef72758a13e28d2ae17a07522993e773

Observation 52cd2ab0-f22e-4a3b-8213-f8f36c2ea07c · outbound

This paper cites Can language models learn to listen?,.

Human Motion Video Generation: A Survey Can language models learn to listen?,

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.941478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.941478Z digest=sha256:1776b8c9dd41a6e5216cda44160856ac9c8b6f9750d62b19441f04840098208e

Observation fae80bfd-7c96-4782-88fc-93930b81ffc6 · outbound

This paper cites Intercontrol: Zero-shot human interaction generation by controlling every joint,.

Human Motion Video Generation: A Survey Intercontrol: Zero-shot human interaction generation by controlling every joint,

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.987297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.987297Z digest=sha256:46e79b372dbc0a452820b1c0140c57aa01966caedeba6718a1df6f1dfe0913a9

Observation ef22c8f3-2093-4d31-81c7-6e7635981c47 · outbound

This paper cites Digital life project: Autonomous 3d characters with social intelligence,.

Human Motion Video Generation: A Survey Digital life project: Autonomous 3d characters with social intelligence,

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.061656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.061656Z digest=sha256:34ee31d1b080edcc47b50bd73306ad2c05efd41e92a2391416664448e602161c

Observation ab34fd6c-d07c-4b82-8c22-208514e94cc1 · outbound

This paper cites Style-Preserving Lip Sync via Audio-Aware Style Reference.

Human Motion Video Generation: A Survey Style-Preserving Lip Sync via Audio-Aware Style Reference

Reference 92

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.393948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T10:36:55.147557Z digest=sha256:f866b6bf255c52c51f1e35108ff2ed9d59e8292c4a501059552f85115a3e472e

Observation cef580a5-5bc0-4f25-9d36-dc9475c1b393 · outbound

This paper cites Dae-talker: High fidelity speech-driven talking face generation with diffusion autoencoder,.

Human Motion Video Generation: A Survey Dae-talker: High fidelity speech-driven talking face generation with diffusion autoencoder,

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.219526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.219526Z digest=sha256:9fab26f90d7798947d62cf4f5f3b9378a960774a544ed5c5b2ee1091c9ff9fb5

Observation 3940f10e-57e5-4fed-9b1f-410f2e027b65 · outbound

This paper cites High-fidelity generalized emotional talking face generation with multi-modal emotion space learning,.

Human Motion Video Generation: A Survey High-fidelity generalized emotional talking face generation with multi-modal emotion space learning,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.341313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.341313Z digest=sha256:be98352e4ea3a8f42cfd94b0b58e3b0df4b716872a8c1946f8f6c591145035e1

Observation 6f19b458-6964-4118-b3e8-de5a03a081fd · outbound

This paper cites Do as i do: Pose guided human motion copy,.

Human Motion Video Generation: A Survey Do as i do: Pose guided human motion copy,

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.437623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.437623Z digest=sha256:c930ed72a389649d4e37a0e196f89085e1a5f38c2bd8bf0b97b86a4325322cdb

Observation a0884bf4-d635-46bb-87a5-e0fe8da07a45 · outbound

This paper cites Magicpose: Realistic human poses and facial expressions retargeting with identity-aware diffusion,.

Human Motion Video Generation: A Survey Magicpose: Realistic human poses and facial expressions retargeting with identity-aware diffusion,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.499737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.499737Z digest=sha256:3805fc2c62ee9a604393e35ae62eaa42950c1d2dbba127f65eb4a76ece1027f4

Observation 0aa3248d-5c1b-43a6-8876-4d88a106f95d · outbound

This paper cites DreaMoving: A Human Video Generation Framework based on Diffusion Models.

Human Motion Video Generation: A Survey DreaMoving: A Human Video Generation Framework based on Diffusion Models

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.542297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.542297Z digest=sha256:cc11aaa0333ce9ae02c8da593a92bdc521df3af8ad24e2d38c96262e5a19f70d

Observation cb9c6960-cff5-4172-adef-5badaf5e1108 · outbound

This paper cites Disentangling Foreground and Background Motion for Enhanced Realism in Human Video Generation.

Human Motion Video Generation: A Survey Disentangling Foreground and Background Motion for Enhanced Realism in Human Video Generation

Reference 98

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.352828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T10:36:55.595181Z digest=sha256:3945747a12836d1a10b092c17d3083da1b661ad2169d55cb8a5ddd58689457e0

Observation 6172c528-1170-46b3-a805-c2950208d190 · outbound

This paper cites MotionFollower: Editing Video Motion via Lightweight Score-Guided Diffusion.

Human Motion Video Generation: A Survey MotionFollower: Editing Video Motion via Lightweight Score-Guided Diffusion

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.671513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.671513Z digest=sha256:ef425431efdf61460ecfc183673d7c3f23cfddc99d0995bf519f2cd923b8ab4f

Observation 67aa83bb-1e4d-451d-8d3c-178e608f0a1e · outbound

This paper cites UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation.

Human Motion Video Generation: A Survey UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.691493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.691493Z digest=sha256:bae500da77e00623068124d02114e783bfdf6bce65687b94f945240c8305c8d3

Pith citing papers

No inbound Pith citation observations are available.