Pith. sign in

Paper Citation Record · LEDGER

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

As of 10 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 4 inbound Pith citation observations for arXiv:2501.18898.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.18898 v3

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T22:06:54.887927Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T16:24:07.533226Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T03:06:29.692095Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact0
  • verified fuzzy43
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 637d77c7-32b8-4076-8cc1-0a86675c8eb2 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets, 2023.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets, 2023

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.636490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.686959Z digest=sha256:14a0910644265b94f14ccadfc2cf8e6c62fa6b1cd4569fd3de8baaf9a646f0d6

Observation 5a082a49-65e1-458c-9e3e-62e226cec4a3 · outbound

This paper cites Nonver- bal Behaviors, Persuasion, and Credibility.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Nonver- bal Behaviors, Persuasion, and Credibility

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.626822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.690773Z digest=sha256:622f4346b4d203854bcf8b3f834e1a62160f826f2dabde2bbb5383db8fc0ff5c

Observation 6ed24e08-0c2d-4564-bd19-19e250ce5e0d · outbound

This paper cites Everybody Dance Now.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Everybody Dance Now

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.616756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.694173Z digest=sha256:bbfa2d44054fc2290ed25e1534a1ffa3754a8ea78f250b6b5caa83c94fe93d90

Observation 3bf10411-b2dd-4e33-b467-ef1915b5701c · outbound

This paper cites Enabling synergistic full-body control in prompt-based co-speech motion generation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Enabling synergistic full-body control in prompt-based co-speech motion generation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.607020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.697466Z digest=sha256:62b420f76fd09476148079d909cb38e2303f1973b9a57853f3f06d9ffa4fbaee

Observation 841495ec-0992-4e87-8bbc-4a0821c32dd3 · outbound

This paper cites DiffSHEG: A Diffusion-Based Ap- proach for Real-Time Speech-driven Holistic 3D Expression and Gesture Generation, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling DiffSHEG: A Diffusion-Based Ap- proach for Real-Time Speech-driven Holistic 3D Expression and Gesture Generation, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.700476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.700476Z digest=sha256:d1b29b176ef5144cc033e7caa4b4a9ae85e13bca19d8f2e9469e9673456956cf

Observation 30eaaba3-2f89-4121-bfb3-748b25a7f51d · outbound

This paper cites WavLM: Large-Scale Self- Supervised Pre-Training for Full Stack Speech Processing.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling WavLM: Large-Scale Self- Supervised Pre-Training for Full Stack Speech Processing

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.590944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.703474Z digest=sha256:a93e5564532a119bcc0790b69b0847e88e8faf790e51ee43b87728efd3a29996

Observation d83ab87a-ddfb-4016-9034-716a51e5215f · outbound

This paper cites MotionLCM: Real-time Controllable Motion Generation via Latent Consistency Model.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling MotionLCM: Real-time Controllable Motion Generation via Latent Consistency Model

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.706587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.706587Z digest=sha256:3501e22eef29c06480363123b863710f6a4a1852e50611577ab736b3248b0a8d

Observation 54765c55-b93f-493a-a6d8-df82b0501609 · outbound

This paper cites The Interplay Between Gesture and Speech in the Production of Referring Expressions: Investigating the Tradeoff Hypothe- sis.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling The Interplay Between Gesture and Speech in the Production of Referring Expressions: Investigating the Tradeoff Hypothe- sis

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.580515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.710168Z digest=sha256:ddb70760e7dea6970577093276c21fdbc1ebaf62199582c15856a586ec10b539

Observation aaf16f7d-bf14-4d5d-bfe0-7a3bb08be096 · outbound

This paper cites Diffusion-based co-speech gesture genera- tion using joint text and audio representation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Diffusion-based co-speech gesture genera- tion using joint text and audio representation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.570947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.713193Z digest=sha256:35103d5934a7f2dff855d0ad3f04f7982c8ecee59f58e0b2a8932afba6903c26

Observation 44f4a4f2-2708-457a-84be-272c892807b3 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.716207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.716207Z digest=sha256:28afad1c786ed9b0b7b03852bb5bdde8164a2e7fa7d3e946b3ccc2f147bb79c9

Observation f726766b-00e7-4829-a445-30510941293c · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale, 2021.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling An image is worth 16x16 words: Transformers for image recognition at scale, 2021

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.719776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.719776Z digest=sha256:7db237e6a1657c98f21b10bd65d58663631b524c9fbc992a75ec7515ccabb91c

Observation 9837a269-ab17-49c2-8f2e-7f3c5c4f49b4 · outbound

This paper cites Scaling rectified flow trans- formers for high-resolution image synthesis, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Scaling rectified flow trans- formers for high-resolution image synthesis, 2024

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.555274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.722759Z digest=sha256:56bea82b00ef64ced6e3831cfbe5a7daa507291b4120e15fdfbbf4af4135ca95

Observation c77b73ec-b354-4ef2-8416-a64ea7e21b20 · outbound

This paper cites One step diffusion via shortcut models, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling One step diffusion via shortcut models, 2024

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.545844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.725469Z digest=sha256:6b6ad6669c36cb4128c86a325b105b2b213dd928309f5cdf75957546916e2b33

Observation fdb0bb53-f0c6-4e5f-83bf-f1c4e5f6df8a · outbound

This paper cites Eraseanything: Enabling concept erasure in rectified flow transformers.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Eraseanything: Enabling concept erasure in rectified flow transformers

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.536960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.728178Z digest=sha256:134a9772aa207065594f1dbfd9b1b290ab9ea796745fc3c97d73658ac98b14be

Observation 9ed8d118-8a65-45e0-834f-004c4b279f97 · outbound

This paper cites Ginosar, A.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Ginosar, A

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.527499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.731020Z digest=sha256:6eb544de6e3a54ca28eaf052cd91d4c92f92166d1f108471a72758145fe1e7dc

Observation 81ad22ba-4613-4a10-a304-c74e5b2a7576 · outbound

This paper cites Momask: Generative masked model- ing of 3d human motions.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Momask: Generative masked model- ing of 3d human motions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.518538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.734221Z digest=sha256:ec3979eb004a5e514d89a21ab43ae9dee1720d59e01ae5a46a1192eba3a38a47

Observation d7427ef7-a4e0-45fe-8ef1-539d07cb259c · outbound

This paper cites Learning Speech-driven 3D Conversational Gestures from Video.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Learning Speech-driven 3D Conversational Gestures from Video

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.737193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.737193Z digest=sha256:5fb5f64e4b07d067edbfcdf3b06a9b3425d0b8cb670ab5d17f09c111b8304c42

Observation ac5f981c-7fa8-4f31-b20c-9ef6b0c7b640 · outbound

This paper cites Denoising diffu- sion probabilistic models, 2020.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Denoising diffu- sion probabilistic models, 2020

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.509139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.740672Z digest=sha256:456b9a1d8dfc34b02210ed08872e9bec89aa043db1c4e86eab804795ad00da17

Observation a8af7099-0f85-4b23-a505-55a1750477b3 · outbound

This paper cites Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.744241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.744241Z digest=sha256:d1afc9c561ec83a1d63679554f3a398ca1859368aa976f01ce3d9ffcf3359815

Observation 68f4e74c-0cd4-42ae-904a-31b33401b8ac · outbound

This paper cites Modeling and driving human body soundfields through acoustic primitives, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Modeling and driving human body soundfields through acoustic primitives, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.500361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.747430Z digest=sha256:027bb28d649c4fa9ea47a396de4edc3f922f45d2758b2f52fc659c6892dc964c

Observation 379df774-178b-4bd2-90ad-7b1cb603be65 · outbound

This paper cites Autoregressive Image Generation Using Residual Quantization, 2022.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Autoregressive Image Generation Using Residual Quantization, 2022

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.487530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.750470Z digest=sha256:32ae579ff35a8b50a4daf93aa55b243ebaa4f421970c79bbbd9f0ae1e3e94df5

Observation 56ae1ec7-820e-467d-8131-3a731be96359 · outbound

This paper cites Improving the training of rectified flows, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Improving the training of rectified flows, 2024

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.476806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.753381Z digest=sha256:78df6c78c40c0a7cccffe95fac70735a0c114d035eea0f400e4ab46266a020a6

Observation 568cde0b-d341-4d88-9b57-c3a7b85ee2fc · outbound

This paper cites Audio2gestures: Generating diverse gestures from speech audio with conditional varia- tional autoencoders.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Audio2gestures: Generating diverse gestures from speech audio with conditional varia- tional autoencoders

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.467114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.756284Z digest=sha256:76e568ccef0df2a22aac8c32b40e35462caa64c9f7722f7f8c5f71150156364f

Observation 0d24deb0-2043-419d-8bbb-7fbe14c42eb8 · outbound

This paper cites Set you straight: Auto-steering denoising trajectories to sidestep unwanted concepts, 2025.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Set you straight: Auto-steering denoising trajectories to sidestep unwanted concepts, 2025

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.457255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.759464Z digest=sha256:fa98e2e621849220b556f4abeba27219dedb02da7d2a20ac668822a28486c191

Observation 58f91817-ed80-4501-b264-21215d60d3cc · outbound

This paper cites AI Choreographer: Music Conditioned 3D Dance Generation with AIST++.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling AI Choreographer: Music Conditioned 3D Dance Generation with AIST++

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.447926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.762454Z digest=sha256:7f828ca67317939b4b042ad38b6ca03cf7cc7327c95959c0938632c2648f2fb2

Observation 1728fd61-8f87-4d1b-929e-8b309d1735e0 · outbound

This paper cites Flow Matching for Generative Modeling.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Flow Matching for Generative Modeling

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.765513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.765513Z digest=sha256:6cd938f4a1de1643f10cbd9f90e463243b694e1545c5bc050d4df2f849cd97b3

Observation 5116fdc1-5a4c-4075-9583-401ecb747f3a · outbound

This paper cites DisCo: Disentan- gled Implicit Content and Rhythm Learning for Diverse Co- Speech Gestures Synthesis.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling DisCo: Disentan- gled Implicit Content and Rhythm Learning for Diverse Co- Speech Gestures Synthesis

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.438437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.769338Z digest=sha256:5d20a4b4446a4a5abc82290aa1eed585e122bb6be9440f65f503ce1ce3ec6dae

Observation 422ac153-a138-4a25-9f3a-4540f888768c · outbound

This paper cites BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.772799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.772799Z digest=sha256:0da6495f9212246b8b7c548be162835368c81a3f70035e7c41539f83bd0dfe23

Observation d896c8ca-c937-4397-864d-99078a18186e · outbound

This paper cites EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.776187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.776187Z digest=sha256:61c892bfee130c461165b90530dcd6a5c816d88d388bf9989e4a807659e2467c

Observation 69adc1e6-d90a-4afc-8edb-c2a8677e7f93 · outbound

This paper cites TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.779538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.779538Z digest=sha256:92ca059987455e0baa0d78fefa58b3612d4e258abe39c41004b1128b42195f5b

Observation ba3e63c8-5161-4386-83a7-5733a2b0a3ff · outbound

This paper cites Semges: Semantics-aware co-speech gesture genera- tion using semantic coherence and relevance learning, 2025.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Semges: Semantics-aware co-speech gesture genera- tion using semantic coherence and relevance learning, 2025

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.783050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.783050Z digest=sha256:d1bf0f768c26ccde1b1bf5953865ad14977275ac646bc4fdc113046c0bc084f3

Observation 54e4a74d-be3d-43ee-b232-bdd70226a08a · outbound

This paper cites Intentional gesture: Deliver your intentions with gestures for speech, 2025.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Intentional gesture: Deliver your intentions with gestures for speech, 2025

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.418618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.786015Z digest=sha256:7726f563c6be28123ba98f824289e011fe26ce64472fa0985e920cc4b170f57a

Observation fa9593b5-2b2a-4cda-9240-ca5ce3d395c5 · outbound

This paper cites Contextual gesture: Co- speech gesture video generation through context-aware ges- ture representation, 2025.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Contextual gesture: Co- speech gesture video generation through context-aware ges- ture representation, 2025

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.407651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.789045Z digest=sha256:77bdb34e7e11b321db6826d0d119a6f8cde68cb160adcb6cac21de0f2d092dc1

Observation 182f8e85-7274-4125-9388-f8fb9c00d6b5 · outbound

This paper cites Flow straight and fast: Learning to generate and transfer data with rectified flow, 2022.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Flow straight and fast: Learning to generate and transfer data with rectified flow, 2022

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.397456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.791964Z digest=sha256:a49a1212f7f6e698e2d297cc50ebb7f1afc676a67ca1512da1280ac5e0e45584

Observation 1eb41990-733d-432d-b42f-3ed6cb5766a9 · outbound

This paper cites Learning Hierarchical Cross-Modal Association for Co-Speech Gesture Generation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Learning Hierarchical Cross-Modal Association for Co-Speech Gesture Generation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.388270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.794981Z digest=sha256:77ff00381306c597b9604ed6afd7a5f2beeb727e315c73588609310076502447

Observation ca132df2-8a79-4611-98b8-11cd3bc42eec · outbound

This paper cites Instaflow: One step is enough for high-quality diffusion- based text-to-image generation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Instaflow: One step is enough for high-quality diffusion- based text-to-image generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.797819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.797819Z digest=sha256:94de7b211f6bcd6634fcc781b98927d61e75b3c7438d132a1cd6b23d8f092f26

Observation c26b7603-ec1b-4614-9746-d89b6c0a0df7 · outbound

This paper cites Towards Variable and Coordinated Holistic Co-Speech Motion Generation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Towards Variable and Coordinated Holistic Co-Speech Motion Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.800710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.800710Z digest=sha256:11e2511cc313bb20980d262a1a187e5454f007103471ffbef146cb842fd05fa1

Observation 0c84542d-725c-42eb-a9fb-d58fba833d90 · outbound

This paper cites Tf-icon: Diffusion-based training-free cross-domain image composi- tion.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Tf-icon: Diffusion-based training-free cross-domain image composi- tion

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.372749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.803883Z digest=sha256:7ec3980abefd93674acdf1c42fe1faa9148f46db4b66fe5cd79f6c142634fb33

Observation 2944c756-ae37-4293-ac49-ad9df169b087 · outbound

This paper cites Mace: Mass concept erasure in diffusion models.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Mace: Mass concept erasure in diffusion models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.363735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.806870Z digest=sha256:dff557789a9792c4a4d1a0eee32059eae7ed6130f0c93ce3819aa94a9768b92f

Observation 07fcb2ac-0a53-417f-a035-6933fb22f84c · outbound

This paper cites Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.809929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.809929Z digest=sha256:75f61338bf1367025a84cab309331e9fcf2109bf3706a783af4185a357cae64b

Observation d8d5f3d5-3454-423e-a64f-1bcd3178db41 · outbound

This paper cites Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.813050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.813050Z digest=sha256:3ccabdf1be70561e56b9257fe3619b379caeee26ec89e4f9c548510b1f2675b2

Observation d39aa1b6-656e-47b5-a7ec-5e9b119384a2 · outbound

This paper cites DCTdiff: Intriguing Properties of Image Generative Modeling in the DCT Space.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling DCTdiff: Intriguing Properties of Image Generative Modeling in the DCT Space

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.816203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.816203Z digest=sha256:49c8871f07ddadf4a4cf094c59ce41dac9601af91ae39bd64a81236662785d44

Observation d3e1d58f-dcdc-4c86-99a4-6d57577bb4cf · outbound

This paper cites an unresolved cited work.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-09T22:06:55.353523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.819510Z digest=sha256:f869bcdf0c17ffe47d81a860cc2862870883f6f9c8f800d37ddd6c67f549f576

Observation 2209d4d5-50e2-4571-a7bf-79ab993047ac · outbound

This paper cites MM-Diffusion: Learning Multi-Modal Diffusion Models for Joint Audio and Video Generation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling MM-Diffusion: Learning Multi-Modal Diffusion Models for Joint Audio and Video Generation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.343642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.822378Z digest=sha256:4fcea93145059462af135a3e5fbaaee35c8c45595d488726690d122d557e6012

Observation b0f3a893-4573-4d70-bea3-53d438d6984b · outbound

This paper cites Fast high- resolution image synthesis with latent adversarial diffusion distillation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Fast high- resolution image synthesis with latent adversarial diffusion distillation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.333948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.825093Z digest=sha256:f992ce887d1f5f2ad3776f089b68f5b62334830d5973af672d5eaadd5bf1ee08

Observation d99ad8c5-b5fd-4c17-9887-06c79f954712 · outbound

This paper cites Talking face video generation with editable expression.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Talking face video generation with editable expression

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.323884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.827824Z digest=sha256:94e9a15ef804eb82e494e333d08166b2a64328037ed095e04be3683084b1097b

Observation e28b0df6-1c6b-45c8-9190-83ca8a846444 · outbound

This paper cites Fsft-net: face transfer video generation with few-shot views.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Fsft-net: face transfer video generation with few-shot views

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.314573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.830773Z digest=sha256:9d5e34bf0312aabc16ab7477c3ad2466f7789c1fd41ef32e96dc309c9148011b

Observation e3d5744c-14a3-4117-8015-8f5437e7659c · outbound

This paper cites Emotional listener portrait: Neural lis- tener head generation with emotion.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Emotional listener portrait: Neural lis- tener head generation with emotion

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.304564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.833651Z digest=sha256:6a378d40760963277a879dfda166059d8d8627e967ec0d4bddffc9e646490ad6

Observation 179a2de4-7b0f-4bf7-b714-6e502b61acfc · outbound

This paper cites Texttoon: Real-time text toonify head avatar from single video.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Texttoon: Real-time text toonify head avatar from single video

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.295521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.836452Z digest=sha256:5c17f6c9a209dcf6ea63fe14b1e84c84308b044ed74c6c053ae2ddfeb32ff1c7

Observation 4ed8c51d-25b6-4447-a1d9-0742cd9d9add · outbound

This paper cites Tri 2-plane: Thinking head avatar via fea- ture pyramid.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Tri 2-plane: Thinking head avatar via fea- ture pyramid

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.285873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.839250Z digest=sha256:7373cb43a6fc6d971f80e108f91b8fb76d239cdab463b77eb074bc58fb301138

Observation 47e1e9da-1858-494a-a880-f463b87a7e61 · outbound

This paper cites Adaptive super resolution for one-shot talking-head genera- tion.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Adaptive super resolution for one-shot talking-head genera- tion

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.274157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.842320Z digest=sha256:d00ece8fa60ba2c7deb9f0eaf50f16af78a44aa47ea0012b8863b99e79c3b8bc

Observation c68fbbdb-688b-439b-a4b9-d23e3876dab3 · outbound

This paper cites Consistency Models.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Consistency Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.845059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.845059Z digest=sha256:935425ffed577163b6e4d303735c3a0f6833e51d136ed6fd9b735cae2d89ae81

Observation 6103794e-2834-49a4-a325-fc6be3e0c102 · outbound

This paper cites Generative ai for cel- animation: A survey.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Generative ai for cel- animation: A survey

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.848096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.848096Z digest=sha256:31fa86a7163206bbc75e3e8b4f71d9e3240d3c323adb3af86b9a9355b2f276ee

Observation e82fde43-8164-498b-81e6-04fb3d647794 · outbound

This paper cites Gomez, Lukasz Kaiser, and Illia Polosukhin.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Gomez, Lukasz Kaiser, and Illia Polosukhin

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.263654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.851533Z digest=sha256:398184c9746db231729a8240c73340be7351f30758b95d4f7bcd55e391901ba4

Observation f583dae9-6b0b-4394-8682-ff72135f8e3b · outbound

This paper cites Rectified diffusion: Straightness is not your need in rectified flow, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Rectified diffusion: Straightness is not your need in rectified flow, 2024

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.854414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.854414Z digest=sha256:dc3526f396e4f270756a26420a22df0cdf924c1f08b332c405693d397f6ed832

Observation cbf548e5-b55c-47ec-a3de-dddc717326dc · outbound

This paper cites High-Resolution Im- age Synthesis and Semantic Manipulation with Conditional GANs.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling High-Resolution Im- age Synthesis and Semantic Manipulation with Conditional GANs

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.245732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.857488Z digest=sha256:7b742d314491f41037fd010c2886ea397716eb6bd1adb6906a37616536bcf59a

Observation 786cbaae-f059-4f48-9c95-fa881fd781c9 · outbound

This paper cites Codetalker: Speech-driven 3d facial animation with discrete motion prior.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Codetalker: Speech-driven 3d facial animation with discrete motion prior

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.235619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.860587Z digest=sha256:9eceeba1b6d5fc84820b95baa42e25909809d4be6dc1cf13239f4916cb590341

Observation 845aa018-0f1b-4530-b1bf-d44075617c81 · outbound

This paper cites Chain of generation: Multi-modal gesture synthesis via cascaded conditional control, 2023.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Chain of generation: Multi-modal gesture synthesis via cascaded conditional control, 2023

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.863520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.863520Z digest=sha256:0e06adcd55340633aa5e5dca655629561689952c34d9a05fcd9d2b4a1acc70ca

Observation 231c510e-48ec-4ce0-b892-cae56c00f4f9 · outbound

This paper cites Mambatalk: Ef- ficient holistic gesture synthesis with selective state space models, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Mambatalk: Ef- ficient holistic gesture synthesis with selective state space models, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.219864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.866681Z digest=sha256:c0f52734b3c5351ee2ec63b1dd82b8782096bff36ce4cbc56938b792fa78692f

Observation 0496e81e-0b8b-43f1-9a89-0306db409c92 · outbound

This paper cites Generating Holistic 3D Human Motion from Speech.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Generating Holistic 3D Human Motion from Speech

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.210578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.869588Z digest=sha256:c95d2fb29854633be34010a4ab7866e3282eab2b2dade412c72e856ed6ffb63f

Observation 89cb8a26-edeb-4562-97fa-7e438cf976e6 · outbound

This paper cites One-step diffusion with distribution matching distillation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling One-step diffusion with distribution matching distillation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.872671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.872671Z digest=sha256:3c8cbfb7ec10322c7826fa245c3bc5cd4ac197f1df4987da1b74a1854c6627f9

Observation 20ac6884-6604-479d-a8de-8b4640fc453a · outbound

This paper cites Speech Ges- ture Generation from the Trimodal Context of Text, Audio, and Speaker Identity.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Speech Ges- ture Generation from the Trimodal Context of Text, Audio, and Speaker Identity

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.193366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.875804Z digest=sha256:bdeecfba2ebeb604d43d37d13c8052e95001a7a95b92020e95094dd9554fc16c

Observation d8da7915-5de4-4030-b594-491fa069e379 · outbound

This paper cites Kinmo: Kinematic-aware human motion understanding and generation, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Kinmo: Kinematic-aware human motion understanding and generation, 2024

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.878794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.878794Z digest=sha256:23b802ac108e9d4ea1284c4e8df9d648b8075aa1479a077af5f0743a3be48c4d

Observation e51e51d7-7dc0-496b-8b3a-16a497c7d6ff · outbound

This paper cites Semantic gestic- ulator: Semantics-aware co-speech gesture synthesis, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Semantic gestic- ulator: Semantics-aware co-speech gesture synthesis, 2024

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.175575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.881930Z digest=sha256:e39a0b98be3f441deade1330e07533b447693c3b603eca8520012f3e53c0bc3e

Observation 8f2aef97-122d-4bef-a7b1-91da7177a8f1 · outbound

This paper cites Slimflow: Training smaller one-step diffusion models with rectified flow, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Slimflow: Training smaller one-step diffusion models with rectified flow, 2024

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.165417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.884864Z digest=sha256:337dccc3232d5a259463de10427d51c7a30c42725d19ad47ed5cf1d91e653f49

Observation 965b128b-ec83-4191-a037-db4d3d5a7a69 · outbound

This paper cites Oftsr: One-step flow for image super- resolution with tunable fidelity-realism trade-offs, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Oftsr: One-step flow for image super- resolution with tunable fidelity-realism trade-offs, 2024

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.155249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.887927Z digest=sha256:13aa3921fdadba464d4de8e2cbd4e3052c898c2ac66c4af44a55d9036abf72f0

Pith citing papers

Observation fd85416b-08b0-41ab-b5ea-7b50414bd5c9 · inbound

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation cites this paper.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.533226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.533226Z digest=sha256:c1cc6a1671b6e678eb2cb00a3cad1b154cba899bc23c034734b9da664cc7139c

Observation ddea5203-9589-4bda-8c20-036c43049e49 · inbound

Reality Check: How Avatar and Face Representation Affect the Perceptual Evaluation of Synthesized Gestures cites this paper.

Reality Check: How Avatar and Face Representation Affect the Perceptual Evaluation of Synthesized Gestures GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:11.047976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T03:39:22.419881Z digest=sha256:2a4de495222e5bae9ab0c8dcbb775db5a9788680b18a5bf62583dfd46e998ed7

Observation 7f4801b6-c754-4a63-8bd7-c4e8186e4b7d · inbound

Reality Check: How Avatar and Face Representation Affect the Perceptual Evaluation of Synthesized Gestures cites this paper.

Reality Check: How Avatar and Face Representation Affect the Perceptual Evaluation of Synthesized Gestures GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:21:16.775911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T02:19:11.967579Z digest=sha256:cb391cd412b8f6c773110180db25ac503a460810257ab62253e6fe937f80ab23

Observation 7afbf16e-15e8-4d16-a9d1-74406252815a · inbound

DyaPlex: Full-Duplex Speech-Motion Model for Dyadic Interaction cites this paper.

DyaPlex: Full-Duplex Speech-Motion Model for Dyadic Interaction GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:06:29.694605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T10:22:30.431160Z digest=sha256:e5dd7f5939adce5094e993c46d927bfbe548f06712ad6b179571df102d754169