Pith. sign in

Paper Citation Record · LEDGER

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

As of 10 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 4 inbound Pith citation observations for arXiv:2501.18898.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.18898 v3

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T22:06:54.887927Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T16:24:07.533226Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T03:06:29.692095Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact0
  • verified fuzzy43
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 637d77c7-32b8-4076-8cc1-0a86675c8eb2 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets, 2023.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets, 2023

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.636490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.686959Z digest=sha256:1280f53ccc43e8851c79a8b96c3e35580244988276e2e7859a6e6181a89dfe4b

Observation 5a082a49-65e1-458c-9e3e-62e226cec4a3 · outbound

This paper cites Nonver- bal Behaviors, Persuasion, and Credibility.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Nonver- bal Behaviors, Persuasion, and Credibility

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.626822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.690773Z digest=sha256:e7c4cb5aa681a8ee66ad4d9c7ab4296586bc56534a5a42d4019ecffed82c4944

Observation 6ed24e08-0c2d-4564-bd19-19e250ce5e0d · outbound

This paper cites Everybody Dance Now.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Everybody Dance Now

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.616756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.694173Z digest=sha256:14d034bfb5d27e5bf6d7f37df3d4d0b6c10d2b7e9fb935275b36eb9ae5f08293

Observation 3bf10411-b2dd-4e33-b467-ef1915b5701c · outbound

This paper cites Enabling synergistic full-body control in prompt-based co-speech motion generation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Enabling synergistic full-body control in prompt-based co-speech motion generation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.607020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.697466Z digest=sha256:7686ea8132916a2baefe4b6943fb4ff17b8c9129b2ec8d250f120326ecad4c6e

Observation 841495ec-0992-4e87-8bbc-4a0821c32dd3 · outbound

This paper cites DiffSHEG: A Diffusion-Based Ap- proach for Real-Time Speech-driven Holistic 3D Expression and Gesture Generation, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling DiffSHEG: A Diffusion-Based Ap- proach for Real-Time Speech-driven Holistic 3D Expression and Gesture Generation, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.700476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.700476Z digest=sha256:30c5146ca8712af5bcec401ead8fecfc596a1f38b049a9018eae35dd60d44049

Observation 30eaaba3-2f89-4121-bfb3-748b25a7f51d · outbound

This paper cites WavLM: Large-Scale Self- Supervised Pre-Training for Full Stack Speech Processing.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling WavLM: Large-Scale Self- Supervised Pre-Training for Full Stack Speech Processing

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.590944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.703474Z digest=sha256:8a2c8092c7c6a835bfa15feffb6262c8868b850a1237fadf7be4b8e362fd83d2

Observation d83ab87a-ddfb-4016-9034-716a51e5215f · outbound

This paper cites MotionLCM: Real-time Controllable Motion Generation via Latent Consistency Model.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling MotionLCM: Real-time Controllable Motion Generation via Latent Consistency Model

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.706587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.706587Z digest=sha256:34f0fce536b3ef0a9877169b4942204a45de31ebcb37013538c38adfdc00250a

Observation 54765c55-b93f-493a-a6d8-df82b0501609 · outbound

This paper cites The Interplay Between Gesture and Speech in the Production of Referring Expressions: Investigating the Tradeoff Hypothe- sis.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling The Interplay Between Gesture and Speech in the Production of Referring Expressions: Investigating the Tradeoff Hypothe- sis

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.580515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.710168Z digest=sha256:c6a7cb892caaa0d00e82aefb75cfae6aa1ed91407cecaddfa290b4dceee38c62

Observation aaf16f7d-bf14-4d5d-bfe0-7a3bb08be096 · outbound

This paper cites Diffusion-based co-speech gesture genera- tion using joint text and audio representation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Diffusion-based co-speech gesture genera- tion using joint text and audio representation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.570947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.713193Z digest=sha256:a9d136f63594f54316d10b959094d5729b8023eb07c8ea64a4a3020b27eaf3eb

Observation 44f4a4f2-2708-457a-84be-272c892807b3 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.716207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.716207Z digest=sha256:df12b8ca7ab5231530c56bb069f86084f76d1524e5dcd2715867a36a1c08daef

Observation f726766b-00e7-4829-a445-30510941293c · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale, 2021.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling An image is worth 16x16 words: Transformers for image recognition at scale, 2021

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.719776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.719776Z digest=sha256:fa20f44eb8ee1bc51faa3ea5e30e5aa673054da259d8e88c7649073690aaf648

Observation 9837a269-ab17-49c2-8f2e-7f3c5c4f49b4 · outbound

This paper cites Scaling rectified flow trans- formers for high-resolution image synthesis, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Scaling rectified flow trans- formers for high-resolution image synthesis, 2024

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.555274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.722759Z digest=sha256:76bb256c895b86f76699df39ad7712e785ed3b34b34e9025f5c1f44ce52a820e

Observation c77b73ec-b354-4ef2-8416-a64ea7e21b20 · outbound

This paper cites One step diffusion via shortcut models, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling One step diffusion via shortcut models, 2024

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.545844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.725469Z digest=sha256:1a6197369b3c6e1e69f4ae401134684b2a3b194044a201d9af8d12849720d62c

Observation fdb0bb53-f0c6-4e5f-83bf-f1c4e5f6df8a · outbound

This paper cites Eraseanything: Enabling concept erasure in rectified flow transformers.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Eraseanything: Enabling concept erasure in rectified flow transformers

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.536960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.728178Z digest=sha256:4731db61c34c9323984de634c93d2c000472aad93cf2967579e0f799a3f4ec24

Observation 9ed8d118-8a65-45e0-834f-004c4b279f97 · outbound

This paper cites Ginosar, A.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Ginosar, A

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.527499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.731020Z digest=sha256:8694af7a37225f5dc1dd909612254ef32223d8ea82d82be0b592237a941738b9

Observation 81ad22ba-4613-4a10-a304-c74e5b2a7576 · outbound

This paper cites Momask: Generative masked model- ing of 3d human motions.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Momask: Generative masked model- ing of 3d human motions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.518538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.734221Z digest=sha256:aa141562e4593817747e8f7c9faf7767a36a48548ebd1f2050d02bc51ebeb487

Observation d7427ef7-a4e0-45fe-8ef1-539d07cb259c · outbound

This paper cites Learning Speech-driven 3D Conversational Gestures from Video.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Learning Speech-driven 3D Conversational Gestures from Video

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.737193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.737193Z digest=sha256:55eae3c3e4998aaf10c77cb8bccf82968ea017e382275c3f5d6752c9f9ef6623

Observation ac5f981c-7fa8-4f31-b20c-9ef6b0c7b640 · outbound

This paper cites Denoising diffu- sion probabilistic models, 2020.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Denoising diffu- sion probabilistic models, 2020

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.509139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.740672Z digest=sha256:bcd393d1a92aa93faee0cef1c612f0169512c90a1879a9e38c4b903ef28e9032

Observation a8af7099-0f85-4b23-a505-55a1750477b3 · outbound

This paper cites Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.744241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.744241Z digest=sha256:6157230c92710386a31fd1950eb0852f0989bdb448506f528591bb2f599b78cd

Observation 68f4e74c-0cd4-42ae-904a-31b33401b8ac · outbound

This paper cites Modeling and driving human body soundfields through acoustic primitives, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Modeling and driving human body soundfields through acoustic primitives, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.500361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.747430Z digest=sha256:94a6370999e8c01df7bd217ec04d63e9a20797a9b4878b6add79f06d0fb5bead

Observation 379df774-178b-4bd2-90ad-7b1cb603be65 · outbound

This paper cites Autoregressive Image Generation Using Residual Quantization, 2022.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Autoregressive Image Generation Using Residual Quantization, 2022

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.487530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.750470Z digest=sha256:535661286ed41769cdfcc397441c0577850993e90380a26942629a94a2dfe99e

Observation 56ae1ec7-820e-467d-8131-3a731be96359 · outbound

This paper cites Improving the training of rectified flows, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Improving the training of rectified flows, 2024

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.476806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.753381Z digest=sha256:21bb09c016fb4a350aba87ec8bf3e69c894d55cd8f6c11717bd487fda019170a

Observation 568cde0b-d341-4d88-9b57-c3a7b85ee2fc · outbound

This paper cites Audio2gestures: Generating diverse gestures from speech audio with conditional varia- tional autoencoders.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Audio2gestures: Generating diverse gestures from speech audio with conditional varia- tional autoencoders

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.467114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.756284Z digest=sha256:d37627f4511df841e70531a7bd4daba756bd6db84771ef5835ea16e75fa375ef

Observation 0d24deb0-2043-419d-8bbb-7fbe14c42eb8 · outbound

This paper cites Set you straight: Auto-steering denoising trajectories to sidestep unwanted concepts, 2025.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Set you straight: Auto-steering denoising trajectories to sidestep unwanted concepts, 2025

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.457255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.759464Z digest=sha256:dba5bc1d0d1e360e41aba05cb61f18ada6c8b76986b29109410f75917bc748c4

Observation 58f91817-ed80-4501-b264-21215d60d3cc · outbound

This paper cites AI Choreographer: Music Conditioned 3D Dance Generation with AIST++.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling AI Choreographer: Music Conditioned 3D Dance Generation with AIST++

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.447926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.762454Z digest=sha256:e37e9bfe4761e800ec26deed91c7740b170479802490d240d4410c13b8d0ace4

Observation 1728fd61-8f87-4d1b-929e-8b309d1735e0 · outbound

This paper cites Flow Matching for Generative Modeling.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Flow Matching for Generative Modeling

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.765513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.765513Z digest=sha256:5a7f4b3c1f4b3763f8beea5a6c55318b8d3a640d698988b918b0c275f7baa7d4

Observation 5116fdc1-5a4c-4075-9583-401ecb747f3a · outbound

This paper cites DisCo: Disentan- gled Implicit Content and Rhythm Learning for Diverse Co- Speech Gestures Synthesis.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling DisCo: Disentan- gled Implicit Content and Rhythm Learning for Diverse Co- Speech Gestures Synthesis

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.438437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.769338Z digest=sha256:9667d0dbadc4e5c50de6fccd517e3136ffad52b165e0138621c58ee62413ff22

Observation 422ac153-a138-4a25-9f3a-4540f888768c · outbound

This paper cites BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.772799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.772799Z digest=sha256:a825f9853b1a1738309b52aae8e00a1e9fc1c9ef88db0345552688499bac010a

Observation d896c8ca-c937-4397-864d-99078a18186e · outbound

This paper cites EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.776187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.776187Z digest=sha256:8167cc330aee5db0880a9d4061d58c0255b8a25511b16c38d750cb99c6b67cd0

Observation 69adc1e6-d90a-4afc-8edb-c2a8677e7f93 · outbound

This paper cites TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.779538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.779538Z digest=sha256:2aecf4a426e4bd4ff5ac75bb7f642253a4ec03afa72b92e49003cb5bc33c6f0b

Observation ba3e63c8-5161-4386-83a7-5733a2b0a3ff · outbound

This paper cites Semges: Semantics-aware co-speech gesture genera- tion using semantic coherence and relevance learning, 2025.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Semges: Semantics-aware co-speech gesture genera- tion using semantic coherence and relevance learning, 2025

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.783050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.783050Z digest=sha256:2c59b8f72de34a8722e28be8f85e764788c6ceb3ccc0b09e52ec2608aba5f5d3

Observation 54e4a74d-be3d-43ee-b232-bdd70226a08a · outbound

This paper cites Intentional gesture: Deliver your intentions with gestures for speech, 2025.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Intentional gesture: Deliver your intentions with gestures for speech, 2025

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.418618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.786015Z digest=sha256:21bbc5727940a1c2e536114e7486c7514da3e9ad4c51e96ad7fc2ea998cf05ff

Observation fa9593b5-2b2a-4cda-9240-ca5ce3d395c5 · outbound

This paper cites Contextual gesture: Co- speech gesture video generation through context-aware ges- ture representation, 2025.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Contextual gesture: Co- speech gesture video generation through context-aware ges- ture representation, 2025

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.407651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.789045Z digest=sha256:983608cc961e8565f3c7868e103025ffa6d7e61e44f6dacdc9430cafaf985a27

Observation 182f8e85-7274-4125-9388-f8fb9c00d6b5 · outbound

This paper cites Flow straight and fast: Learning to generate and transfer data with rectified flow, 2022.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Flow straight and fast: Learning to generate and transfer data with rectified flow, 2022

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.397456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.791964Z digest=sha256:52d38a9c916c7a69533ab75508124feb0aadd2fc28671ab07bcf0b1fae9a880a

Observation 1eb41990-733d-432d-b42f-3ed6cb5766a9 · outbound

This paper cites Learning Hierarchical Cross-Modal Association for Co-Speech Gesture Generation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Learning Hierarchical Cross-Modal Association for Co-Speech Gesture Generation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.388270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.794981Z digest=sha256:e7641a240738e7986480013c1c501dd48e718894f97bf9ea2d9a6604e5895867

Observation ca132df2-8a79-4611-98b8-11cd3bc42eec · outbound

This paper cites Instaflow: One step is enough for high-quality diffusion- based text-to-image generation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Instaflow: One step is enough for high-quality diffusion- based text-to-image generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.797819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.797819Z digest=sha256:ae87984bb0eb29d8e71fc1fbf69e3563fb746025e12405ec447e1adfefb45edf

Observation c26b7603-ec1b-4614-9746-d89b6c0a0df7 · outbound

This paper cites Towards Variable and Coordinated Holistic Co-Speech Motion Generation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Towards Variable and Coordinated Holistic Co-Speech Motion Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.800710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.800710Z digest=sha256:c75e842b2435d4e4aa407729a0c814bd41a2e28e7f184be1ec104d604fe7fa03

Observation 0c84542d-725c-42eb-a9fb-d58fba833d90 · outbound

This paper cites Tf-icon: Diffusion-based training-free cross-domain image composi- tion.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Tf-icon: Diffusion-based training-free cross-domain image composi- tion

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.372749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.803883Z digest=sha256:a820776e33b5e1f3223f235a5cac5a7002757d63ea1c22a240c9c2ab2f5a3733

Observation 2944c756-ae37-4293-ac49-ad9df169b087 · outbound

This paper cites Mace: Mass concept erasure in diffusion models.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Mace: Mass concept erasure in diffusion models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.363735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.806870Z digest=sha256:fda38c78f59d5baa554047e4997fb4930c99544f17980244a26e07fe41dc4752

Observation 07fcb2ac-0a53-417f-a035-6933fb22f84c · outbound

This paper cites Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.809929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.809929Z digest=sha256:e2b88be84010ca99b1545845789f28a2d2981694cbf12beda30151298bb9472b

Observation d8d5f3d5-3454-423e-a64f-1bcd3178db41 · outbound

This paper cites Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.813050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.813050Z digest=sha256:91613cd78e11c07967135dd759717d1f7032163c135007ff790810e442cfea61

Observation d39aa1b6-656e-47b5-a7ec-5e9b119384a2 · outbound

This paper cites DCTdiff: Intriguing Properties of Image Generative Modeling in the DCT Space.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling DCTdiff: Intriguing Properties of Image Generative Modeling in the DCT Space

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.816203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.816203Z digest=sha256:16406e5f579eea79db98c8e43832e99bf21fb7076727f56e6dd43454e5b8c1d2

Observation d3e1d58f-dcdc-4c86-99a4-6d57577bb4cf · outbound

This paper cites an unresolved cited work.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-09T22:06:55.353523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.819510Z digest=sha256:745ec56839e53c2769a674d0437947349610def3a07bd4d202f2eaf1106dbe42

Observation 2209d4d5-50e2-4571-a7bf-79ab993047ac · outbound

This paper cites MM-Diffusion: Learning Multi-Modal Diffusion Models for Joint Audio and Video Generation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling MM-Diffusion: Learning Multi-Modal Diffusion Models for Joint Audio and Video Generation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.343642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.822378Z digest=sha256:5257b95ba6cb7b6da5e3cba93e521aaed084be3d0f71eee0a2af125a62e0ec56

Observation b0f3a893-4573-4d70-bea3-53d438d6984b · outbound

This paper cites Fast high- resolution image synthesis with latent adversarial diffusion distillation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Fast high- resolution image synthesis with latent adversarial diffusion distillation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.333948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.825093Z digest=sha256:e86c0aacd34a9245d339a2487f53117e72902289e223a83dc184d46cb569c190

Observation d99ad8c5-b5fd-4c17-9887-06c79f954712 · outbound

This paper cites Talking face video generation with editable expression.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Talking face video generation with editable expression

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.323884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.827824Z digest=sha256:b0855f51524c2bb13bd8893ca416e4098e725bd62c4fb3f9a530af6e57e83147

Observation e28b0df6-1c6b-45c8-9190-83ca8a846444 · outbound

This paper cites Fsft-net: face transfer video generation with few-shot views.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Fsft-net: face transfer video generation with few-shot views

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.314573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.830773Z digest=sha256:a8f03504702a181edd8f0b3d962e2313686a686e3492ed4063090cb3a66e254e

Observation e3d5744c-14a3-4117-8015-8f5437e7659c · outbound

This paper cites Emotional listener portrait: Neural lis- tener head generation with emotion.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Emotional listener portrait: Neural lis- tener head generation with emotion

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.304564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.833651Z digest=sha256:223c97fc8420953f4c84419bdca2b7a200c08f518c4561f842a4d6e953fc4645

Observation 179a2de4-7b0f-4bf7-b714-6e502b61acfc · outbound

This paper cites Texttoon: Real-time text toonify head avatar from single video.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Texttoon: Real-time text toonify head avatar from single video

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.295521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.836452Z digest=sha256:478e4387c3c58637bb664378a60901df99a0c5f57e10d5ef2e88a3a8c711e64e

Observation 4ed8c51d-25b6-4447-a1d9-0742cd9d9add · outbound

This paper cites Tri 2-plane: Thinking head avatar via fea- ture pyramid.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Tri 2-plane: Thinking head avatar via fea- ture pyramid

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.285873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.839250Z digest=sha256:17c294834e16ef7b28fe8843cfc0d701b53c6bd9b7f95c77da6ac32b047a0c45

Observation 47e1e9da-1858-494a-a880-f463b87a7e61 · outbound

This paper cites Adaptive super resolution for one-shot talking-head genera- tion.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Adaptive super resolution for one-shot talking-head genera- tion

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.274157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.842320Z digest=sha256:3b3f1e17d9edc788bbd333b75a58be8d19688aa6152575337bae35c97b56d777

Observation c68fbbdb-688b-439b-a4b9-d23e3876dab3 · outbound

This paper cites Consistency Models.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Consistency Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.845059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.845059Z digest=sha256:e9a5944ad7a72ba6e32e7b5f5d4d8560b23eaf64a23c1de50a13b2e3873bd273

Observation 6103794e-2834-49a4-a325-fc6be3e0c102 · outbound

This paper cites Generative ai for cel- animation: A survey.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Generative ai for cel- animation: A survey

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.848096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.848096Z digest=sha256:0f14c078fc28246c441571e1f8be0eac55e78278390848b6a2832a4965846fad

Observation e82fde43-8164-498b-81e6-04fb3d647794 · outbound

This paper cites Gomez, Lukasz Kaiser, and Illia Polosukhin.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Gomez, Lukasz Kaiser, and Illia Polosukhin

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.263654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.851533Z digest=sha256:14c17484e3cfa4a36f83ee119a991a2c2601a37bc6456313ad77f95fe8e279ea

Observation f583dae9-6b0b-4394-8682-ff72135f8e3b · outbound

This paper cites Rectified diffusion: Straightness is not your need in rectified flow, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Rectified diffusion: Straightness is not your need in rectified flow, 2024

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.854414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.854414Z digest=sha256:2c30b06789aa0e9556b7657976ef23d430d89a2e9c175115d822cfba488e4ea6

Observation cbf548e5-b55c-47ec-a3de-dddc717326dc · outbound

This paper cites High-Resolution Im- age Synthesis and Semantic Manipulation with Conditional GANs.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling High-Resolution Im- age Synthesis and Semantic Manipulation with Conditional GANs

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.245732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.857488Z digest=sha256:ff5f94a6f08727c6376456d586a260e00fbadea1cccd2136b062974f378cdb3a

Observation 786cbaae-f059-4f48-9c95-fa881fd781c9 · outbound

This paper cites Codetalker: Speech-driven 3d facial animation with discrete motion prior.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Codetalker: Speech-driven 3d facial animation with discrete motion prior

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.235619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.860587Z digest=sha256:4dab5013743e3dffb19c9f42fba2e9a0f1ecb4a2160640f0c8cfb6965ebfdaa9

Observation 845aa018-0f1b-4530-b1bf-d44075617c81 · outbound

This paper cites Chain of generation: Multi-modal gesture synthesis via cascaded conditional control, 2023.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Chain of generation: Multi-modal gesture synthesis via cascaded conditional control, 2023

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.863520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.863520Z digest=sha256:932ef9790a4a1bbd4576337ad84f975262cf60b856b1b3d89133f1a1d0f922af

Observation 231c510e-48ec-4ce0-b892-cae56c00f4f9 · outbound

This paper cites Mambatalk: Ef- ficient holistic gesture synthesis with selective state space models, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Mambatalk: Ef- ficient holistic gesture synthesis with selective state space models, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.219864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.866681Z digest=sha256:4f783c5cdb15f134eadcaafbeb7f512fcea1224a2bfcc558cca1e98c2baa18cd

Observation 0496e81e-0b8b-43f1-9a89-0306db409c92 · outbound

This paper cites Generating Holistic 3D Human Motion from Speech.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Generating Holistic 3D Human Motion from Speech

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.210578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.869588Z digest=sha256:19ee9f3d7aa905d1f8b6f8bf61afa18a7b6e1a431e446fb19aa571accce08770

Observation 89cb8a26-edeb-4562-97fa-7e438cf976e6 · outbound

This paper cites One-step diffusion with distribution matching distillation.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling One-step diffusion with distribution matching distillation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.872671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.872671Z digest=sha256:794a774d3007713e61c89d42364a176fcf68ebdcc1a7b849b9e7717200a257a1

Observation 20ac6884-6604-479d-a8de-8b4640fc453a · outbound

This paper cites Speech Ges- ture Generation from the Trimodal Context of Text, Audio, and Speaker Identity.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Speech Ges- ture Generation from the Trimodal Context of Text, Audio, and Speaker Identity

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.193366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.875804Z digest=sha256:0e633e0d4e3ed8d5e246bba7348ad684f02d3ad43d9ebbd72748ce8f78f0ddd2

Observation d8da7915-5de4-4030-b594-491fa069e379 · outbound

This paper cites Kinmo: Kinematic-aware human motion understanding and generation, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Kinmo: Kinematic-aware human motion understanding and generation, 2024

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.878794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.878794Z digest=sha256:393d9b8028808eeecbe6b68786bd290a1760a329c86e333db1b4502b05423522

Observation e51e51d7-7dc0-496b-8b3a-16a497c7d6ff · outbound

This paper cites Semantic gestic- ulator: Semantics-aware co-speech gesture synthesis, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Semantic gestic- ulator: Semantics-aware co-speech gesture synthesis, 2024

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.175575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.881930Z digest=sha256:46e8b54a55311e78664cd0b25e08f0c77d7b812d33fe87c9d11cb39507332902

Observation 8f2aef97-122d-4bef-a7b1-91da7177a8f1 · outbound

This paper cites Slimflow: Training smaller one-step diffusion models with rectified flow, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Slimflow: Training smaller one-step diffusion models with rectified flow, 2024

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.165417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.884864Z digest=sha256:62bc9dda160d2fca689d1146d159e7ff836b48f7741c9915454a594cdc6b16f0

Observation 965b128b-ec83-4191-a037-db4d3d5a7a69 · outbound

This paper cites Oftsr: One-step flow for image super- resolution with tunable fidelity-realism trade-offs, 2024.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling Oftsr: One-step flow for image super- resolution with tunable fidelity-realism trade-offs, 2024

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:06:55.155249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:06:54.887927Z digest=sha256:8e200addd1c4c2aa4b44245c861e87425a3467867708aedb323b6756215885a2

Pith citing papers

Observation fd85416b-08b0-41ab-b5ea-7b50414bd5c9 · inbound

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation cites this paper.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.533226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.533226Z digest=sha256:96ed31e26083a8e7b795a62502c4c2675a9ee10fae3ae8a58665d89a196e5938

Observation ddea5203-9589-4bda-8c20-036c43049e49 · inbound

Reality Check: How Avatar and Face Representation Affect the Perceptual Evaluation of Synthesized Gestures cites this paper.

Reality Check: How Avatar and Face Representation Affect the Perceptual Evaluation of Synthesized Gestures GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:11.047976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T03:39:22.419881Z digest=sha256:280e674480ee600d7d708dcb1ebc22fcedcb41eabc9edb81b41fe9859c5dc508

Observation 7f4801b6-c754-4a63-8bd7-c4e8186e4b7d · inbound

Reality Check: How Avatar and Face Representation Affect the Perceptual Evaluation of Synthesized Gestures cites this paper.

Reality Check: How Avatar and Face Representation Affect the Perceptual Evaluation of Synthesized Gestures GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:21:16.775911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T02:19:11.967579Z digest=sha256:706c7a186a9dae246b2310bf32f79b3ccf24bf428c00426e235fad20552f83b2

Observation 7afbf16e-15e8-4d16-a9d1-74406252815a · inbound

DyaPlex: Full-Duplex Speech-Motion Model for Dyadic Interaction cites this paper.

DyaPlex: Full-Duplex Speech-Motion Model for Dyadic Interaction GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:06:29.694605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T10:22:30.431160Z digest=sha256:71224c42f218a5e0d261c8e39ee0ebacc3340b9fb5a615ec41356ded0d78f8f2