Pith. sign in

Paper Citation Record · LEDGER

TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2410.04221.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.04221 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T22:06:54.779538Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T08:11:03.375642Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 69adc1e6-d90a-4afc-8edb-c2a8677e7f93 · inbound

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling cites this paper.

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T22:06:54.779538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:06:54.779538Z digest=sha256:2aecf4a426e4bd4ff5ac75bb7f642253a4ec03afa72b92e49003cb5bc33c6f0b

Observation eb1ee01c-a96d-4f98-97e6-e1d5a1088a54 · inbound

HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation cites this paper.

HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T21:16:16.523335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:16:16.523335Z digest=sha256:d9243bbbd146e46f12220e805850a41cddf06f215a33d94820840f7f9b0f8503

Observation 5c6eb645-3586-49ee-8797-43e93e7a6763 · inbound

MotionRAG-Diff: A Retrieval-Augmented Diffusion Framework for Long-Term Music-to-Dance Generation cites this paper.

MotionRAG-Diff: A Retrieval-Augmented Diffusion Framework for Long-Term Music-to-Dance Generation TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:22:38.599410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:22:38.599410Z digest=sha256:93daf624f980c91620e2cbced613fc92b09b0ff8a637b6279496caaf740e7e3f

Observation 9370856a-bb87-47d2-87dd-38fe730e3878 · inbound

Multi-human Interactive Talking Dataset cites this paper.

Multi-human Interactive Talking Dataset TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T04:46:04.127815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:46:04.127815Z digest=sha256:dd58e3f509f19d2a6b317a795dfad5359836fa16493ee823ce8929e82d56532e

Observation 5edae754-ffb5-4aaf-93f2-d7af468a7428 · inbound

LiveGesture Streamable Co-Speech Gesture Generation Model cites this paper.

LiveGesture Streamable Co-Speech Gesture Generation Model TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:11:03.380850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T16:46:58.944000Z digest=sha256:94e34cf36adabb2f08898629589b064939d117af8664551154880192836c07df