Pith. sign in

Paper Citation Record · LEDGER

TSTMotion: Training-free Scene-aware Text-to-motion Generation

As of 19 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 1 inbound Pith citation observation for arXiv:2505.01182.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.01182 v2

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:30:08.666034Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T12:31:22.605420Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-09T12:31:22.996104Z

Reference resolution

21 of 21 outbound references displayed

  • verified exact2
  • verified fuzzy12
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fd3315f5-7378-4237-9702-28c649bd874a · outbound

This paper cites Human Motion Diffusion Model.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Human Motion Diffusion Model

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.570066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.570066Z digest=sha256:cf27b49ee3e5df0986358c2630b67efd53ae80bfb9549a74e2d7b5613a45360b

Observation d0d205bb-2b98-4fad-9b65-1ae2d76af6f2 · outbound

This paper cites MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model.

TSTMotion: Training-free Scene-aware Text-to-motion Generation MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.575943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.575943Z digest=sha256:892a9572695c05471937d39b8b961b115a8f7ff16432d1a71cc7c7e64f8c75bd

Observation e2ea81fd-e9af-45f8-b9f9-c37363aa2b04 · outbound

This paper cites Humanise: Language-conditioned human motion generation in 3d scenes,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Humanise: Language-conditioned human motion generation in 3d scenes,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:09.050825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.581072Z digest=sha256:a2730ba7447e3adaa91afe2fceb27fe775bee3625c3dbb9ccf6cc9e9313aa855

Observation 8c4601ae-b98f-4d31-a7b3-8f030b49eda6 · outbound

This paper cites Unified human-scene interaction via prompted chain-of-contacts,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Unified human-scene interaction via prompted chain-of-contacts,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:09.029927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.585863Z digest=sha256:340ffaa7779b7a7eeaa05323ace6f53f19239e8a494ca8c06fe78037bf4af9b2

Observation 716d6ba9-50fc-4aa4-9f34-680d1962ae0b · outbound

This paper cites Scaling up dynamic human-scene interaction modeling,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Scaling up dynamic human-scene interaction modeling,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:09.013724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.590267Z digest=sha256:a7c7b453755115a7f47255e949c5861742a399d30e6f009f77d7189e69387f36

Observation a6ba2fe0-1bc6-47c8-9548-322343ba007e · outbound

This paper cites Move as You Say, Interact as You Can: Language-guided Human Motion Generation with Scene Affordance.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Move as You Say, Interact as You Can: Language-guided Human Motion Generation with Scene Affordance

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-16T04:30:08.815613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.594698Z digest=sha256:9a353fb2eb3e9775f7fe81fd87a744136da47d5dfe2e754c5526fa981d60d780

Observation bfc4b223-1e06-47bd-b239-88653af3e64f · outbound

This paper cites Stochastic scene-aware motion prediction,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Stochastic scene-aware motion prediction,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.996801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.599749Z digest=sha256:9d03686230e5f997d8309a42a08f5ebacd2521a975b5c0ba1c6140c9f8c1cc73

Observation 1f00d3aa-c95e-4fcf-9c2e-b1767874bcec · outbound

This paper cites Synthesizing Diverse Human Motions in 3D Indoor Scenes.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Synthesizing Diverse Human Motions in 3D Indoor Scenes

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-16T04:30:08.791477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.603907Z digest=sha256:15633f3a7a8e05ed5b30a4f3dd7a73d77a3c920498361b9a73f4a8d9dc651e44

Observation 3c318909-9016-4e78-afa6-f8a1662c8f94 · outbound

This paper cites Generating diverse and natural 3d human motions from text,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Generating diverse and natural 3d human motions from text,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.981152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.608284Z digest=sha256:a68d6dce53cf23e98f488e6bb97cec52a7ccd318dd09d454a0bb00546ef10417

Observation 34bee9e1-2f46-464a-a823-4e68cbf7a68e · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Chain-of-thought prompting elicits reasoning in large language models,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.965539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.612499Z digest=sha256:ac8db41004f9eca9c1763548c62c5efd0d2bd313e7af52ef7b18c7874929cbd1

Observation ed01635c-9c89-4851-9623-d37211f2dabd · outbound

This paper cites Visual language maps for robot navigation,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Visual language maps for robot navigation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.949133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.616570Z digest=sha256:d712cb932febc9dcc01d08e42bd5a4e53d05b78519c0b8d09edcd5b3eb96ad9b

Observation 518fd29e-d7ad-4995-890a-159f650a59bf · outbound

This paper cites Tabllm: Few-shot classification of tabular data with large language models,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Tabllm: Few-shot classification of tabular data with large language models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.930160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.621528Z digest=sha256:8237a32b51cd847e702db3ac3f5f49ebed1adaf4741d966e039825d68dd8397d

Observation 54c274b1-f1b8-4b45-8ca3-0c7afa4a38ba · outbound

This paper cites Language models are few-shot learners,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Language models are few-shot learners,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.913206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.626397Z digest=sha256:0b8aa8c951b893cbab7e98b5d6173ba70d2cd7e2fc526d794810790cd44ac3db

Observation 13398b34-1b1b-4fac-b1c3-1585fe712cc9 · outbound

This paper cites Diffusion Posterior Sampling for General Noisy Inverse Problems.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Diffusion Posterior Sampling for General Noisy Inverse Problems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.631348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.631348Z digest=sha256:57b0cff4d4bafefa6ab18fab412ddd621d293526b0545b0a7a57b5af8261ca37

Observation a0b71c18-cd0e-425b-96f2-c47bad26a9f0 · outbound

This paper cites Ex- pressive body capture: 3D hands, face, and body from a single image,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Ex- pressive body capture: 3D hands, face, and body from a single image,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.896164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.636071Z digest=sha256:f604786c3bd1851466dcca1aee428daf95bc6b8264ea215b6a0921f843980a24

Observation 5fff4d97-50ce-4ba3-95bf-fee379698b1d · outbound

This paper cites Enhancing Human-Centered Dynamic Scene Understanding via Multiple LLMs Collaborated Reasoning.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Enhancing Human-Centered Dynamic Scene Understanding via Multiple LLMs Collaborated Reasoning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.641750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.641750Z digest=sha256:6d07f1be834ef938144efef09745003d40ed7776b8073482b9cf97da825e14ac

Observation 812ac5f0-b8c7-4317-81ad-978ac90dc382 · outbound

This paper cites Recognize anything: A strong image tagging model,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Recognize anything: A strong image tagging model,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.880240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.646732Z digest=sha256:0d9a83bd83d692f3f5040f65bb0c993966e0a41d0ce683e3198ea1067e32ec78

Observation 62d47bc0-47f7-4715-8099-4fe9dddd01e4 · outbound

This paper cites OpenIns3D: Snap and Lookup for 3D Open-vocabulary Instance Segmentation.

TSTMotion: Training-free Scene-aware Text-to-motion Generation OpenIns3D: Snap and Lookup for 3D Open-vocabulary Instance Segmentation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.651493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.651493Z digest=sha256:baa441a5069179386f0f4d9d0b9e102304137a3583893d51ac8dbcf2633c6434

Observation a5d525c9-f8b4-408e-9cd7-326d0be5b0b8 · outbound

This paper cites GPT-4 Technical Report.

TSTMotion: Training-free Scene-aware Text-to-motion Generation GPT-4 Technical Report

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.656276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.656276Z digest=sha256:5abbdb981ac5295b4480ed5f425f311f8b138e570b5d497487341e3879f5080b

Observation 4c17d583-9859-4e26-b80b-34b2c3a4d773 · outbound

This paper cites OmniControl: Control Any Joint at Any Time for Human Motion Generation.

TSTMotion: Training-free Scene-aware Text-to-motion Generation OmniControl: Control Any Joint at Any Time for Human Motion Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.661146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.661146Z digest=sha256:31c251b797a5d396e3277625facf1bd11f9bd9a3ea3465ae10816e94fd5cac07

Observation b30f74f7-eded-4866-ae67-14fd6d5e0003 · outbound

This paper cites Resolving 3D human pose ambiguities with 3D scene con- straints,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Resolving 3D human pose ambiguities with 3D scene con- straints,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.864474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:30:08.666034Z digest=sha256:952689646dc932176bc7b977d064a1201bf3edaf8d9afb2562d0955a34d3b113

Pith citing papers

Observation c5fd513e-8e1d-4682-82d2-a1e2526fe9a9 · inbound

MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm cites this paper.

MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm TSTMotion: Training-free Scene-aware Text-to-motion Generation

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-09T12:31:22.999369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T12:31:22.605420Z digest=sha256:64f51d7c5f9c8f3c4eef26e2b81066902d0d3577d3d0a9cd86535616990d3aca