Pith. sign in

Paper Citation Record · LEDGER

TSTMotion: Training-free Scene-aware Text-to-motion Generation

As of 17 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 1 inbound Pith citation observation for arXiv:2505.01182.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.01182 v2

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:30:08.666034Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T12:31:22.605420Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-09T12:31:22.996104Z

Reference resolution

21 of 21 outbound references displayed

  • verified exact2
  • verified fuzzy12
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fd3315f5-7378-4237-9702-28c649bd874a · outbound

This paper cites Human Motion Diffusion Model.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Human Motion Diffusion Model

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.570066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.570066Z digest=sha256:17a4bc875ae8349ac78f0757628a8c730476ba0dac3f1aa9579fc2b3cbd1b825

Observation d0d205bb-2b98-4fad-9b65-1ae2d76af6f2 · outbound

This paper cites MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model.

TSTMotion: Training-free Scene-aware Text-to-motion Generation MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.575943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.575943Z digest=sha256:e1556e91d1431343923218b1adb434dbf1484aa8b1dddd56bf7b71a560ecae86

Observation e2ea81fd-e9af-45f8-b9f9-c37363aa2b04 · outbound

This paper cites Humanise: Language-conditioned human motion generation in 3d scenes,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Humanise: Language-conditioned human motion generation in 3d scenes,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:09.050825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.581072Z digest=sha256:12b595fb40e6a909987f6e829d718d4f60ee69a795a4fd3ebec1edd1b648d94b

Observation 8c4601ae-b98f-4d31-a7b3-8f030b49eda6 · outbound

This paper cites Unified human-scene interaction via prompted chain-of-contacts,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Unified human-scene interaction via prompted chain-of-contacts,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:09.029927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.585863Z digest=sha256:757cd6c7a25a2471b251a528ed4da3b5895e4c9dd3e3b2f5b87e339637a28992

Observation 716d6ba9-50fc-4aa4-9f34-680d1962ae0b · outbound

This paper cites Scaling up dynamic human-scene interaction modeling,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Scaling up dynamic human-scene interaction modeling,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:09.013724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.590267Z digest=sha256:fc624c27d51fffdba2741482306b775dba25b098f8f7841b58adddcb67490778

Observation a6ba2fe0-1bc6-47c8-9548-322343ba007e · outbound

This paper cites Move as You Say, Interact as You Can: Language-guided Human Motion Generation with Scene Affordance.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Move as You Say, Interact as You Can: Language-guided Human Motion Generation with Scene Affordance

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-16T04:30:08.815613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.594698Z digest=sha256:885ededd1169e3c36b7f94bf67f164309b3b064d908576cb7bbed24d9eb1c914

Observation bfc4b223-1e06-47bd-b239-88653af3e64f · outbound

This paper cites Stochastic scene-aware motion prediction,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Stochastic scene-aware motion prediction,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.996801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.599749Z digest=sha256:0542c432bb0f6ea6af5a8e956f6c5a90309d051e3d1786edd4d50c3ad74b5f78

Observation 1f00d3aa-c95e-4fcf-9c2e-b1767874bcec · outbound

This paper cites Synthesizing Diverse Human Motions in 3D Indoor Scenes.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Synthesizing Diverse Human Motions in 3D Indoor Scenes

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-16T04:30:08.791477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.603907Z digest=sha256:63a7d08ee0ff6447f3f2fe91476ae14b9f4ebd8147fb9760fcfac66c4293bfd0

Observation 3c318909-9016-4e78-afa6-f8a1662c8f94 · outbound

This paper cites Generating diverse and natural 3d human motions from text,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Generating diverse and natural 3d human motions from text,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.981152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.608284Z digest=sha256:2bdbd4a1cf72c7e0694e8adbc941a0434e8c36fb94e0a3dc937b512a928ded7c

Observation 34bee9e1-2f46-464a-a823-4e68cbf7a68e · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Chain-of-thought prompting elicits reasoning in large language models,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.965539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.612499Z digest=sha256:a2b32e436d69e489cbb0dd5645932f5eb39f1cc30cdb0d6fc4afa59f667ca2fe

Observation ed01635c-9c89-4851-9623-d37211f2dabd · outbound

This paper cites Visual language maps for robot navigation,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Visual language maps for robot navigation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.949133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.616570Z digest=sha256:9819817f30441105ae480a8dfa4c1531f8bf3f1eb9ff4843ab46d58747a01461

Observation 518fd29e-d7ad-4995-890a-159f650a59bf · outbound

This paper cites Tabllm: Few-shot classification of tabular data with large language models,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Tabllm: Few-shot classification of tabular data with large language models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.930160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.621528Z digest=sha256:57518be1bc680fe898d0d47f7b01aaa3f096153a8a680d61d280d2ba3eb2d413

Observation 54c274b1-f1b8-4b45-8ca3-0c7afa4a38ba · outbound

This paper cites Language models are few-shot learners,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Language models are few-shot learners,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.913206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.626397Z digest=sha256:6e4012c685aa7ee454d346f2d5d20a717f27a45521de8df8fbc8bfe390fdadb0

Observation 13398b34-1b1b-4fac-b1c3-1585fe712cc9 · outbound

This paper cites Diffusion Posterior Sampling for General Noisy Inverse Problems.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Diffusion Posterior Sampling for General Noisy Inverse Problems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.631348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.631348Z digest=sha256:284c455f047af40b2afb7e9d84a0e8378a827ea97f11778042edd4f2593aa0e4

Observation a0b71c18-cd0e-425b-96f2-c47bad26a9f0 · outbound

This paper cites Ex- pressive body capture: 3D hands, face, and body from a single image,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Ex- pressive body capture: 3D hands, face, and body from a single image,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.896164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.636071Z digest=sha256:4dfe5e44e3a8e884a04dbe785933adcead446e6ede8fd3e61fda005490c5793d

Observation 5fff4d97-50ce-4ba3-95bf-fee379698b1d · outbound

This paper cites Enhancing Human-Centered Dynamic Scene Understanding via Multiple LLMs Collaborated Reasoning.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Enhancing Human-Centered Dynamic Scene Understanding via Multiple LLMs Collaborated Reasoning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.641750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.641750Z digest=sha256:6d50c98456b0c61dc3a2c412a701e76e1f36c531dbdb97874322ac470f3fbfde

Observation 812ac5f0-b8c7-4317-81ad-978ac90dc382 · outbound

This paper cites Recognize anything: A strong image tagging model,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Recognize anything: A strong image tagging model,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.880240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.646732Z digest=sha256:e26f6c67ff26edb182891bd7a73baa58ac9d628031868df7f85c0c4e42731e55

Observation 62d47bc0-47f7-4715-8099-4fe9dddd01e4 · outbound

This paper cites OpenIns3D: Snap and Lookup for 3D Open-vocabulary Instance Segmentation.

TSTMotion: Training-free Scene-aware Text-to-motion Generation OpenIns3D: Snap and Lookup for 3D Open-vocabulary Instance Segmentation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.651493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.651493Z digest=sha256:abf8c537c0dc5241cae099017dd9aeeac9839b02078da359e16b83cb68e85759

Observation a5d525c9-f8b4-408e-9cd7-326d0be5b0b8 · outbound

This paper cites GPT-4 Technical Report.

TSTMotion: Training-free Scene-aware Text-to-motion Generation GPT-4 Technical Report

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.656276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.656276Z digest=sha256:80fb4f75efe893baf8633f8005171c82fc94096418a61d11197ae66adbfc8015

Observation 4c17d583-9859-4e26-b80b-34b2c3a4d773 · outbound

This paper cites OmniControl: Control Any Joint at Any Time for Human Motion Generation.

TSTMotion: Training-free Scene-aware Text-to-motion Generation OmniControl: Control Any Joint at Any Time for Human Motion Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T04:30:08.661146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:30:08.661146Z digest=sha256:5712c15146fbb771942189a66f733004753c064f993b4909622fb9e142243ac2

Observation b30f74f7-eded-4866-ae67-14fd6d5e0003 · outbound

This paper cites Resolving 3D human pose ambiguities with 3D scene con- straints,.

TSTMotion: Training-free Scene-aware Text-to-motion Generation Resolving 3D human pose ambiguities with 3D scene con- straints,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:30:08.864474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T04:30:08.666034Z digest=sha256:493033d34047247282334cc7e470a55ab1bcb73ff4c38c722c6700b0e74502df

Pith citing papers

Observation c5fd513e-8e1d-4682-82d2-a1e2526fe9a9 · inbound

MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm cites this paper.

MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm TSTMotion: Training-free Scene-aware Text-to-motion Generation

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-09T12:31:22.999369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-09T12:31:22.605420Z digest=sha256:e27920a6542c5fe44f8fd26e7854690664f9cf04fcdcc4f5a60c3e5ab725aea9