Pith. sign in

Paper Citation Record · LEDGER

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis

As of 21 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 1 inbound Pith citation observation for arXiv:2506.17912.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.17912 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:27:45.288936Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-13T20:16:35.411267Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T20:18:13.202646Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact1
  • verified fuzzy29
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 90a763a1-4952-46ba-bd87-8a8c9bbd7aad · outbound

This paper cites Motion Flow Matching for Human Motion Synthesis and Editing.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Motion Flow Matching for Human Motion Synthesis and Editing

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:40.790842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:40.790842Z digest=sha256:2a8fcd05bb46e11b77d6405f23d2d543a64de8213440e1167cb5a6a6cc4dbc90

Observation 57669eca-3d83-4974-9a06-9ffab7a0cc2d · outbound

This paper cites GPT-4 Technical Report.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:40.884564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:40.884564Z digest=sha256:3d505cf3bfd39bd7600d59dd1fc40a1fee3479484cd0cdfb12e8b27aa594a203

Observation 1f6e173e-f83a-4c82-8e5d-4701e13d8fb2 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736, 2022.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736, 2022

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:40.953922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:40.953922Z digest=sha256:ed3eba82430d7f10d0734a4ad2c6d0a375efcbb4b6d1e4a4914308c9c37c1167

Observation 816bef85-a25e-4377-abc5-bd7efa5c1b7a · outbound

This paper cites Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:46.006530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:41.030977Z digest=sha256:ac8d0e9651fdbca1b7689cacd3119a6509c554a4b9b0bde2f7b5f645d9165ed5

Observation 7a526be9-24ed-4d37-92df-ea0fc9ce24b0 · outbound

This paper cites Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901, 2020.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901, 2020

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:41.111558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:41.111558Z digest=sha256:3ada739e8a3b687f4ca9f3b43f79e414733fbec1fae722aee3aa93c550f2cfb1

Observation 78b91fcb-316a-40c2-a690-9e5016e15408 · outbound

This paper cites Long- term human motion prediction with scene context.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Long- term human motion prediction with scene context

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.985661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:41.177602Z digest=sha256:1b9a08dba9183f9687199b469acd41688764833ef6d9e6aa72a0cb4b5d49e4f3

Observation 3ce38eb4-dcee-4427-8e8c-c372f81d34f5 · outbound

This paper cites Cmu graphics lab motion capture database, 2003.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Cmu graphics lab motion capture database, 2003

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.973174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:41.269465Z digest=sha256:4b6817ef302bc74b7accd758434256d0cc787091e2d9699d904358871177dff8

Observation ecd3cc6c-2636-43ba-8afd-fe32c3b7ee89 · outbound

This paper cites Matterport3D: Learning from RGB-D Data in Indoor Environments.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Matterport3D: Learning from RGB-D Data in Indoor Environments

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:41.329444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:41.329444Z digest=sha256:a7d61d6900a420802723f439f78b698d4e14d538970f07b5b14c095b774035a3

Observation 7696776e-6263-4799-a737-657b688c2ee8 · outbound

This paper cites Executing your commands via motion diffusion in latent space.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Executing your commands via motion diffusion in latent space

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:41.368660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:41.368660Z digest=sha256:70fd5cb4c969a4cc9a9cbff4f314984c10a5b41d2091f4f3ebdd30be32444a20

Observation 5de944e7-8c27-4f3c-a07f-2b2f8248296f · outbound

This paper cites Generative adversarial graph convolutional networks for human action synthesis.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Generative adversarial graph convolutional networks for human action synthesis

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.951577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:41.450727Z digest=sha256:dc7ef59a0a479564b7d7acac881c080ffc042117b9e19258e8e967d38978b02a

Observation b013ad27-0cd9-4466-88a4-132ced0e390c · outbound

This paper cites Learning individual styles of conversational gesture.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Learning individual styles of conversational gesture

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.939695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:41.496393Z digest=sha256:f278d515877169d75f2288274f6177e6c7b0a889993cd6a449692cd10029b70d

Observation 41102238-5526-4254-9e40-284c4da4ad7d · outbound

This paper cites Generative adversarial nets.Advances in neural information processing systems, 27, 2014.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Generative adversarial nets.Advances in neural information processing systems, 27, 2014

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:41.563826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:41.563826Z digest=sha256:a7bb43e819eb0d6e375c5d42036e5e13bdfae2d4cff5dde77c0004e23998cc54

Observation e96f607d-db40-4bd0-8e15-07a904e7c950 · outbound

This paper cites Momask: Generative masked modeling of 3d human motions.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Momask: Generative masked modeling of 3d human motions

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:41.655390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:41.655390Z digest=sha256:467945876065db34388a240be03be5ec4225fc4a746233694b2c01dc01dd306c

Observation f6e7729f-74e0-4049-ba57-4573d64075bc · outbound

This paper cites Generating diverse and natural 3d human motions from text.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Generating diverse and natural 3d human motions from text

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.902606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:41.746624Z digest=sha256:7ad0e7d5354e83ea0e83388463c93da2155fcd40928cc3bf8658476e1106faee

Observation 053a7caf-3f12-40ed-b05e-44155b30ce1e · outbound

This paper cites Action2motion: Conditioned generation of 3d human motions.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Action2motion: Conditioned generation of 3d human motions

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:41.807511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:41.807511Z digest=sha256:f740d912268834097821b5be34ac8df4b9e7ef241c959487b2d1bbf41d3da01e

Observation 9d8e9e12-bc80-40e8-bf60-71324fefe8a8 · outbound

This paper cites Stochastic scene-aware motion prediction.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Stochastic scene-aware motion prediction

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.880994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:41.897057Z digest=sha256:89b2ad2fa6885e31281e207866a5d8faf24c4017e9f04a4002dfc3b5d9c3631f

Observation 2d05f54c-11dc-4994-bee4-fa4265f943a8 · outbound

This paper cites Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:42.014387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:42.014387Z digest=sha256:764be16e78e9c38146f1ea1ecb3577b231caadcbc52ff34ed3f63b9ebf98a524

Observation 5d02a334-0f00-49fd-b8eb-6c78d08a11b6 · outbound

This paper cites Avatarclip: zero-shot text-driven generation and animation of 3d avatars.ACM Transactions on Graphics (TOG), 41(4):1–19, 2022.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Avatarclip: zero-shot text-driven generation and animation of 3d avatars.ACM Transactions on Graphics (TOG), 41(4):1–19, 2022

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.852164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:42.112072Z digest=sha256:e7d5acb21afa91d1897670dd819df60c7715e0df804ca58a779420f22c9c3d76

Observation 2c1ea0fd-ffd0-4bb6-a55e-352daccc1e77 · outbound

This paper cites Motiongpt: Human motion as a foreign language.Advances in Neural Information Processing Systems, 36:20067–20079, 2023.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Motiongpt: Human motion as a foreign language.Advances in Neural Information Processing Systems, 36:20067–20079, 2023

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:42.194414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:42.194414Z digest=sha256:c08cee4ada6d2613fb8fc52ae30a46f01b1477b24b57356553c902a0e3c49bb2

Observation 654560e6-27ae-4bfc-af7f-be440fa7ce2f · outbound

This paper cites Auto-Encoding Variational Bayes.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Auto-Encoding Variational Bayes

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:42.307504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:42.307504Z digest=sha256:24e43bdbd63c5a70c56a25194193be3f05b72117966a1879cb6c1231609b1729

Observation 32780bfb-008c-42f1-a852-2bd3bf9cf993 · outbound

This paper cites Generating images with multimodal language models.Advances in Neural Information Processing Systems, 36:21487–21506, 2023.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Generating images with multimodal language models.Advances in Neural Information Processing Systems, 36:21487–21506, 2023

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.833196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:42.414811Z digest=sha256:40bf5674d03ba7b2464dd82d2164995cfbcac337867031f068d6c8510233ea48

Observation ff3a0ceb-25e9-42d8-abee-24ed46e8185d · outbound

This paper cites Large language models are zero-shot reasoners.Advances in neural information processing systems, 35:22199–22213, 2022.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Large language models are zero-shot reasoners.Advances in neural information processing systems, 35:22199–22213, 2022

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:42.495149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:42.495149Z digest=sha256:55d980436258a265fc3a33ae4f6eb7ffc552103446eb58eb13ab73a8cba50338

Observation 25609d26-928c-4375-bb12-5065c8a8095f · outbound

This paper cites Analyzing input and output representations for speech-driven gesture generation.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Analyzing input and output representations for speech-driven gesture generation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.807207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:42.550917Z digest=sha256:41cb297b2076e4a0f3944b2f4a886ddfaab71b1b15a158b0abb186d5352edc8d

Observation 79e64c09-2848-47f4-9918-95d86b730bc5 · outbound

This paper cites Au- dio2gestures: Generating diverse gestures from speech audio with conditional variational autoencoders.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Au- dio2gestures: Generating diverse gestures from speech audio with conditional variational autoencoders

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.792121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:42.689624Z digest=sha256:086b988b581b6b5d3fb69cd809a72aec4d5c0e92048546e798e970e6a70a1801

Observation d87e30b9-27d3-484c-babf-53422a71627e · outbound

This paper cites Infinite Motion: Extended Motion Generation via Long Text Instructions.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Infinite Motion: Extended Motion Generation via Long Text Instructions

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:27:45.407544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:42.736779Z digest=sha256:27799a845edb3f910872cbede75b54be638cbdae979ba8a9d3403e46491e84dc

Observation ecdf3310-10a9-4923-b9ac-b9e703403049 · outbound

This paper cites Ai choreographer: Music conditioned 3d dance generation with aist++.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Ai choreographer: Music conditioned 3d dance generation with aist++

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.775916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:42.856241Z digest=sha256:017012330f736b70be881de3ee35bee2def59e9272b1fa79dd64b7c586bdb363

Observation 33abe0e4-c369-459e-a5be-8dd24423838e · outbound

This paper cites A survey of multimodel large language models.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis A survey of multimodel large language models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:42.936840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:42.936840Z digest=sha256:5d91df552fd0c12a39d9f2d5ffd8f3c4e4c595305b83309fb78308a418394476

Observation b00ffc27-4117-40dd-b262-b58435d78abd · outbound

This paper cites Flow Matching for Generative Modeling.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Flow Matching for Generative Modeling

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:43.009675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:43.009675Z digest=sha256:0949f20a98f9314746318be9fedfb554a0dd1699a182f129a57bb1a293588a24

Observation 8b592e95-dba2-4068-a28e-05b0e1a81b38 · outbound

This paper cites Beat: A large-scale semantic and emotional multi-modal dataset for conversa- tional gestures synthesis.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Beat: A large-scale semantic and emotional multi-modal dataset for conversa- tional gestures synthesis

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:43.127714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:43.127714Z digest=sha256:263f3a3adc8f4b61c357f325f9286bf0ffcf2c569c8a72c24506858985c8fe72

Observation 95e183fa-59da-4890-9eb4-b387a9854c97 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:43.227851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:43.227851Z digest=sha256:f858dad8c9586dfa4313bccba19d899eb8ad5d0e1a65e237aa57c96616bfc66f

Observation b074d865-de4b-473e-be3f-77e84cecbd71 · outbound

This paper cites Amass: Archive of motion capture as surface shapes.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Amass: Archive of motion capture as surface shapes

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.740960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:43.370490Z digest=sha256:8ba5c25026625c5ef607dc2f439032d67694b6562cde41ee042264b6f8479c6b

Observation dafbfee2-9fa1-492e-8948-23691e63c9d9 · outbound

This paper cites The kit whole-body human motion database.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis The kit whole-body human motion database

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:43.440147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:43.440147Z digest=sha256:8068f44321189888a9180dd250f47abaab0093e6de6c8d42f6541754911b03de

Observation 4f1350f8-20f1-47af-ad4b-0d602b3eb3c7 · outbound

This paper cites Gan-based reactive motion synthesis with class-aware discriminators for human–human interaction.Computers & Graphics, 102:634–645, 2022.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Gan-based reactive motion synthesis with class-aware discriminators for human–human interaction.Computers & Graphics, 102:634–645, 2022

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.722450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:43.571777Z digest=sha256:f7ace907de69ef322a61539f4d114f0327620ccbc7ca2ef11e7b679999861bbc

Observation 9c951a02-b31e-4534-b591-8f36d55fed0d · outbound

This paper cites Action-conditioned 3d human motion synthesis with transformer vae.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Action-conditioned 3d human motion synthesis with transformer vae

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.708698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:43.675630Z digest=sha256:d217b931cf959d19d6e5c1e826f5db85a62c665690f9d4f4cfe85766890bee36

Observation e68d06b3-b7a4-4f46-9da2-5b8a3a7add91 · outbound

This paper cites Bamm: Bidirectional autoregressive motion model.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Bamm: Bidirectional autoregressive motion model

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.695175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:43.787534Z digest=sha256:4c6379f6e547cf3fba30b5c86573ca5bc58b3b778b9b73d37e5d06d9fd111332

Observation d9dacdf2-b652-4c66-9f2a-a738cfa11e26 · outbound

This paper cites Babel: Bodies, action and behavior with english labels.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Babel: Bodies, action and behavior with english labels

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.681363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:43.948822Z digest=sha256:c18f1c182b81663960fe1d60cce609c666013bb3b3893b3af711cd36c205d04a

Observation 1f2469ff-d6d8-417d-bbb7-8ec95f704fde · outbound

This paper cites Learning transferable visual models from natural language supervision.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Learning transferable visual models from natural language supervision

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:43.968345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:43.968345Z digest=sha256:c6935f1656a67e49758f98ef9b1e16e9741c5a67f430ebd04465bddb4b452cfd

Observation d82057e1-efa5-4331-92ed-b515759d9f69 · outbound

This paper cites U-net: Convolutional networks for biomedical image segmentation.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis U-net: Convolutional networks for biomedical image segmentation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:44.043467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:44.043467Z digest=sha256:0d17c9dd7f0518d05a4dbb85993de0b39197f5e8f52d706e0bbc71d681d8f3b7

Observation 31529926-e519-4b02-a244-9a8a2030b739 · outbound

This paper cites Human motion diffusion as a generative prior.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Human motion diffusion as a generative prior

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:44.142452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:44.142452Z digest=sha256:343ee2372d2477cd4e70b16fe63423b85c568cd9ca833b0bc8f2570ace3cb1a6

Observation 9bdc2a60-e9bc-497c-840b-5855fc38fc15 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Score-Based Generative Modeling through Stochastic Differential Equations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:44.186577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:44.186577Z digest=sha256:b2c1f8a40e6b04d558af470a6dd6d24fd3369b8aa6989413989507050b1cc1c3

Observation 8bc70c01-ae56-4358-9f5a-7df27e6577ef · outbound

This paper cites Goal: Generating 4d whole-body motion for hand-object grasping.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Goal: Generating 4d whole-body motion for hand-object grasping

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.648816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:44.318731Z digest=sha256:21b9a6f907841a3debc46b78d5f4d7a734b5434ff0eb8d66fdc76463def0c871

Observation d831c7a9-40fb-42f7-8dc7-45d51cdbcd4e · outbound

This paper cites Motionclip: Exposing human motion generation to clip space.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Motionclip: Exposing human motion generation to clip space

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.636282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:44.431258Z digest=sha256:3a599f9e969fc7c42a9d4f91b409e25b43a1caa21d417eada85577e5a0de449e

Observation b8635515-863b-43b7-b7f6-aa9f4b806f47 · outbound

This paper cites Human motion diffusion model.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Human motion diffusion model

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:44.557955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:44.557955Z digest=sha256:bf5381f74188a60aed31a0ae3b7d2f0fbf6d9950c9f505fe7ddcc1cd245ac240

Observation 4e695a6a-a812-4826-a7c8-001b8cafed4f · outbound

This paper cites Neural discrete representation learning.Advances in neural information processing systems, 30, 2017.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Neural discrete representation learning.Advances in neural information processing systems, 30, 2017

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:44.720922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:44.720922Z digest=sha256:0d7c7100d80a6aab2aa038c5ecec3ce478f3c74695f18770e71bb58da46b42d6

Observation 7a406796-3a34-4788-9783-7bbac4e63607 · outbound

This paper cites Scene-aware generative network for human motion synthesis.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Scene-aware generative network for human motion synthesis

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.600223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:44.891281Z digest=sha256:3a10c826e278cfce141b30331017e7cacd43071267cba899eed65d4c7e458703

Observation 44dab7cc-acd6-4a79-9d92-9c0cb37d0b3f · outbound

This paper cites Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:45.028067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:45.028067Z digest=sha256:cb98e750246e0e8eed89f5b7d9a0f9af9a287b19ac890f47de95a2e6b55903c2

Observation e9c64be4-d6a7-4840-b204-a360bf2d2205 · outbound

This paper cites MotionGPT-2: A General-Purpose Motion-Language Model for Motion Generation and Understanding.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis MotionGPT-2: A General-Purpose Motion-Language Model for Motion Generation and Understanding

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:45.119692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:45.119692Z digest=sha256:d599bf37b2c51a551d64927896c6874650a8e92087561b393815e976d739eab4

Observation 042684c2-8fef-4d0e-9957-3cc89415d60a · outbound

This paper cites Emergent Abilities of Large Language Models.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Emergent Abilities of Large Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:45.217350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:45.217350Z digest=sha256:08c69741b01cc69f62987b54159e957517232df4e5af6476d1b36d5ad035ade5

Observation 34dfba78-1941-413e-94be-e7b21e2a4062 · outbound

This paper cites What are diffusion models?lilianweng.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis What are diffusion models?lilianweng

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.587048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:45.241093Z digest=sha256:a0a9faf4bce549383d3eb529a8dd37b420c651c5416c060905bc2860f3777e2d

Observation 638ead98-4c4a-4941-ad2f-509994d3bf05 · outbound

This paper cites Motion- agent: A conversational framework for human motion generation with LLMs.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Motion- agent: A conversational framework for human motion generation with LLMs

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.573623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:45.244205Z digest=sha256:63a598354fc5ae374f5e62ce606225a460a7099277c5f7205851a712b9476506

Observation c74c5ab4-2db8-40c0-aaf3-b51086f03b0d · outbound

This paper cites Saga: Stochastic whole-body grasping with contact.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Saga: Stochastic whole-body grasping with contact

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.560929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:45.247621Z digest=sha256:970fcd0fa80f4163cf6779465751f4bcd02b0dcc3f1afc060f1d0fbcbf2ed6c2

Observation 35a731e4-c5aa-450d-9f43-2249fe1a2292 · outbound

This paper cites Actformer: A gan-based transformer towards general action-conditioned 3d human motion generation.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Actformer: A gan-based transformer towards general action-conditioned 3d human motion generation

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.546818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:45.250954Z digest=sha256:17926f132c1cd826c4a42ed9041614a112179ccbe6bebe97c555481e32b46e6c

Observation 3a5542d0-66f4-4b15-a776-fec27c11f0b1 · outbound

This paper cites Qpgesture: Quantization-based and phase-guided motion matching for natural speech- driven gesture generation.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Qpgesture: Quantization-based and phase-guided motion matching for natural speech- driven gesture generation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.532734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:45.253922Z digest=sha256:eb3532e382a86513d391192618232bf474d47f29a6d921294e10c8be05c76de2

Observation 77c05760-ddcd-4ade-84b1-06193add16bd · outbound

This paper cites Structure-aware human- action generation.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Structure-aware human- action generation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.518658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:45.257282Z digest=sha256:b54105682ce581141c21aa792662fd7c3df7d10ccbb46b7a9af20a2a73527f4d

Observation 02036f03-6e50-4726-acf0-08ef561ae583 · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:45.261034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:45.261034Z digest=sha256:0524d2494f7a492fbed7181b4f498b7c4db42348844b6d04691feedc09f55f2f

Observation 1ed6f148-46f4-4da3-b350-93188cca6e25 · outbound

This paper cites T2m-gpt: Generating human motion from textual descriptions with discrete representations.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis T2m-gpt: Generating human motion from textual descriptions with discrete representations

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:45.264481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:45.264481Z digest=sha256:0cd6982084f142d0ad3c89098a6762e6342c18863da6a816be59acb7122fa120

Observation 1534295e-ad2a-41a2-9952-083cc4e500c8 · outbound

This paper cites MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:45.267627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:45.267627Z digest=sha256:7c9f9d138c9f09c93aa2b942ffc92b11c11435ba9c5b314e06108f147a57c8bc

Observation d9d41000-93b0-4a9c-a6cb-628446c0ad75 · outbound

This paper cites Remodiffuse: Retrieval-augmented motion diffusion model.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Remodiffuse: Retrieval-augmented motion diffusion model

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.497864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:45.271788Z digest=sha256:68ceab4593ab95aaa090398db3c132af0129f5897cc59f7a1c43c0d405b73ffe

Observation 0452802d-75d5-45c7-838f-1aba360c4ef1 · outbound

This paper cites Finemogen: Fine-grained spatio-temporal motion generation and editing.NeurIPS, 2023.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Finemogen: Fine-grained spatio-temporal motion generation and editing.NeurIPS, 2023

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:45.275289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:45.275289Z digest=sha256:9b4a61b664c756d3e046619a715182c31d3860047cc9d475f3b31b2a2fbc89f7

Observation 0b83881c-b5db-45fa-8ec5-3f2e9e434c08 · outbound

This paper cites Tinyllama: An open-source small language model, 2024.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Tinyllama: An open-source small language model, 2024

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:45.279431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:45.279431Z digest=sha256:bd57b96a1c3cca648003ce44e3f8b2f63ee1a77741d399aa8c0299f97419ebb5

Observation 7440acb5-c7bb-4952-844e-ea77fc968eba · outbound

This paper cites Large language models as commonsense knowledge for large-scale task planning.Advances in Neural Information Processing Systems, 36:31967– 31987, 2023.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis Large language models as commonsense knowledge for large-scale task planning.Advances in Neural Information Processing Systems, 36:31967– 31987, 2023

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.470701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:45.282495Z digest=sha256:f5a2f6005a5c57555df16ad94e88eb8f9410c348ed0b75d82fcf042f1ec4433b

Observation e255058e-104e-46f3-92d1-93423f7ee31e · outbound

This paper cites LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:45.285511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:45.285511Z digest=sha256:4b7035a44fea959337a75cda8dc54d93d3035fbc0111b6a24738883aae90f3b7

Observation 4aab62f8-9254-4e92-8f11-be20860702c7 · outbound

This paper cites base” for “residual.

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis base” for “residual

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:27:45.458664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:27:45.288936Z digest=sha256:dcbb58f426946d3f2808ebee8a9041adf61b0a96cc50a141f25b9f52688bdf52

Pith citing papers

Observation bbbe2e75-3a77-4fa9-bc92-ab3620eeb88b · inbound

SentiAvatar: Towards Expressive and Interactive Digital Humans cites this paper.

SentiAvatar: Towards Expressive and Interactive Digital Humans PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:18:13.204278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T20:16:35.411267Z digest=sha256:dd87ec7f9dd3e853cc92e327e448f7288ba0b45c7a85322820ce54f7b0ad99a4