Pith. sign in

Paper Citation Record · LEDGER

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling

As of 9 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 1 inbound Pith citation observation for arXiv:2502.08556.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.08556 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T04:41:08.519954Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T15:44:02.709823Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T09:56:04.285692Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact5
  • verified fuzzy32
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1a1830a2-d5fd-4548-b3a4-918e82ab271e · outbound

This paper cites GPT-4 Technical Report.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T04:41:08.377108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:41:08.377108Z digest=sha256:904cc136c640439c1b43782c8b625851cf50f56cd716cb5646b1be9f732bc0b3

Observation d0f3106d-b488-4ea5-bff3-d073a315d7ec · outbound

This paper cites A morphable model for the synthesis of 3d faces.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling A morphable model for the synthesis of 3d faces

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:09.418696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.388121Z digest=sha256:f511b05a2dfc699f526c3fe160918a447790e0d3acd286e506684a5310e839a9

Observation 9b62267d-bf78-4841-ae6f-6e0d3b0ef9a2 · outbound

This paper cites Instructpix2pix: Learning to follow image editing instructions.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Instructpix2pix: Learning to follow image editing instructions

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:09.392174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.391669Z digest=sha256:3aa16dc9ea15939899b775a8c4c57e3c44b8962373defcb27f12ecf5b39ed6cb

Observation 68db0431-19e6-4f70-b238-ff4cfbfd8ca3 · outbound

This paper cites Smpler-x: Scaling up expressive human pose and shape estimation.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Smpler-x: Scaling up expressive human pose and shape estimation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:09.340011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.394613Z digest=sha256:0f2b2f5e3b26db592513b757b50c8dac9759afd02279c66e5f24403b3131795f

Observation 32cf0edd-7385-46cc-b956-126e241d7e5b · outbound

This paper cites Sofgan: A portrait image generator with dynamic styling.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Sofgan: A portrait image generator with dynamic styling

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:09.224513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.397844Z digest=sha256:9404e09d815a1a119fb48eb8b0e120f32bd6de6b47c1ff1aad7f0a72f9116258

Observation 668e820b-f5fe-468d-ace7-1cb4f268bdb1 · outbound

This paper cites The Language of Motion: Unifying Verbal and Non-verbal Language of 3D Human Motion.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling The Language of Motion: Unifying Verbal and Non-verbal Language of 3D Human Motion

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T04:41:08.400750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:41:08.400750Z digest=sha256:1a74be8e496ad1ef360acd2c8ae0b64abfaf5430ce2edc5b00dbe2c9f1add392

Observation 25e9178b-06ff-455f-a36b-c67b6067d624 · outbound

This paper cites Unihcp: A unified model for human-centric perceptions.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Unihcp: A unified model for human-centric perceptions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:09.092314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.403764Z digest=sha256:862ef966750086b62470444401c3d932ff837e365afce5388b6335e0673c0181

Observation 2c08b5d9-8326-461d-9bea-bacd7bd35446 · outbound

This paper cites Ag3d: Learning to generate 3d avatars from 2d image col- lections.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Ag3d: Learning to generate 3d avatars from 2d image col- lections

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:09.000012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.406354Z digest=sha256:a635271b6e1fccc871bb7b2d0050d050f3b32afc3c22994c2e904b414f7b9776

Observation 3ce47ec6-abae-4151-a1fa-e9f7d1ae4e55 · outbound

This paper cites Bringing Robots Home: The Rise of AI Robots in Consumer Electronics.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Bringing Robots Home: The Rise of AI Robots in Consumer Electronics

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-08T04:41:08.663866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.409033Z digest=sha256:ad35d60fa7dc77dbf81a718753ca551ac074e1a98fee857edbd1a9df64d2c0f0

Observation 26ea5f40-56fb-4af6-b24a-a15b6429f6ff · outbound

This paper cites Foundation Models for 3D Humans,.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Foundation Models for 3D Humans,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.978794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.412412Z digest=sha256:04ebf6f8303e97e1cd910307e7b6a48c79f4a71bd61f2d468c6f8bfe1b1b0b92

Observation 1e2bb3b0-5d9f-4dd5-9f03-cd9497290edb · outbound

This paper cites Chatpose: Chatting about 3d human pose.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Chatpose: Chatting about 3d human pose

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.971308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.415119Z digest=sha256:c560b5d6904256c1dce00cbb12b36db6263c16fb0beadb252543a6e5a1530e24

Observation 51cdde9c-17a8-4469-9fae-0796167e5570 · outbound

This paper cites Stylegan-human: A data-centric odyssey of human generation.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Stylegan-human: A data-centric odyssey of human generation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.962826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.417883Z digest=sha256:5f16bd8c0a9c52f97e5e9b8322dc7d60048696bfa80992e38e5262e9999c2aaa

Observation c179ddcc-5116-405a-869b-b06b9dd8c5ea · outbound

This paper cites Instruct-reid: A multi- purpose person re-identification task with instructions.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Instruct-reid: A multi- purpose person re-identification task with instructions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.945694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.423883Z digest=sha256:6c8748151513693629c8287e479a181a63cacdc98d09117ec7a744be42932093

Observation f2a3aca0-9532-44e7-8bae-6c7ed63de1dc · outbound

This paper cites Versatile multi-modal pre-training for human-centric perception.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Versatile multi-modal pre-training for human-centric perception

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.937719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.426774Z digest=sha256:14877bff5515df76fa16b88745c75378954e5ef25f77d89ac3df16035f1187cd

Observation 3c50a14d-6a28-4a26-8cf0-931df301697a · outbound

This paper cites Animate anyone: Consistent and control- lable image-to-video synthesis for character animation.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Animate anyone: Consistent and control- lable image-to-video synthesis for character animation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.929237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.429347Z digest=sha256:9a41db12a306d15ee4f835dc18639f250098d9d1038732c87e4ffa04e28ee829

Observation bcf3f852-e480-42f0-b09a-17a317120d32 · outbound

This paper cites RefHCM: A Unified Model for Referring Perceptions in Human-Centric Scenarios.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling RefHCM: A Unified Model for Referring Perceptions in Human-Centric Scenarios

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-08T04:41:08.652133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.431780Z digest=sha256:09bc0c637d8eb34ba0474920ebcef26da8c0b1747062a1b2e6f23d6eb376cdb1

Observation e5524322-98c8-42f8-a06e-c0a84a2206f2 · outbound

This paper cites Motiongpt: Human motion as a foreign language.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Motiongpt: Human motion as a foreign language

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.920111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.435204Z digest=sha256:d7b8e5768c3f9d2d7f31d8a3973a1f426c76875df02ae6e8f2184d0dc34f73ff

Observation 7d44acf9-6e04-4ece-816a-1f87c0349b6c · outbound

This paper cites You only learn one query: learning unified human query for single-stage multi-person multi-task human-centric perception.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling You only learn one query: learning unified human query for single-stage multi-person multi-task human-centric perception

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.907445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.438420Z digest=sha256:1481b061ed33835a66193e650e9e1a240b2b39290ec825cb7a022e93216f7ddd

Observation 2812a0b5-a4c3-420a-af3d-797bd727cf4c · outbound

This paper cites Humansd: A native skeleton-guided diffusion model for human image generation.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Humansd: A native skeleton-guided diffusion model for human image generation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.898770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.441074Z digest=sha256:14c93108eb425ae14ed67bc42f73d9fd8b67bd2acaa91dac0ae13a4fe8356887

Observation 4fb847e5-10f6-402c-b44e-b61631a0d4e5 · outbound

This paper cites Superpadl: Scaling language- directed physics-based control with progressive supervised distillation.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Superpadl: Scaling language- directed physics-based control with progressive supervised distillation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.846654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.444889Z digest=sha256:704b1d2ed3d65acb03a54d0f69ef178f20b489d2601a942a404363d32c502a57

Observation 598fbca9-d1cb-46d8-b0c4-d0e7165a1960 · outbound

This paper cites Fashion- vdm: Video diffusion model for virtual try-on.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Fashion- vdm: Video diffusion model for virtual try-on

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.807616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.447499Z digest=sha256:4f423745bd244f172a67e92ed7a8cc3e630dde14f4176808d25542796fc97806

Observation 9731220b-9840-4e9f-a763-8998a8a61f8d · outbound

This paper cites Sapiens: Foundation for human vision models.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Sapiens: Foundation for human vision models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.799224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.450493Z digest=sha256:3f19f34defad14376733a5a77960c3b1c7efc050ed0f5a8ac663a854121032f6

Observation cad1d602-586b-4794-9275-770b85e7a134 · outbound

This paper cites DreamHuman: Animatable 3D Avatars from Text.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling DreamHuman: Animatable 3D Avatars from Text

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-08T04:41:08.640613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.453676Z digest=sha256:721f7101b5759846b1b8ee39d1b7d01c533e51622627cf2b6ecd2eafc5b901d6

Observation 58e59c0b-f90e-4ba0-b0f8-8734290830df · outbound

This paper cites Blip-2: Bootstrapping language-image pre- training with frozen image encoders and large language models.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Blip-2: Bootstrapping language-image pre- training with frozen image encoders and large language models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.790976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.457353Z digest=sha256:0da68f1d53f6dd6ff3a6300d389fb7be0df390917080558bfc3e02b52cfcc3f1

Observation deb65451-07b8-4f89-b9a5-20ee598e4a3e · outbound

This paper cites UniPose: A Unified Multimodal Framework for Human Pose Comprehension, Generation and Editing.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling UniPose: A Unified Multimodal Framework for Human Pose Comprehension, Generation and Editing

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T04:41:08.459981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:41:08.459981Z digest=sha256:082b02543a4bd4dec4eab9519348497666ad3aa1a22983fd7c106dea97b4e9b5

Observation b1e3a915-0735-49ae-b470-3ee2df4cb797 · outbound

This paper cites Motion-x: A large-scale 3d expressive whole-body human motion dataset.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Motion-x: A large-scale 3d expressive whole-body human motion dataset

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.782830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.463040Z digest=sha256:c02b7d36719c79e77507d0a81dd3c699d2df9d230ddcf4dcfdeef40c992017fa

Observation af07394e-ec87-4a02-849b-c17cbed32da1 · outbound

This paper cites ChatHuman: Chatting about 3D Humans with Tools.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling ChatHuman: Chatting about 3D Humans with Tools

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T04:41:08.465666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:41:08.465666Z digest=sha256:8c8915bd256e679802a27cc041909cf0a01a43da66830052b4d0ddbf1e3c8b52

Observation 3e8d6917-fac6-44a2-8139-c7ef4e1a200d · outbound

This paper cites Omnihuman-1: Rethinking the scaling-up of one- stage conditioned human animation models,.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Omnihuman-1: Rethinking the scaling-up of one- stage conditioned human animation models,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.774839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.468306Z digest=sha256:ef5cee0f614a8f7461a18c832f9f15cbcdba6f6b2e03e4131e0e6cc1958775c3

Observation 76cfbeeb-ab2e-41f0-8f02-b8f94ec52a98 · outbound

This paper cites Visual instruction tuning.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Visual instruction tuning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.766155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.474089Z digest=sha256:4b826b4c7131d397b40e8fa30a6619939867fdf49e62786cda1773417ca3e452

Observation 59dec457-e653-43bc-af4e-f45b5c87d6de · outbound

This paper cites M$^3$GPT: An Advanced Multimodal, Multitask Framework for Motion Comprehension and Generation.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling M$^3$GPT: An Advanced Multimodal, Multitask Framework for Motion Comprehension and Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T04:41:08.476445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:41:08.476445Z digest=sha256:6a19fcf4b5e2ada0f4081f4337eccb1393cdb66ee5ba9be912710fa3545d805c

Observation 077e29e9-043b-4834-8f28-965af779d315 · outbound

This paper cites Efficient multi-modal human-centric contrastive pre-training with a pseudo body-structured prior.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Efficient multi-modal human-centric contrastive pre-training with a pseudo body-structured prior

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.757947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.479097Z digest=sha256:245ac9e6737cfa9c3431b4dd26e6c2e87f0cac1e326bddaf54f27c62937bbd5d

Observation 31941097-486a-428a-83e5-80d369282312 · outbound

This paper cites Drag your gan: Interactive point-based manipulation on the generative image manifold.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Drag your gan: Interactive point-based manipulation on the generative image manifold

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.749243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.481712Z digest=sha256:3a47aa94bafb6cbc22a5df62f407ecf543ce62f4736ed67c9fbda893ba9c7fae

Observation 42f79f43-4765-41c8-bdf3-1fb6e08eb202 · outbound

This paper cites 360-degree human video generation with 4d diffusion transformer.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling 360-degree human video generation with 4d diffusion transformer

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.740218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.484447Z digest=sha256:a29038b6bdb41455877e0d6a49ea0d43ef3104bfce7fd6fc3e6707c99848f9b1

Observation 6ff272ab-2943-475d-a030-bbba13f42eb2 · outbound

This paper cites Humanbench: To- wards general human-centric perception with projector as- sisted pretraining.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Humanbench: To- wards general human-centric perception with projector as- sisted pretraining

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.730286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.486921Z digest=sha256:26cc138598af09182eda90c0f71660f7b5025fa9bf9f81992a28be94ffe80d86

Observation f8244cab-e9bc-4662-bb34-007dedf8a584 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling LLaMA: Open and Efficient Foundation Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T04:41:08.489625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:41:08.489625Z digest=sha256:305c4dcabfe3c1cd6ed19ba0d5f1cda8b0ef22dbfe9de9c77598b99f82434af3

Observation 628b6da8-4e69-454c-8f60-45ae1ad5d54d · outbound

This paper cites Hulk: A Universal Knowledge Translator for Human-Centric Tasks.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Hulk: A Universal Knowledge Translator for Human-Centric Tasks

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T04:41:08.492729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:41:08.492729Z digest=sha256:564b965537cf79184605b6818102ab6451077809f81de23a84336da2f91a7e7c

Observation d036d4b3-77bd-4635-83ad-d8adbac3ca22 · outbound

This paper cites FaceGPT: Self-supervised Learning to Chat about 3D Human Faces.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling FaceGPT: Self-supervised Learning to Chat about 3D Human Faces

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-08T04:41:08.584013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.495779Z digest=sha256:25ae3f1536232c17dbf0db2881ee2bd3f36292b16716f4ac85ac493c16f5e5e0

Observation bcbbc957-0c2d-4229-a817-922fb49ecf40 · outbound

This paper cites MotionGPT-2: A General-Purpose Motion-Language Model for Motion Generation and Understanding.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling MotionGPT-2: A General-Purpose Motion-Language Model for Motion Generation and Understanding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T04:41:08.499116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:41:08.499116Z digest=sha256:4e03e868b435d161d3e334daac4b3c8c8bd9903b3728fcd2cceeecaacfd17cc3

Observation af5f1fb6-f059-4e8c-b57e-e36e7b61a9dc · outbound

This paper cites Aniportraitgan: animatable 3d portrait generation from 2d image collections.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Aniportraitgan: animatable 3d portrait generation from 2d image collections

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.722090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.502014Z digest=sha256:ec94bb598772f392e9168b9cc42d8d53d067cc9fa6bcd90202734c489acc622d

Observation 4c4a901d-d771-4107-8346-564711dc1b0c · outbound

This paper cites Motion-Agent: A Conversational Framework for Human Motion Generation with LLMs.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Motion-Agent: A Conversational Framework for Human Motion Generation with LLMs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T04:41:08.504904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:41:08.504904Z digest=sha256:dc281ba5e813436ecb4404ba799dcf11f830e93b395deedb5c3dee062c2e8a04

Observation 3c267c3e-08ae-4aa1-9654-4ec216898056 · outbound

This paper cites Get3dhuman: Lifting stylegan-human into a 3d generative model using pixel- aligned reconstruction priors.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Get3dhuman: Lifting stylegan-human into a 3d generative model using pixel- aligned reconstruction priors

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.713742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.507741Z digest=sha256:1892320d8219a4f7615f1bca5427a520359235e1f710b2ca3a9db06f0dae619a

Observation 27b19fc3-d202-4fd5-8e75-814db2c649fb · outbound

This paper cites HumanVLA: Towards Vision-Language Directed Object Rearrangement by Physical Humanoid.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling HumanVLA: Towards Vision-Language Directed Object Rearrangement by Physical Humanoid

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T04:41:08.510578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:41:08.510578Z digest=sha256:3c2aab96dd05a7c4337d49c860511c5a479671aa87ecef698cedd678ba16030b

Observation 861fa56b-80ea-4dd2-8c1c-92cdf519b075 · outbound

This paper cites StyleHumanCLIP: Text-guided Garment Manipulation for StyleGAN-Human.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling StyleHumanCLIP: Text-guided Garment Manipulation for StyleGAN-Human

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-08T04:41:08.547968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.513566Z digest=sha256:b9636918c65c83f6e50fafd035c49b3dec2bcc65d397b345a0d43f75a0306d4d

Observation abcf1d4f-85cf-4137-9ef2-069522cb4d2b · outbound

This paper cites Hap: Structure-aware masked image modeling for human-centric perception.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Hap: Structure-aware masked image modeling for human-centric perception

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.705280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.517546Z digest=sha256:2a2f49db4719445de6a01d69038cb875609ca8c38a471bccf0c3a10c00d79fb7

Observation eb642ab2-4af6-4676-903b-ae9515111fe9 · outbound

This paper cites Avatargpt: All-in-one framework for motion un- derstanding planning generation and beyond.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Avatargpt: All-in-one framework for motion un- derstanding planning generation and beyond

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.695948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.519954Z digest=sha256:f6e17b4f66a6174752b859e24190a5b18cca8caeb1d83d67db60a4067e806261

Observation d7aed88c-a3d2-4c9c-ad70-69d97b6723a5 · outbound

This paper cites Unitedhuman: Har- nessing multi-source data for high-resolution human gener- ation.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Unitedhuman: Har- nessing multi-source data for high-resolution human gener- ation

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:08.954775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.420447Z digest=sha256:a4ab7ee4cc4c94a4849ae8fa4726a730f2ca3574b271c01613eb8352bf5e3eaf

Observation a4cf39ba-56f8-4266-b381-1350e9f52606 · outbound

This paper cites Cross- view and cross-pose completion for 3d human understand- ing.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling Cross- view and cross-pose completion for 3d human understand- ing

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T04:41:09.432138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T04:41:08.380323Z digest=sha256:88515da7758304ffbd958be73060673ce2974500633b3bbfdcd674f13ee7a85e

Observation bb6479c3-a526-4dd0-89ac-590ea0b17a3d · outbound

This paper cites ChatGarment: Garment Estimation, Generation and Editing via Large Language Models.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling ChatGarment: Garment Estimation, Generation and Editing via Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T04:41:08.383912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:41:08.383912Z digest=sha256:72e12202bda727736f4019af30cba618bfd7f8ec5ab5648362d21a8d4a8dc037

Observation 83b6b5f0-890c-4ea3-9b97-504d4cd99d08 · outbound

This paper cites HyperHuman: Hyper-Realistic Human Generation with Latent Structural Diffusion.

Human-Centric Foundation Models: Perception, Generation and Agentic Modeling HyperHuman: Hyper-Realistic Human Generation with Latent Structural Diffusion

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-08T04:41:08.471237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:41:08.471237Z digest=sha256:a20f61c3eda94e3837eb87afae585a0cfd793624d3f77f65bc55a99964f36916

Pith citing papers

Observation ecec6641-0f10-4b85-9e6b-078593bce66c · inbound

VGA-Bench: A Unified Benchmark and Multi-Model Framework for Video Aesthetics and Generation Quality Evaluation cites this paper.

VGA-Bench: A Unified Benchmark and Multi-Model Framework for Video Aesthetics and Generation Quality Evaluation Human-Centric Foundation Models: Perception, Generation and Agentic Modeling

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:04.291715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:44:02.709823Z digest=sha256:80a5ee80508878414a7f2452d6a89d9313b7d9c8ae3c9db8c7a2f68cb8fe48e5