Pith. sign in

Paper Citation Record · LEDGER

Sapiens: Foundation for Human Vision Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2408.12569.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.12569 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:53:00.932093Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:48:32.980001Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 06f81859-b2c0-45bc-8a19-15ccdac53684 · inbound

Generative Physical AI in Vision: A Survey cites this paper.

Generative Physical AI in Vision: A Survey Sapiens: Foundation for Human Vision Models

Reference 239

Resolution
unresolved
no resolver link, observed 2026-08-10T18:53:00.932093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:53:00.932093Z digest=sha256:d8c7ef8db7b9fbb2dcb49d47d3cc595e9ca8a129c3e4b557f6ef3bba1066ae82

Observation caa24bc2-58d8-48b4-b717-e5cceb5bdbca · inbound

HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation cites this paper.

HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Sapiens: Foundation for Human Vision Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T21:16:16.496425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:16:16.496425Z digest=sha256:87cd7fa0191aed5a87c23b92997f5201ce2dfb494b2f45db3533ba58fc07b8a1

Observation 4134d363-e3ab-498e-ad66-b05f272d8a05 · inbound

PoseBH: Prototypical Multi-Dataset Training Beyond Human Pose Estimation cites this paper.

PoseBH: Prototypical Multi-Dataset Training Beyond Human Pose Estimation Sapiens: Foundation for Human Vision Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:41.572907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:41.572907Z digest=sha256:c5208d8918ed8a5c61ef246965c675576a46bf967dafc9f160c74f180565b913

Observation 14e38772-f43b-48b1-a47e-406ecd8c8adb · inbound

Total-Editing: Head Avatar with Editable Appearance, Motion, and Lighting cites this paper.

Total-Editing: Head Avatar with Editable Appearance, Motion, and Lighting Sapiens: Foundation for Human Vision Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:27.939782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:27.939782Z digest=sha256:e593a606a1c3a3ee7f6599d642e8b46bdea426da19fbaa4623f2837035d8cb7f

Observation 062eb3d0-d154-444c-8b9b-5e2e1887a938 · inbound

HuGeDiff: 3D Human Generation via Diffusion with Gaussian Splatting cites this paper.

HuGeDiff: 3D Human Generation via Diffusion with Gaussian Splatting Sapiens: Foundation for Human Vision Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T10:50:50.117287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:50:50.117287Z digest=sha256:b3dc9cfec056e19861dc5685193825b40ac339aa2a46ec75b84cf98abee51b71

Observation e7518576-ac00-44e3-aad1-8dfeedd98dfa · inbound

Controllable and Expressive One-Shot Video Head Swapping cites this paper.

Controllable and Expressive One-Shot Video Head Swapping Sapiens: Foundation for Human Vision Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:42:33.980220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:42:33.980220Z digest=sha256:c05d46cf41356798bbcf4cc3f682cec9df90d1197a541067208b1303b3b6d474

Observation 816c7bcd-ba28-4ed8-8187-abb5b8b292a9 · inbound

DreamCube: 3D Panorama Generation via Multi-plane Synchronization cites this paper.

DreamCube: 3D Panorama Generation via Multi-plane Synchronization Sapiens: Foundation for Human Vision Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:05.494749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:35:05.494749Z digest=sha256:62db82bc41b2a0a0e660875b40909219a4a1602024c0746271d3f284e3208acc

Observation df1e7e3f-843f-4dd4-87b8-45e8c909f101 · inbound

Video Virtual Try-on with Conditional Diffusion Transformer Inpainter cites this paper.

Video Virtual Try-on with Conditional Diffusion Transformer Inpainter Sapiens: Foundation for Human Vision Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T22:33:30.636955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:33:30.636955Z digest=sha256:caa30773aadb2ec2a3eb95db72f53fd8d8f3a6dc4e0968b19f08047ac4fa68e7

Observation 659b81d6-7980-4d5f-8377-d2320644712b · inbound

CuriosAI Submission to the EgoExo4D Proficiency Estimation Challenge 2025 cites this paper.

CuriosAI Submission to the EgoExo4D Proficiency Estimation Challenge 2025 Sapiens: Foundation for Human Vision Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:16.473720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:18:16.473720Z digest=sha256:3c2866aa9f2794c7b0c5246e35615cca4aaa4cb5e65e2701878b1167980c7648

Observation 19b506b2-4eb5-49ae-927e-1a82223fd318 · inbound

EgoAnimate: Generating Human Animations from Egocentric top-down Views cites this paper.

EgoAnimate: Generating Human Animations from Egocentric top-down Views Sapiens: Foundation for Human Vision Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:05:52.105572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:05:52.105572Z digest=sha256:605bf1fc1a45ee1c5c41e52dd75513e22de1b627d22af352709a440a5b1272aa

Observation 6ef0743c-0968-4fe6-b29d-cadef456ac28 · inbound

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models cites this paper.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Sapiens: Foundation for Human Vision Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.528734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.528734Z digest=sha256:4a2e339f4906e2fd1ced3aea397d9e974dfc35ae250a2ad08682cf39e7298fa8

Observation b280cab7-c547-4bc9-a41c-462ccaa37c3d · inbound

Part Segmentation of Human Meshes via Multi-View Human Parsing cites this paper.

Part Segmentation of Human Meshes via Multi-View Human Parsing Sapiens: Foundation for Human Vision Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:03:12.936155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:03:12.936155Z digest=sha256:34531454d6562e5db5914213b590548478f207243ce2cba5704d6c98b3e4a5fb

Observation e306f04c-b886-48a6-9290-c1047cab6dcb · inbound

Delay-constrained re-entry governs large-scale brain seizures and other network pathologies cites this paper.

Delay-constrained re-entry governs large-scale brain seizures and other network pathologies Sapiens: Foundation for Human Vision Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T23:49:54.834936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:49:54.834936Z digest=sha256:e6d49cca5ed6bc1d84f153e2faf8733009a7904a1ebf4ea82c2280ec8bf96195

Observation 02209304-27e1-4d7f-8908-97dfcb40cae8 · inbound

Generative Video Matting cites this paper.

Generative Video Matting Sapiens: Foundation for Human Vision Models

Reference 2002

Resolution
unresolved
no resolver link, observed 2026-08-05T21:55:44.834958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:55:44.834958Z digest=sha256:72cd89be6e57b8f0aa48cc0d03666eeb3a5b71e6cce2056e459ff01179901a5e

Observation aefc38d3-8410-471e-ae5e-979f2c89fec7 · inbound

InfinityHuman: Towards Long-Term Audio-Driven Human cites this paper.

InfinityHuman: Towards Long-Term Audio-Driven Human Sapiens: Foundation for Human Vision Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.163455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.163455Z digest=sha256:814e9181fcd13cd925a3b0313745e5d4024b72f38b8b253c6cf96f3a6019805f

Observation fca705c6-4480-412b-a557-e8e826d19aac · inbound

GRMM: Real-Time High-Fidelity Gaussian Morphable Head Model with Learned Residuals cites this paper.

GRMM: Real-Time High-Fidelity Gaussian Morphable Head Model with Learned Residuals Sapiens: Foundation for Human Vision Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T11:54:57.300729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:54:57.300729Z digest=sha256:37f7aea335df9e54600bc8b7ca6e21a9123cc878ce56f1bb88fe03ee63577fbb

Observation 90b452a8-e535-40ca-be4c-ed19577865d5 · inbound

STROKEVISION-BENCH: A Multimodal Video And 2D Pose Benchmark For Tracking Stroke Recovery cites this paper.

STROKEVISION-BENCH: A Multimodal Video And 2D Pose Benchmark For Tracking Stroke Recovery Sapiens: Foundation for Human Vision Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T11:29:52.407904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:29:52.407904Z digest=sha256:3173ae8dda41b99464ceccec924bf051f42cb9ef71063007fa0cb61e8a3d5ee7

Observation 099940c9-e8a8-43ea-b9ca-db73581b73a9 · inbound

SAM 3: Segment Anything with Concepts cites this paper.

SAM 3: Segment Anything with Concepts Sapiens: Foundation for Human Vision Models

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T20:25:11.397091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T20:22:46.220021Z digest=sha256:9b608d912ed6321f48c746969c4014589fc42deba8cc97108a6136aed6616bde

Observation dfdef90a-a544-4d51-a081-eecaf0b00038 · inbound

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification cites this paper.

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification Sapiens: Foundation for Human Vision Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:31:22.021978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T23:30:14.969895Z digest=sha256:cff15e615ccec86807e954e2f53fede4090df4d15c408622dd30dcefefa9ded8

Observation e69ec915-a80e-4e02-9723-a121158ab61b · inbound

InverseDraping: Recovering Sewing Patterns from 3D Garment Surfaces via BoxMesh Bridging cites this paper.

InverseDraping: Recovering Sewing Patterns from 3D Garment Surfaces via BoxMesh Bridging Sapiens: Foundation for Human Vision Models

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:18:13.097908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T20:16:53.260771Z digest=sha256:e8fa1bcc73585f3c73b16fd707d42dd389892fc824372412d57592df8b9676eb

Observation 759c3397-c114-4a79-ad84-dbe731f2562a · inbound

HVG-3D: Bridging Real and Simulation Domains for 3D-Conditional Hand-Object Interaction Video Synthesis cites this paper.

HVG-3D: Bridging Real and Simulation Domains for 3D-Conditional Hand-Object Interaction Video Synthesis Sapiens: Foundation for Human Vision Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:18:30.120331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-14T00:14:36.537688Z digest=sha256:f57be3f09a904be9aec80e1b197698428fa98dda6dd8c84c8321368e8d0a8d5f

Observation 91286b49-0ae6-4e2d-a812-95909ae48b8c · inbound

GeoRelight: Learning Joint Geometrical Relighting and Reconstruction with Flexible Multi-Modal Diffusion Transformers cites this paper.

GeoRelight: Learning Joint Geometrical Relighting and Reconstruction with Flexible Multi-Modal Diffusion Transformers Sapiens: Foundation for Human Vision Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:36:10.152128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T01:19:12.431761Z digest=sha256:9cc67ca8aafb7a095aa7efcfc3385db673e59ae42c2ba1911cbcd9707b85427d

Observation 9b6e0ef0-3ed2-466e-9441-c7cdc8af5759 · inbound

MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction cites this paper.

MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction Sapiens: Foundation for Human Vision Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:28.624552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T22:06:58.480052Z digest=sha256:72e496a0eed634b3e5673f2d1bdd0fcc1ed8fa0b7a39c59743a321106a0154cc

Observation de71c350-fa05-417a-a583-34bfb38b363e · inbound

SignMAE: Segmentation-Driven Self-Supervised Learning for Sign Language Recognition cites this paper.

SignMAE: Segmentation-Driven Self-Supervised Learning for Sign Language Recognition Sapiens: Foundation for Human Vision Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:55:30.827729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T19:23:07.258575Z digest=sha256:fab126fdb32ec6dcbd26af8b08c69abdc24a50a2bafcc49193b9c76c282e8a9c

Observation 988a177d-061c-447c-aa48-b3054b09a0df · inbound

Agentic Pipeline for Self-Synchronized Multiview Joint Angle Monitoring in Uncalibrated Environments cites this paper.

Agentic Pipeline for Self-Synchronized Multiview Joint Angle Monitoring in Uncalibrated Environments Sapiens: Foundation for Human Vision Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:33:45.707145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T21:33:29.195617Z digest=sha256:26e63d690ba8c3bb55668a8b4bcc9653897fabef98d5ce4af133de1f9f45dce8

Observation eeb02796-ae3e-4c1c-aa43-2524eb78d273 · inbound

BodyReLux: Temporally Consistent Full-Body Video Relighting cites this paper.

BodyReLux: Temporally Consistent Full-Body Video Relighting Sapiens: Foundation for Human Vision Models

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:54:45.882987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-22T08:52:32.185764Z digest=sha256:c11105056671ee6abdb138acc31f965f3ef46c81a5c56a142b7868b77ac6f5de

Observation 1a233636-7f61-423c-922b-32cc3ff3f7a1 · inbound

Flex4DHuman: Flexible Multi-view Video Diffusion for 4D Human Reconstruction cites this paper.

Flex4DHuman: Flexible Multi-view Video Diffusion for 4D Human Reconstruction Sapiens: Foundation for Human Vision Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:48:32.982171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T06:54:11.233650Z digest=sha256:d0ee5234762a6d8a3d790c8f2cbe9846f70bdfa4a4a73837bad027b10aaf4d1f