Pith. sign in

Paper Citation Record · LEDGER

Sapiens: Foundation for Human Vision Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2408.12569.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.12569 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:16:16.496425Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:48:32.980001Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation caa24bc2-58d8-48b4-b717-e5cceb5bdbca · inbound

HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation cites this paper.

HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Sapiens: Foundation for Human Vision Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T21:16:16.496425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:16:16.496425Z digest=sha256:0c4426d2dedcc4c5175df706b2ea689f68eb4f42a760af91ea991b1b8c71e6df

Observation 4134d363-e3ab-498e-ad66-b05f272d8a05 · inbound

PoseBH: Prototypical Multi-Dataset Training Beyond Human Pose Estimation cites this paper.

PoseBH: Prototypical Multi-Dataset Training Beyond Human Pose Estimation Sapiens: Foundation for Human Vision Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:41.572907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:41.572907Z digest=sha256:dc641509391bcf0cf6a7b1f0745010f687ac88fa6dd5340bcbabe47231b8153d

Observation 14e38772-f43b-48b1-a47e-406ecd8c8adb · inbound

Total-Editing: Head Avatar with Editable Appearance, Motion, and Lighting cites this paper.

Total-Editing: Head Avatar with Editable Appearance, Motion, and Lighting Sapiens: Foundation for Human Vision Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:27.939782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:27.939782Z digest=sha256:efb4482d2c246c59f31b86f90b75478fc320a7d3ea1927cd41e61cc257089e0d

Observation 062eb3d0-d154-444c-8b9b-5e2e1887a938 · inbound

HuGeDiff: 3D Human Generation via Diffusion with Gaussian Splatting cites this paper.

HuGeDiff: 3D Human Generation via Diffusion with Gaussian Splatting Sapiens: Foundation for Human Vision Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T10:50:50.117287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:50:50.117287Z digest=sha256:987bf2b1170937a11eac738f24e90fae67c80d819d18602e06ff1c8aebd849de

Observation e7518576-ac00-44e3-aad1-8dfeedd98dfa · inbound

Controllable and Expressive One-Shot Video Head Swapping cites this paper.

Controllable and Expressive One-Shot Video Head Swapping Sapiens: Foundation for Human Vision Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:42:33.980220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:42:33.980220Z digest=sha256:f2b400d770eb93867b0bf8da9ca8e41569d8f843f9e821e56428b51e697bc622

Observation 816c7bcd-ba28-4ed8-8187-abb5b8b292a9 · inbound

DreamCube: 3D Panorama Generation via Multi-plane Synchronization cites this paper.

DreamCube: 3D Panorama Generation via Multi-plane Synchronization Sapiens: Foundation for Human Vision Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:05.494749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:35:05.494749Z digest=sha256:f30698b5cceaf7b5af2862b1b5ab0d798d76ef73583654b075c1c35b222c4e04

Observation df1e7e3f-843f-4dd4-87b8-45e8c909f101 · inbound

Video Virtual Try-on with Conditional Diffusion Transformer Inpainter cites this paper.

Video Virtual Try-on with Conditional Diffusion Transformer Inpainter Sapiens: Foundation for Human Vision Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T22:33:30.636955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:33:30.636955Z digest=sha256:1aab412afa68da036c9ce18dfe5165c94a9b308e906ea8ab9b79c0c73a085030

Observation 659b81d6-7980-4d5f-8377-d2320644712b · inbound

CuriosAI Submission to the EgoExo4D Proficiency Estimation Challenge 2025 cites this paper.

CuriosAI Submission to the EgoExo4D Proficiency Estimation Challenge 2025 Sapiens: Foundation for Human Vision Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:16.473720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:18:16.473720Z digest=sha256:4896fd81152e0c496beab090aa4789929d676e75c5439af6e10841a0bb390246

Observation 19b506b2-4eb5-49ae-927e-1a82223fd318 · inbound

EgoAnimate: Generating Human Animations from Egocentric top-down Views cites this paper.

EgoAnimate: Generating Human Animations from Egocentric top-down Views Sapiens: Foundation for Human Vision Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:05:52.105572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:05:52.105572Z digest=sha256:a055f3081dbc2f04e362fd3cd54fcbc12dd4aa957eeac525a19302ba7e78d6ac

Observation 6ef0743c-0968-4fe6-b29d-cadef456ac28 · inbound

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models cites this paper.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Sapiens: Foundation for Human Vision Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.528734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.528734Z digest=sha256:05cf280346208c4f4123395aa9612ed78f6aca2125bba809df6aba88d5402289

Observation b280cab7-c547-4bc9-a41c-462ccaa37c3d · inbound

Part Segmentation of Human Meshes via Multi-View Human Parsing cites this paper.

Part Segmentation of Human Meshes via Multi-View Human Parsing Sapiens: Foundation for Human Vision Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:03:12.936155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:03:12.936155Z digest=sha256:04c8a4fab2f5006ec113f45ff81293d8520479998292bdbc0c7a3ca9bc3e4e33

Observation e306f04c-b886-48a6-9290-c1047cab6dcb · inbound

Delay-constrained re-entry governs large-scale brain seizures and other network pathologies cites this paper.

Delay-constrained re-entry governs large-scale brain seizures and other network pathologies Sapiens: Foundation for Human Vision Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T23:49:54.834936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:49:54.834936Z digest=sha256:b0d8dc7e65b3a5dbc676260ac575a30d662d1ea3d113335fe047efcac3653d6e

Observation 02209304-27e1-4d7f-8908-97dfcb40cae8 · inbound

Generative Video Matting cites this paper.

Generative Video Matting Sapiens: Foundation for Human Vision Models

Reference 2002

Resolution
unresolved
no resolver link, observed 2026-08-05T21:55:44.834958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:55:44.834958Z digest=sha256:94b2fe90f35eb3109079a61f83c6104707c6f32dcd6c427170c35a60fb83fc31

Observation aefc38d3-8410-471e-ae5e-979f2c89fec7 · inbound

InfinityHuman: Towards Long-Term Audio-Driven Human cites this paper.

InfinityHuman: Towards Long-Term Audio-Driven Human Sapiens: Foundation for Human Vision Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.163455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.163455Z digest=sha256:f2175db4d20aab0172d2f817b599636d54ab45a13775ec276a71dd441a61541a

Observation fca705c6-4480-412b-a557-e8e826d19aac · inbound

GRMM: Real-Time High-Fidelity Gaussian Morphable Head Model with Learned Residuals cites this paper.

GRMM: Real-Time High-Fidelity Gaussian Morphable Head Model with Learned Residuals Sapiens: Foundation for Human Vision Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T11:54:57.300729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:54:57.300729Z digest=sha256:b418882a6c067cce357b51eafe43d2d56962d8606dc78ea74b5b06c394931716

Observation 90b452a8-e535-40ca-be4c-ed19577865d5 · inbound

STROKEVISION-BENCH: A Multimodal Video And 2D Pose Benchmark For Tracking Stroke Recovery cites this paper.

STROKEVISION-BENCH: A Multimodal Video And 2D Pose Benchmark For Tracking Stroke Recovery Sapiens: Foundation for Human Vision Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T11:29:52.407904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:29:52.407904Z digest=sha256:0558d3b93fa6e702afc649c61bf990247771a1bea3959860b192e58431db5ac0

Observation 099940c9-e8a8-43ea-b9ca-db73581b73a9 · inbound

SAM 3: Segment Anything with Concepts cites this paper.

SAM 3: Segment Anything with Concepts Sapiens: Foundation for Human Vision Models

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T20:25:11.397091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T20:22:46.220021Z digest=sha256:4ae1a07208a43b9d3e9ad9403e88bc79356b874436df7c9423abd754f27308fc

Observation dfdef90a-a544-4d51-a081-eecaf0b00038 · inbound

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification cites this paper.

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification Sapiens: Foundation for Human Vision Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:31:22.021978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T23:30:14.969895Z digest=sha256:ffa3c5612488489ecf9444d421a1edffecccd7e36fd45b93574c33f7c4a8bc2d

Observation e69ec915-a80e-4e02-9723-a121158ab61b · inbound

InverseDraping: Recovering Sewing Patterns from 3D Garment Surfaces via BoxMesh Bridging cites this paper.

InverseDraping: Recovering Sewing Patterns from 3D Garment Surfaces via BoxMesh Bridging Sapiens: Foundation for Human Vision Models

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:18:13.097908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T20:16:53.260771Z digest=sha256:26f0cd20f18e4cdb4be81645b6d8376e9a6f31f7fe4e18383add54e5fad46e05

Observation 759c3397-c114-4a79-ad84-dbe731f2562a · inbound

HVG-3D: Bridging Real and Simulation Domains for 3D-Conditional Hand-Object Interaction Video Synthesis cites this paper.

HVG-3D: Bridging Real and Simulation Domains for 3D-Conditional Hand-Object Interaction Video Synthesis Sapiens: Foundation for Human Vision Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:18:30.120331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T00:14:36.537688Z digest=sha256:727a7ab0856463738b96837eca16fed5f27f61fa54882822f8e783274e48afe2

Observation 91286b49-0ae6-4e2d-a812-95909ae48b8c · inbound

GeoRelight: Learning Joint Geometrical Relighting and Reconstruction with Flexible Multi-Modal Diffusion Transformers cites this paper.

GeoRelight: Learning Joint Geometrical Relighting and Reconstruction with Flexible Multi-Modal Diffusion Transformers Sapiens: Foundation for Human Vision Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:36:10.152128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T01:19:12.431761Z digest=sha256:c835e262b477055836b88a337535d024715ba69b57ea5f7d0f955a6994e06f9a

Observation 9b6e0ef0-3ed2-466e-9441-c7cdc8af5759 · inbound

MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction cites this paper.

MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction Sapiens: Foundation for Human Vision Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:28.624552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T22:06:58.480052Z digest=sha256:67ab83db3fcf177fdde2500fa62f096e8e3f7782cbb544b53e10e44aa7fba2c4

Observation de71c350-fa05-417a-a583-34bfb38b363e · inbound

SignMAE: Segmentation-Driven Self-Supervised Learning for Sign Language Recognition cites this paper.

SignMAE: Segmentation-Driven Self-Supervised Learning for Sign Language Recognition Sapiens: Foundation for Human Vision Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:55:30.827729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T19:23:07.258575Z digest=sha256:e64204237ef5c20f81fd202455c278d265f0e2fffb6f04ad013b8d3c1d47799e

Observation 988a177d-061c-447c-aa48-b3054b09a0df · inbound

Agentic Pipeline for Self-Synchronized Multiview Joint Angle Monitoring in Uncalibrated Environments cites this paper.

Agentic Pipeline for Self-Synchronized Multiview Joint Angle Monitoring in Uncalibrated Environments Sapiens: Foundation for Human Vision Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:33:45.707145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T21:33:29.195617Z digest=sha256:dbfb40b84d39e991cafa1c2192f20f5a8c1edd77b40914d5f85783856903c4d8

Observation eeb02796-ae3e-4c1c-aa43-2524eb78d273 · inbound

BodyReLux: Temporally Consistent Full-Body Video Relighting cites this paper.

BodyReLux: Temporally Consistent Full-Body Video Relighting Sapiens: Foundation for Human Vision Models

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:54:45.882987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-22T08:52:32.185764Z digest=sha256:3cd1a19cd6ccaa80f9a14a59797d5f7c9024f2a262f274156816c57799b813f2

Observation 1a233636-7f61-423c-922b-32cc3ff3f7a1 · inbound

Flex4DHuman: Flexible Multi-view Video Diffusion for 4D Human Reconstruction cites this paper.

Flex4DHuman: Flexible Multi-view Video Diffusion for 4D Human Reconstruction Sapiens: Foundation for Human Vision Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:48:32.982171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T06:54:11.233650Z digest=sha256:41a1357667f343698e0149d55c6b6c4e252423faaf66155c5e5c4645066eb108