Pith. sign in

Paper Citation Record · LEDGER

Towards Human-level Intelligence via Human-like Whole-Body Manipulation

As of 17 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 5 inbound Pith citation observations for arXiv:2507.17141.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.17141 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:01:48.302258Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T12:54:41.171118Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T12:08:06.368037Z

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2b0f7b73-1b62-43b5-98e6-5171f42d7b2a · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.267873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.267873Z digest=sha256:dd3c13e857851a80be9ecaa747362002b4b56f787d1daa801d973b2688606222

Observation 01bf83cd-8472-4c9e-9f74-2851dec3ac0d · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.274643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.274643Z digest=sha256:9e11d6a4b39365148c0291474399c19cd8a7529ac8fb46e3e71a4ee829ce58a0

Observation 5ee6dd17-23ce-4b32-a598-cfe00ffbc0e5 · outbound

This paper cites BEHAVIOR Robot Suite: Streamlining Real-World Whole-Body Manipulation for Everyday Household Activities.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation BEHAVIOR Robot Suite: Streamlining Real-World Whole-Body Manipulation for Everyday Household Activities

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.282301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.282301Z digest=sha256:fa2a44d3d5251c04054be9e5d62be6f234d02c94470e98e6dbd58bcd34de3c69

Observation 4e3da22d-4d4c-4f90-8c16-8dc0c50be2d6 · outbound

This paper cites an unresolved cited work.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:01:48.409435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T15:01:48.285012Z digest=sha256:963481d366d3d373a0f5fb0b54e2fd4961c80d68eb96c72f50d3629ffc989610

Observation 6c259e54-5c90-4886-8292-e2122eb90125 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.289297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.289297Z digest=sha256:399fb0e6b8b9ae2459a4a2d250085c0349d0561f6ea84b9b1a9b7be123f93553

Observation f8beccea-b818-4671-8b6e-a9f41a99bb7f · outbound

This paper cites Creating Multimodal Interactive Agents with Imitation and Self-Supervised Learning.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation Creating Multimodal Interactive Agents with Imitation and Self-Supervised Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.293869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.293869Z digest=sha256:bbf5163eda21a7a07917eb58cc5c65cd60c45254c57fd1fcfbd96b8ebc630aca

Observation c0c266d8-bd49-42e8-ba05-7a49719e50ac · outbound

This paper cites Gemini Robotics: Bringing AI into the Physical World.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation Gemini Robotics: Bringing AI into the Physical World

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.296876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.296876Z digest=sha256:ba3e916da7b3c826f24469fd1d55bdf759edada9642b516dfae26f114685b4c8

Observation 6fb58d83-d35d-4e5a-a845-c10e1f118060 · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.302258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.302258Z digest=sha256:5d6be6daa1173ff7c898c3dcdada6a62827c5504c4c72f7d8e2a23c838a55d01

Observation 568521f4-86e4-4c0c-9b98-971ebf4975f9 · outbound

This paper cites 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.299852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.299852Z digest=sha256:2b022e621998aa584a8d5d984191f31c33f30e1d7acaf332158adfba3297911a

Observation bb851c8b-7f14-437b-b6c4-201183b1f4ab · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.279619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.279619Z digest=sha256:590687a871699e00829fbc47cdc464f544cb427a143f3d3ee0a6b0e6ade8ef05

Observation 28e90636-98c5-430f-a027-361b01fd9cfe · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation DINOv2: Learning Robust Visual Features without Supervision

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.291422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.291422Z digest=sha256:5dd0160a8b9cde82bfc8ae64f5241eaf38e4609d47a874eb46f072c5996ad3e7

Observation 0ded6602-4027-4b6e-b618-c2da1cb3899d · outbound

This paper cites Learning with 3D rotations, a hitchhiker's guide to SO(3).

Towards Human-level Intelligence via Human-like Whole-Body Manipulation Learning with 3D rotations, a hitchhiker's guide to SO(3)

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.277154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.277154Z digest=sha256:9fba2e22bfbf04bf58a9822e1c6bdc1b3b9ff8a7d045fa0d315e2b1ecf71f6c6

Observation ade8f8f6-1d36-4790-9147-462b03494103 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.287098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.287098Z digest=sha256:ac9520380a37baa54d57f2a51f62ecbb288632c8b65545d748cd274f44d8f39b

Observation 08323709-7dd0-48ff-a5ae-b4252357730b · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.271563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.271563Z digest=sha256:4d723e59f417dc5bd4eca19c8839b0bef6a97212e650058be120a2ac393c1600

Pith citing papers

Observation 9409a926-7b0b-4fc2-ab04-3c52d1410668 · inbound

When to Trust Imagination: Adaptive Action Execution for World Action Models cites this paper.

When to Trust Imagination: Adaptive Action Execution for World Action Models Towards Human-level Intelligence via Human-like Whole-Body Manipulation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:01:31.228812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T01:22:49.557507Z digest=sha256:9d6af168daf00056ab6da4faf808050284bde501e0181c4f824bb29e73e58065

Observation a96f9eb6-5bc2-4696-b900-3453c0c8184b · inbound

StableVLA: Towards Robust Vision-Language-Action Models without Extra Data cites this paper.

StableVLA: Towards Robust Vision-Language-Action Models without Extra Data Towards Human-level Intelligence via Human-like Whole-Body Manipulation

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:58:15.142464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-20T11:55:22.498055Z digest=sha256:ef1f9d7d025f89fe8cd12970de7105d7e34404114adca5433dc3d9c0a18d2d49

Observation fc780535-d83b-4c15-9972-f5881b1384e6 · inbound

SAFE-Pruner: Semantic Attention-Guided Future-Aware Token Pruning for Efficient Vision-Language-Action Manipulation cites this paper.

SAFE-Pruner: Semantic Attention-Guided Future-Aware Token Pruning for Efficient Vision-Language-Action Manipulation Towards Human-level Intelligence via Human-like Whole-Body Manipulation

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T08:43:14.979187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-29T08:40:10.152344Z digest=sha256:a20b70bede89fe742ca2e7a3e29a92154db462d269d4dfa0c4ca7377b07e125c

Observation 3dc1fbf9-1497-4d4b-8955-db0ca6abbf64 · inbound

SAFE-Pruner: Semantic Attention-Guided Future-Aware Token Pruning for Efficient Vision-Language-Action Manipulation cites this paper.

SAFE-Pruner: Semantic Attention-Guided Future-Aware Token Pruning for Efficient Vision-Language-Action Manipulation Towards Human-level Intelligence via Human-like Whole-Body Manipulation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T12:54:41.171118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:54:41.171118Z digest=sha256:64bd5c7baed3e35d8fb793846ba5692995e5824245a266199e0acf4dc8da99ac

Observation 97405903-b3c0-4430-a10c-5bc2e7650981 · inbound

PhysMani: Physics-principled 3D World Model for Dynamic Object Manipulation cites this paper.

PhysMani: Physics-principled 3D World Model for Dynamic Object Manipulation Towards Human-level Intelligence via Human-like Whole-Body Manipulation

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T12:08:06.369422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-03T12:05:35.255381Z digest=sha256:64b72551ff5f3b8c995a1f2a6d491fda24ed2ed314dbbafd1b165ddf3343daa8