Pith. sign in

Paper Citation Record · LEDGER

Towards Human-level Intelligence via Human-like Whole-Body Manipulation

As of 7 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 5 inbound Pith citation observations for arXiv:2507.17141.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.17141 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:01:48.302258Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T12:54:41.171118Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T12:08:06.368037Z

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2b0f7b73-1b62-43b5-98e6-5171f42d7b2a · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.267873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.267873Z digest=sha256:49da096dcc8ca510d0f1843f236648396224fd3ecd22117c1992bd47640d432e

Observation 01bf83cd-8472-4c9e-9f74-2851dec3ac0d · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.274643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.274643Z digest=sha256:4536cf9016a3545ecc0fa2f5865af572d138b0c4ab9365260effe30d6657d212

Observation 5ee6dd17-23ce-4b32-a598-cfe00ffbc0e5 · outbound

This paper cites BEHAVIOR Robot Suite: Streamlining Real-World Whole-Body Manipulation for Everyday Household Activities.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation BEHAVIOR Robot Suite: Streamlining Real-World Whole-Body Manipulation for Everyday Household Activities

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.282301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.282301Z digest=sha256:338f3b99a71ba91b0b2a5c2f81bbb4bb5b396608885d515eb3b26bda9306ad85

Observation 4e3da22d-4d4c-4f90-8c16-8dc0c50be2d6 · outbound

This paper cites an unresolved cited work.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:01:48.409435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:01:48.285012Z digest=sha256:2262c9bc2ca70e9b9797c817c510db89c48f48b06bf649d08b9c2e3eb09ef017

Observation 6c259e54-5c90-4886-8292-e2122eb90125 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.289297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.289297Z digest=sha256:22113a705469ebe3795f2fabc110623267b2009db0004324284495269aefd3cf

Observation f8beccea-b818-4671-8b6e-a9f41a99bb7f · outbound

This paper cites Creating Multimodal Interactive Agents with Imitation and Self-Supervised Learning.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation Creating Multimodal Interactive Agents with Imitation and Self-Supervised Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.293869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.293869Z digest=sha256:dadcaeb287f376650d1a192c01a1074e09473a9027ae90dc39da986fb02d4fcb

Observation c0c266d8-bd49-42e8-ba05-7a49719e50ac · outbound

This paper cites Gemini Robotics: Bringing AI into the Physical World.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation Gemini Robotics: Bringing AI into the Physical World

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.296876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.296876Z digest=sha256:1e546c6763f45a6da3cff2c61b7cda64ab89af8cfe2698cfa537cb54723cc4f0

Observation 6fb58d83-d35d-4e5a-a845-c10e1f118060 · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.302258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.302258Z digest=sha256:ff9a38b0c0c7f6b7f70be2be6895636bc8d75cdc77665bcfd9fe40014d647c3e

Observation 568521f4-86e4-4c0c-9b98-971ebf4975f9 · outbound

This paper cites 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.299852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.299852Z digest=sha256:d2ccc5e53a9ed9aab3723d584e78f8b766a8b44957ba9c09798c10757848ebd8

Observation bb851c8b-7f14-437b-b6c4-201183b1f4ab · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.279619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.279619Z digest=sha256:291920a58604f7ea4c481e19f0f926d77fb4e5c4d9f54db15f480f3c5d693187

Observation 28e90636-98c5-430f-a027-361b01fd9cfe · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation DINOv2: Learning Robust Visual Features without Supervision

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.291422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.291422Z digest=sha256:7094909ac528bb163a0dac9381f71217dcf2f2d0a24e2274f60a6ebf4c79a97f

Observation 0ded6602-4027-4b6e-b618-c2da1cb3899d · outbound

This paper cites Learning with 3D rotations, a hitchhiker's guide to SO(3).

Towards Human-level Intelligence via Human-like Whole-Body Manipulation Learning with 3D rotations, a hitchhiker's guide to SO(3)

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.277154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.277154Z digest=sha256:d4404c1c000fcfbf3ad77958e09402bd5598373224325bd379e06cdb601c33ca

Observation ade8f8f6-1d36-4790-9147-462b03494103 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.287098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.287098Z digest=sha256:a6e16ef105002d1cdbf7037ad4f01f1912bac3f5c2583e61ce515dd05a6d845d

Observation 08323709-7dd0-48ff-a5ae-b4252357730b · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Towards Human-level Intelligence via Human-like Whole-Body Manipulation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:48.271563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:48.271563Z digest=sha256:2cf08916c5d34f65b18b374fdb99575397b4097b58c7f9d0055123624a11095f

Pith citing papers

Observation 9409a926-7b0b-4fc2-ab04-3c52d1410668 · inbound

When to Trust Imagination: Adaptive Action Execution for World Action Models cites this paper.

When to Trust Imagination: Adaptive Action Execution for World Action Models Towards Human-level Intelligence via Human-like Whole-Body Manipulation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:01:31.228812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T01:22:49.557507Z digest=sha256:b24c6d3e26262fb190a052746c7dad7fcfe8bab76bc4a8ff6b33ea177eb887f2

Observation a96f9eb6-5bc2-4696-b900-3453c0c8184b · inbound

StableVLA: Towards Robust Vision-Language-Action Models without Extra Data cites this paper.

StableVLA: Towards Robust Vision-Language-Action Models without Extra Data Towards Human-level Intelligence via Human-like Whole-Body Manipulation

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:58:15.142464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T11:55:22.498055Z digest=sha256:cd0a956b1094fd1abcbbdec4c3056ff6fa9aad31ff49b17f2d1f4d96feb2b9ca

Observation fc780535-d83b-4c15-9972-f5881b1384e6 · inbound

SAFE-Pruner: Semantic Attention-Guided Future-Aware Token Pruning for Efficient Vision-Language-Action Manipulation cites this paper.

SAFE-Pruner: Semantic Attention-Guided Future-Aware Token Pruning for Efficient Vision-Language-Action Manipulation Towards Human-level Intelligence via Human-like Whole-Body Manipulation

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T08:43:14.979187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T08:40:10.152344Z digest=sha256:345b765d8f815d98dde7fec887e1afc46409e26c1c9a9e49ab1c796c6432da51

Observation 3dc1fbf9-1497-4d4b-8955-db0ca6abbf64 · inbound

SAFE-Pruner: Semantic Attention-Guided Future-Aware Token Pruning for Efficient Vision-Language-Action Manipulation cites this paper.

SAFE-Pruner: Semantic Attention-Guided Future-Aware Token Pruning for Efficient Vision-Language-Action Manipulation Towards Human-level Intelligence via Human-like Whole-Body Manipulation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T12:54:41.171118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:54:41.171118Z digest=sha256:a8af7e25e8a5d2a7fcc511888479372c36757dfc315deac6a913c89049c01d24

Observation 97405903-b3c0-4430-a10c-5bc2e7650981 · inbound

PhysMani: Physics-principled 3D World Model for Dynamic Object Manipulation cites this paper.

PhysMani: Physics-principled 3D World Model for Dynamic Object Manipulation Towards Human-level Intelligence via Human-like Whole-Body Manipulation

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T12:08:06.369422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T12:05:35.255381Z digest=sha256:4c887dd653d1b90b3a2a7c120f5e47c1ff8c742f386ed27e206b3e62c0d245af