Pith. sign in

Paper Citation Record · LEDGER

Vision-based Manipulation from Single Human Video with Open-World Object Graphs

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2405.20321.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.20321 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:26:47.561876Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:27:40.028035Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 97ff7ff7-49a0-499a-ae5b-312c9601bd7e · inbound

DreamGen: Unlocking Generalization in Robot Learning through Video World Models cites this paper.

DreamGen: Unlocking Generalization in Robot Learning through Video World Models Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:50:45.596702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T23:50:45.332466Z digest=sha256:f1b06c86c5006f27723c3a5a05549eeef33a0897a294a2baac0342cf36a5397a

Observation 6c0f72e4-1a9b-4b8a-8ae8-66ee50294210 · inbound

Object-Focus Actor for Data-efficient Robot Generalization Dexterous Manipulation cites this paper.

Object-Focus Actor for Data-efficient Robot Generalization Dexterous Manipulation Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:47.561876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:47.561876Z digest=sha256:81f6ad79bb64005e70006fef88716aade4526895f05925c62cb17f27e1d830ef

Observation 81baae65-98a1-4b7d-9c09-6ffcb5003fea · inbound

FLARE: Robot Learning with Implicit World Modeling cites this paper.

FLARE: Robot Learning with Implicit World Modeling Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:59:08.944575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T15:59:08.846629Z digest=sha256:cdf741964253e3799cdd52363095d0b087e84ebb0647ac55b3883cfb92fe7516

Observation c0a0494e-c379-4aad-b550-59a29ff9c649 · inbound

MimicFunc: Imitating Tool Manipulation from a Single Human Video via Functional Correspondence cites this paper.

MimicFunc: Imitating Tool Manipulation from a Single Human Video via Functional Correspondence Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T19:04:21.942117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:04:21.942117Z digest=sha256:6e343199f01aa64d2fa4045d6742608a9fcc80080627e91d8114c80f6a278dbd

Observation 36dc9322-e1dd-407f-9e1c-463dc66d6069 · inbound

Weakly-Supervised Learning of Dense Functional Correspondences cites this paper.

Weakly-Supervised Learning of Dense Functional Correspondences Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-05T10:37:52.726480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:37:52.726480Z digest=sha256:109b06525088ea0f0af78820e578e73600c43b0ab06bc0ec592e5f588db8652f

Observation f3066f7f-e8a0-4f0c-8bc1-bb9f7a5c211f · inbound

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos cites this paper.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.108784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.108784Z digest=sha256:990a2d2ed5bf4c30278615fcd8a9c9008eafe02ace6863d7fc23dd717318f386

Observation 2b306d1c-84f1-48c2-841b-59cafae07164 · inbound

X-Diffusion: Training Diffusion Policies on Cross-Embodiment Human Demonstrations cites this paper.

X-Diffusion: Training Diffusion Policies on Cross-Embodiment Human Demonstrations Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:40:33.717233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T00:37:12.170711Z digest=sha256:78520d869fb91cd29b5cb3bfa3cd3723718f1bbad4b2fa3232e350069cb24eea

Observation fa048d89-1e3a-422d-9ffd-dfe3e2ee3b1e · inbound

Act, Sense, Act: Learning Active Perception from Large-Scale Egocentric Human Data cites this paper.

Act, Sense, Act: Learning Active Perception from Large-Scale Egocentric Human Data Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T04:37:22.123267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:37:22.123267Z digest=sha256:64cf2beebd69d945a8a85a3b0ebb72a2bd03a3d37b7f1158c80f89ddba08cb75

Observation f1af1576-434a-48ca-b6cc-5321830ea88b · inbound

WARPED: Wrist-Aligned Rendering for Robot Policy Learning from Egocentric Human Demonstrations cites this paper.

WARPED: Wrist-Aligned Rendering for Robot Policy Learning from Egocentric Human Demonstrations Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 132

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:01:02.638257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:14:24.932972Z digest=sha256:ba2f3cd4a3848fcc27ab20372d80d2444443eb521a3e2072715f6bb62dee0cc8

Observation ea2cfcd4-2e15-48f1-9660-6aa81e5fd7cf · inbound

One-Shot Cross-Geometry Skill Transfer through Part Decomposition cites this paper.

One-Shot Cross-Geometry Skill Transfer through Part Decomposition Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:44:38.293188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T10:30:24.436801Z digest=sha256:842bcb81e641c08a32acd3f012cc7e44e0826d09ff588de7e11be924b40f6764

Observation 5c904740-9e62-4a7e-a43b-15029303366b · inbound

Bridging the Embodiment Gap: Disentangled Cross-Embodiment Video Editing cites this paper.

Bridging the Embodiment Gap: Disentangled Cross-Embodiment Video Editing Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T11:01:31.464839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T15:44:25.021995Z digest=sha256:d4c58b0819cd838ad2d58f430a867ce065095b32760a3317271baa39ca5e4f0b

Observation 9a6c31e0-ea67-4404-a1a0-6314c53d5b1e · inbound

MonoDuo: Using One Robot Arm to Learn Bimanual Policies cites this paper.

MonoDuo: Using One Robot Arm to Learn Bimanual Policies Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:23:12.511997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T07:22:17.679021Z digest=sha256:2675c18608ca6872e66b2136698c2382c21c763fc7c97b74d492998f46ee4094

Observation b5f70642-fd6f-4685-ac1f-01b729b0c0fb · inbound

Hand-centric Human-to-Robot Trajectory Transfer from Video Demonstrations via Open-World Contact Localization cites this paper.

Hand-centric Human-to-Robot Trajectory Transfer from Video Demonstrations via Open-World Contact Localization Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:27:40.029445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T13:15:57.832559Z digest=sha256:07aaf5897b4b73786def3392d4fa25ac12bd453b373b34546872da3f10365798

Observation 4a04b760-40a9-4a33-9eda-7df6de9f918c · inbound

ObjRetarget: An Object-Aware Motion Retargeting Framework with Anthropomorphic Arm Constraints and Polyhedral Hand Modeling cites this paper.

ObjRetarget: An Object-Aware Motion Retargeting Framework with Anthropomorphic Arm Constraints and Polyhedral Hand Modeling Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-11T23:40:27.616119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:40:27.616119Z digest=sha256:8117aaa418f07bfc920110117cb85920e801aa99545802c17f6e1135214fdb9b