Pith. sign in

Paper Citation Record · LEDGER

Vision-based Manipulation from Single Human Video with Open-World Object Graphs

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2405.20321.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.20321 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:26:47.561876Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:27:40.028035Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 97ff7ff7-49a0-499a-ae5b-312c9601bd7e · inbound

DreamGen: Unlocking Generalization in Robot Learning through Video World Models cites this paper.

DreamGen: Unlocking Generalization in Robot Learning through Video World Models Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:50:45.596702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T23:50:45.332466Z digest=sha256:825122d3fd4393d4af06eb00a3e6825630b128fe421f1f1cca157e079a9aa5d8

Observation 6c0f72e4-1a9b-4b8a-8ae8-66ee50294210 · inbound

Object-Focus Actor for Data-efficient Robot Generalization Dexterous Manipulation cites this paper.

Object-Focus Actor for Data-efficient Robot Generalization Dexterous Manipulation Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:47.561876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:26:47.561876Z digest=sha256:976258d890a41fec725ff97cf9e7d21581930f1b7d5c27285f6b2c48c2d0a5ee

Observation 81baae65-98a1-4b7d-9c09-6ffcb5003fea · inbound

FLARE: Robot Learning with Implicit World Modeling cites this paper.

FLARE: Robot Learning with Implicit World Modeling Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:59:08.944575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T15:59:08.846629Z digest=sha256:f851b34c30c3b625bf03f19108b8e43131bc9b2b551bbb4e776b7b290475461d

Observation c0a0494e-c379-4aad-b550-59a29ff9c649 · inbound

MimicFunc: Imitating Tool Manipulation from a Single Human Video via Functional Correspondence cites this paper.

MimicFunc: Imitating Tool Manipulation from a Single Human Video via Functional Correspondence Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T19:04:21.942117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:04:21.942117Z digest=sha256:6e343199f01aa64d2fa4045d6742608a9fcc80080627e91d8114c80f6a278dbd

Observation 36dc9322-e1dd-407f-9e1c-463dc66d6069 · inbound

Weakly-Supervised Learning of Dense Functional Correspondences cites this paper.

Weakly-Supervised Learning of Dense Functional Correspondences Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-05T10:37:52.726480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:37:52.726480Z digest=sha256:109b06525088ea0f0af78820e578e73600c43b0ab06bc0ec592e5f588db8652f

Observation f3066f7f-e8a0-4f0c-8bc1-bb9f7a5c211f · inbound

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos cites this paper.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.108784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.108784Z digest=sha256:990a2d2ed5bf4c30278615fcd8a9c9008eafe02ace6863d7fc23dd717318f386

Observation 2b306d1c-84f1-48c2-841b-59cafae07164 · inbound

X-Diffusion: Training Diffusion Policies on Cross-Embodiment Human Demonstrations cites this paper.

X-Diffusion: Training Diffusion Policies on Cross-Embodiment Human Demonstrations Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:40:33.717233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T00:37:12.170711Z digest=sha256:e309134e67bd590da1cae997fda799c0adb71d6203a32bbed57c0a1d46d26adb

Observation fa048d89-1e3a-422d-9ffd-dfe3e2ee3b1e · inbound

Act, Sense, Act: Learning Active Perception from Large-Scale Egocentric Human Data cites this paper.

Act, Sense, Act: Learning Active Perception from Large-Scale Egocentric Human Data Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T04:37:22.123267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:37:22.123267Z digest=sha256:64cf2beebd69d945a8a85a3b0ebb72a2bd03a3d37b7f1158c80f89ddba08cb75

Observation f1af1576-434a-48ca-b6cc-5321830ea88b · inbound

WARPED: Wrist-Aligned Rendering for Robot Policy Learning from Egocentric Human Demonstrations cites this paper.

WARPED: Wrist-Aligned Rendering for Robot Policy Learning from Egocentric Human Demonstrations Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 132

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:01:02.638257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:14:24.932972Z digest=sha256:e71446c13e0f2008168311b2a4e1f46b721e5011684701d5e28e80c832beac2e

Observation ea2cfcd4-2e15-48f1-9660-6aa81e5fd7cf · inbound

One-Shot Cross-Geometry Skill Transfer through Part Decomposition cites this paper.

One-Shot Cross-Geometry Skill Transfer through Part Decomposition Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:44:38.293188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T10:30:24.436801Z digest=sha256:ea7dc7d739f64c75369d7a53a11d453834bfb3029e1d523807c0da8e2babdc3c

Observation 5c904740-9e62-4a7e-a43b-15029303366b · inbound

Bridging the Embodiment Gap: Disentangled Cross-Embodiment Video Editing cites this paper.

Bridging the Embodiment Gap: Disentangled Cross-Embodiment Video Editing Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T11:01:31.464839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T15:44:25.021995Z digest=sha256:b05253b82e41ef9f82f1d4312f1e46f56264e8b83171a6973d3c843c3d7a8e0c

Observation 9a6c31e0-ea67-4404-a1a0-6314c53d5b1e · inbound

MonoDuo: Using One Robot Arm to Learn Bimanual Policies cites this paper.

MonoDuo: Using One Robot Arm to Learn Bimanual Policies Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:23:12.511997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T07:22:17.679021Z digest=sha256:51ac4bd8a6555fb305f61cc9d967ec6fe708b7160350cfb225f689345ee3c6d0

Observation b5f70642-fd6f-4685-ac1f-01b729b0c0fb · inbound

Hand-centric Human-to-Robot Trajectory Transfer from Video Demonstrations via Open-World Contact Localization cites this paper.

Hand-centric Human-to-Robot Trajectory Transfer from Video Demonstrations via Open-World Contact Localization Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:27:40.029445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T13:15:57.832559Z digest=sha256:0f1e7ee3c9cb30dfed6807d853aa350bc814d38a266a458d4bf6527ad11318de

Observation 4a04b760-40a9-4a33-9eda-7df6de9f918c · inbound

ObjRetarget: An Object-Aware Motion Retargeting Framework with Anthropomorphic Arm Constraints and Polyhedral Hand Modeling cites this paper.

ObjRetarget: An Object-Aware Motion Retargeting Framework with Anthropomorphic Arm Constraints and Polyhedral Hand Modeling Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-11T23:40:27.616119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:40:27.616119Z digest=sha256:8117aaa418f07bfc920110117cb85920e801aa99545802c17f6e1135214fdb9b