Pith. sign in

Paper Citation Record · LEDGER

Robots Acquire Manipulation Skills in Seconds from a Single Human Video

As of 8 August 2026, this Paper Citation Record lists 100 of 117 outbound references and 0 inbound Pith citation observations for arXiv:2607.20033.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.20033 v3

Coverage vector

measured 100 of 117 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:06:14.622405Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 117 outbound references displayed

  • verified exact4
  • verified fuzzy0
  • unresolved95
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 94359ae8-8180-4b6a-9a92-35b8e76da95e · outbound

This paper cites Learning agile and dynamic motor skills for legged robots.Science Robotics, 4(26):eaau5872, 2019.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning agile and dynamic motor skills for legged robots.Science Robotics, 4(26):eaau5872, 2019

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.322238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.322238Z digest=sha256:ffd678156bb5b687bbb5904d17ad0314b2c820fa3fe22d45642334df2688c98a

Observation 31383d2c-fb02-4955-be36-b420bf8072af · outbound

This paper cites Learning quadrupedal locomotion over challenging terrain.Science Robotics, 5(47):eabc5986, 2020.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning quadrupedal locomotion over challenging terrain.Science Robotics, 5(47):eabc5986, 2020

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.326173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.326173Z digest=sha256:db13c44d2eaa3701907a0c5ea4af3183fe0a3420625dfd514db840d8d38bc632

Observation e85ff0d9-0444-46cc-b40e-76cfccf17004 · outbound

This paper cites Visual dexterity: In-hand reorientation of novel and complex object shapes.Science Robotics, 8(84):eadc9244, 2023.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Visual dexterity: In-hand reorientation of novel and complex object shapes.Science Robotics, 8(84):eadc9244, 2023

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.329059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.329059Z digest=sha256:0294875c8e4f5dcdd809c90bb16b8d4d27d76de40fba9c803b47e5c4e50656a6

Observation d91eb503-8646-4787-ab3a-267304942f20 · outbound

This paper cites Solving Rubik's Cube with a Robot Hand.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Solving Rubik's Cube with a Robot Hand

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.332530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.332530Z digest=sha256:46db40d1bf5e860d1dd235da99a719a6c48d729b024476ba05b8a908124f7b0e

Observation b2be2375-b5ec-447b-86b5-13d25e217a23 · outbound

This paper cites QT-Opt: Scalable deep reinforcement learning for vision-based robotic manipulation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video QT-Opt: Scalable deep reinforcement learning for vision-based robotic manipulation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.336354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.336354Z digest=sha256:dff26fb0d9f3ac70b8e21b56871bbd5b999fc1c9e598a24a509430eb3344f360

Observation 7a5d49b2-b912-43d9-ba97-5407b2ad3758 · outbound

This paper cites Scaling up multi-task robotic reinforcement learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Scaling up multi-task robotic reinforcement learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.339512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.339512Z digest=sha256:c952861a39e4f5dd964f0834664f7d4ad8e2b0771c3a74831858cc7420eca3d3

Observation 2db30b1a-54e8-4262-91f8-93d5771ac8c7 · outbound

This paper cites A generalized path integral control approach to reinforce- ment learning.Journal of Machine Learning Research, 11:3137–3181, 2010.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video A generalized path integral control approach to reinforce- ment learning.Journal of Machine Learning Research, 11:3137–3181, 2010

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.345777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.345777Z digest=sha256:6f63c445f50f553bd233117fb64bb368d7997bd44eaf3251b311f7d9f4aa542b

Observation 56be6551-0e6d-4b8e-ac45-478c1fa5d384 · outbound

This paper cites Path integral guided policy search.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Path integral guided policy search

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.349219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.349219Z digest=sha256:7dbc54c92a3231b63d2459e4568bfa2b0c3a8d8133af5ba65e07133456c2754f

Observation 42f478a6-c7ee-4ffb-bcaa-f124f707500e · outbound

This paper cites Learning coordinated badminton skills for legged manipulators.Science Robotics, 10(102):eadu3922, 2025.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning coordinated badminton skills for legged manipulators.Science Robotics, 10(102):eadu3922, 2025

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.355179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.355179Z digest=sha256:8c4148c083674f715bb16358ce272f246255c30f9c203c032608d2213f3b3be2

Observation ce5ed6d5-3969-4b24-8c5d-287a8c5c444f · outbound

This paper cites Precise and dexterous robotic manipulation via human-in- the-loop reinforcement learning.Science Robotics, 10(105):eads5033, 2025.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Precise and dexterous robotic manipulation via human-in- the-loop reinforcement learning.Science Robotics, 10(105):eads5033, 2025

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.357852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.357852Z digest=sha256:35eb09dfab0d0df796f65f189cf2c9d10f0cef985ff6ddbd700bbf9faf2fee9a

Observation 58750782-11ed-4e7e-a708-c6d01e743cf2 · outbound

This paper cites Barreiros, Aykut Özgün Önol, Mengchao Zhang, Sam Creasey, Aimee Goncalves, Andrew Beaulieu, Aditya Bhat, Kate M.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Barreiros, Aykut Özgün Önol, Mengchao Zhang, Sam Creasey, Aimee Goncalves, Andrew Beaulieu, Aditya Bhat, Kate M

Reference 11

Resolution
verified exact
doi, observed 2026-08-01T11:08:35.489516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-01T11:06:14.360554Z digest=sha256:e43e3b36768ecb89ef6a352e2e0a2772ba70bfb61cb015db1074f6404a50b91b

Observation ef832a9f-1fad-490c-8c2a-103f579e6ba3 · outbound

This paper cites Is imitation learning the route to humanoid robots?Trends in Cognitive Sciences, 3(6):233–242, 1999.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Is imitation learning the route to humanoid robots?Trends in Cognitive Sciences, 3(6):233–242, 1999

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.363086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.363086Z digest=sha256:eaa11e9f9a356a4b3cbd584a2c9cf2ce4812db10a0c8271ac1c3c11dfb82b81d

Observation 81d501af-ed7c-425d-b38c-aba62af36d4c · outbound

This paper cites Sanketi, Grecia Salazar, Michael S.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Sanketi, Grecia Salazar, Michael S

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.365672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.365672Z digest=sha256:e2c92080964056a3e215ec3f914d391fab23c089a5b838c29f8a0d2b89beeecf

Observation e47ab465-17f2-4f00-96d1-ef4b9c8bc63f · outbound

This paper cites In Proceedings of Robotics: Science and Systems, Los Angeles, CA, USA, 6 2025.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video In Proceedings of Robotics: Science and Systems, Los Angeles, CA, USA, 6 2025

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.368337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.368337Z digest=sha256:ebbc8cda991c5d1be34e3048be4a1f7955584a5bb1481338cbf05f4bd7a147a5

Observation eb283036-c049-4964-b2ed-01884736ce38 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.370745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.370745Z digest=sha256:ae0c6894e72a6340f3c8d3c7af05c5f56f7cb8b9fa0f3182205d18689b7710d0

Observation 1f836854-ada9-4d92-896d-664d4c65a5da · outbound

This paper cites Toward next-generation learned robot manipulation.Science Robotics, 6(54):eabd9461,.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Toward next-generation learned robot manipulation.Science Robotics, 6(54):eabd9461,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.373831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.373831Z digest=sha256:59d1a4ede013cb7b28fc79354621d9f13f42e818d1b33c7635d87624d4192df4

Observation 7446c1bd-1e0e-4fed-b958-4fa8a336a585 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Diffusion policy: Visuomotor policy learning via action diffusion

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.379606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.379606Z digest=sha256:8710445eaceeadd249970ffe67ad9d2930ffa7d48aa0ede538b938b0494ddb32

Observation 876d813f-f882-4067-886d-d3ac3bec386f · outbound

This paper cites Learning fine-grained bimanual manipulation with low-cost hardware.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning fine-grained bimanual manipulation with low-cost hardware

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.382048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.382048Z digest=sha256:b2d40c7e92f2222420cb9e191236807c2484de95b6f3661bae4bf96b44826dee

Observation d1fcdd1e-5b8e-408f-9026-32a24c71b420 · outbound

This paper cites RT-1: Robotics transformer for real-world control at scale.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video RT-1: Robotics transformer for real-world control at scale

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.385066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.385066Z digest=sha256:211b00365bf03fa792d2ce798dfe66f1a71f201274859db7fcc0ed026ec357d4

Observation 8bb048f5-b077-4cc1-b109-24e8d470aeab · outbound

This paper cites Open X-Embodiment: Robotic learning datasets and RT-X models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Open X-Embodiment: Robotic learning datasets and RT-X models

Reference 20

Resolution
malformed identifier
no resolver link, observed 2026-08-01T11:06:14.388416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.388416Z digest=sha256:c89897e872b5b136c8be3eaf3929bdd0bf1db68bc97d642dd3f3f7f819022670

Observation 784b7789-1279-423c-b45d-60a55e13a68d · outbound

This paper cites Octo: An open-source generalist robot policy.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Octo: An open-source generalist robot policy

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.391072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.391072Z digest=sha256:f24e53808e2ce08355d3a7d7bc9912d6f927ce9a9fee59be84b8d5ac7627e3cb

Observation 7274a360-5f3c-4098-acd1-5d1578617171 · outbound

This paper cites From Intention to Execution: Probing the Generalization Boundaries of Vision-Language-Action Models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video From Intention to Execution: Probing the Generalization Boundaries of Vision-Language-Action Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.393672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.393672Z digest=sha256:76fe0a918cf24d602e64d81c685bc0aa183671cc67558fb6101e95136128d73d

Observation ba9bc7c0-aae1-4b8a-8edf-65c29bb08870 · outbound

This paper cites Ricl: Adding in-context adaptability to pre-trained vision-language-action models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Ricl: Adding in-context adaptability to pre-trained vision-language-action models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.396754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.396754Z digest=sha256:b713f1a57d1ae76d5c370598554e745c97aef32cb759a511024fd189fa9fc00c

Observation 408baf13-066f-49c7-b04e-bd321caf8875 · outbound

This paper cites FLaRe: Achieving masterful and adaptive robot policies with large-scale reinforcement learning fine-tuning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video FLaRe: Achieving masterful and adaptive robot policies with large-scale reinforcement learning fine-tuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.399270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.399270Z digest=sha256:5e00e37483745abd2a184057bbef5e08ef203f7386601d3854d5990acbd40015

Observation a1afcf68-1be3-4145-896d-1bdd1d0be6e6 · outbound

This paper cites an unresolved cited work.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.402247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.402247Z digest=sha256:c57cbaf9e60a641b2a1b985bdc63bc4516ac0e7586e1023c2dd48695bd1d252b

Observation be24fb01-4931-4523-a20d-81b6d89f6d2b · outbound

This paper cites Learning by watching: Extracting reusable task knowledge from visual observation of human performance.IEEE Transactions on Robotics and Automation, 10(6):799–822, 1994.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning by watching: Extracting reusable task knowledge from visual observation of human performance.IEEE Transactions on Robotics and Automation, 10(6):799–822, 1994

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.404884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.404884Z digest=sha256:8647a52624d68799a5c56a6e567ea1d2d10f573fb82bd0a4dbe29833197d57f4

Observation 73f553e2-3f32-4fca-ab95-9133b8f05908 · outbound

This paper cites One-shot visual imitation learning via meta-learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video One-shot visual imitation learning via meta-learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.407354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.407354Z digest=sha256:1813d486edb1dfb2b46011b89f65db11ea604ac6d41a19cb911f13dcd9cab5ad

Observation 69b31d01-597f-455d-b3dc-43433d1e98eb · outbound

This paper cites One-shot imitation from observing humans via domain-adaptive meta-learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video One-shot imitation from observing humans via domain-adaptive meta-learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.409817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.409817Z digest=sha256:c940eeac9a799889ea0ce0ea68e9c94ed14f3e8ec8099f1a7f35c30c54ef5f08

Observation 6b99bef7-0c7f-46e8-9acc-c37fe7e00dca · outbound

This paper cites Towards more generalizable one-shot visual imitation learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Towards more generalizable one-shot visual imitation learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.412233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.412233Z digest=sha256:9a9667a2f965aeb5dd984b10f564dbb84fbcd181dd2e6a81d9ec7389b18d6427

Observation 10a83cc3-ad8f-45c6-8c6f-30a7cfeb88dd · outbound

This paper cites Vid2Robot: 23 End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Vid2Robot: 23 End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers

Reference 30

Resolution
verified exact
doi, observed 2026-08-01T11:08:35.235070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-01T11:06:14.414848Z digest=sha256:a891fc42edb137a5d03e233481d12dfe182b236a0bbf3dc595e7831c9644e4d3

Observation 7cc6af32-645c-4526-8f3f-64c6eddb73f8 · outbound

This paper cites Osvi-wm: One-shot visual imitation for unseen tasks using world-model-guided trajectory generation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Osvi-wm: One-shot visual imitation for unseen tasks using world-model-guided trajectory generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.417781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.417781Z digest=sha256:79cd7403f632f0db51562f81affc1220ad0272ffe0a5ddbf82b086068461b84b

Observation b2f1a0b4-c5e9-4fe8-aa35-0436f800d3ae · outbound

This paper cites Remembering the past to imagine the future: the prospective brain.Nature Reviews Neuroscience, 8(9):657–661, 2007.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Remembering the past to imagine the future: the prospective brain.Nature Reviews Neuroscience, 8(9):657–661, 2007

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.420468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.420468Z digest=sha256:fdfef18e46ec506b64ae582bb26824c84d3348b736c41d0e92c03b772ddcdd63

Observation d8b1e285-b93f-47a8-9890-c6ea725e75df · outbound

This paper cites Infant imitation after a 1-week delay: long-term memory for novel acts and multiple stimuli.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Infant imitation after a 1-week delay: long-term memory for novel acts and multiple stimuli

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.422854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.422854Z digest=sha256:95b3ad05749ca8aab63adfeda383cd1b79e143cc99f87c0e0f0c258c29917d61

Observation f9c97278-27f3-4a00-877d-01141e6949c0 · outbound

This paper cites Reinforcement learning and episodic memory in humans and animals: an integrative framework.Annual Review of Psychology, 68:101–128, 2017.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Reinforcement learning and episodic memory in humans and animals: an integrative framework.Annual Review of Psychology, 68:101–128, 2017

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.425807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.425807Z digest=sha256:a893b737efa9b1562b59857d4036e392c29b5f415a1a3fd2c722a4656aebed15

Observation 38020a35-ef05-4b6a-9f53-03a01ce2e132 · outbound

This paper cites Neural simulation of action: a unifying mechanism for motor cognition.NeuroImage, 14(1): S103–S109, 2001.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Neural simulation of action: a unifying mechanism for motor cognition.NeuroImage, 14(1): S103–S109, 2001

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.428457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.428457Z digest=sha256:1e654d155305e93c13a03dae563588f118b32d548f160561154a406cc7c5fcec

Observation 8af4d18e-5d39-435e-b4fd-4e21921b9af2 · outbound

This paper cites A unifying computational framework for motor control and social interaction.Philosophical Transactions of the Royal Society B: Biological Sciences, 358(1431):593–602, 2003.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video A unifying computational framework for motor control and social interaction.Philosophical Transactions of the Royal Society B: Biological Sciences, 358(1431):593–602, 2003

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.431672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.431672Z digest=sha256:1a1dbda80e58d9d6e2e9416659b92882e6e0bc0af3d2330d2c7b9442ad53c117

Observation 0fca39c9-e272-4c56-85f8-d64780b8154a · outbound

This paper cites Neurophysiological mechanisms underlying the understanding and imitation of action.Nature Reviews Neuroscience, 2(9):661–670, 2001.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Neurophysiological mechanisms underlying the understanding and imitation of action.Nature Reviews Neuroscience, 2(9):661–670, 2001

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.434086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.434086Z digest=sha256:aba70e920fe01bd4561604fd5f0267539e98375c4eb9247e5905ba3bd76b3166

Observation be67857c-2a1e-49bd-808a-935bd98a4d1a · outbound

This paper cites Neural circuits underlying imitation learning of hand actions: an event-related fmri study.Neuron, 42(2):323–334,.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Neural circuits underlying imitation learning of hand actions: an event-related fmri study.Neuron, 42(2):323–334,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.436692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.436692Z digest=sha256:59ccb2a59e54cb57e95fddcef2e75215b02326ccfd3db7814fa7fef7e2d43ff0

Observation 12435650-b02d-4012-bee6-23c93cb7b850 · outbound

This paper cites One-shot visual imitation via attributed waypoints and demonstration augmentation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video One-shot visual imitation via attributed waypoints and demonstration augmentation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.443751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.443751Z digest=sha256:edf7230e127e509891a20aa368ab2546cfbcfd8db6479f5f42e881101974803a

Observation 027ea5b4-0fbb-4fbd-bfc5-ced8b2edf762 · outbound

This paper cites Igniting vlms toward the embodied space.arXiv preprint arXiv:2509.11766, 2025.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Igniting vlms toward the embodied space.arXiv preprint arXiv:2509.11766, 2025

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.446310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.446310Z digest=sha256:3d898cc5cfe1e3293041eb904a444a6d7a25472174012539044bc05d384b1005

Observation ffa0874f-a180-4039-a09f-10a890c3a654 · outbound

This paper cites Fast-WAM: Do World Action Models Need Test-time Future Imagination?.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Fast-WAM: Do World Action Models Need Test-time Future Imagination?

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.449461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.449461Z digest=sha256:1b3a230ec4314e2bc09c88f90d8e46eeb6972bfd184fb097f1abc447421a8cc1

Observation 51a79043-76a0-40aa-87ae-8ca72bf81b8e · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Lora: Low-rank adaptation of large language models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.453040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.453040Z digest=sha256:0d990b7197ef1c30e3da16342ca05c1b0ffa933dec5ad14472078f56b5655a58

Observation 18a546c5-4863-4443-b270-8509f0e6d75c · outbound

This paper cites Tenenbaum, and Alberto Rodriguez.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Tenenbaum, and Alberto Rodriguez

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.456412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.456412Z digest=sha256:4375b3bcb801ef6e42a4651a8dcf1574e2202be3e7571b0e7e4540ee17978157

Observation 28a84e25-d66a-41f3-9f52-84fc45183b83 · outbound

This paper cites Representation learning via global temporal alignment and cycle-consistency.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Representation learning via global temporal alignment and cycle-consistency

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.461680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.461680Z digest=sha256:51a08700cee411ebc244c6937f3206e171a0a9c326b399218736a7db8d33fa47

Observation 5d8f974d-26c1-411b-884b-e7c22281e876 · outbound

This paper cites Temporal cycle- consistency learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Temporal cycle- consistency learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.464919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.464919Z digest=sha256:1402c4763a531eded3192151990457b2015b88529ddfb2730e308a022fcb3aa0

Observation a9f95ce4-39e2-4777-9c47-c9c98dc35ef3 · outbound

This paper cites Qwen3-VL Technical Report.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Qwen3-VL Technical Report

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.467761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.467761Z digest=sha256:c583119e3e4d5a649d78dc8530c276317cc06008c88f2c33c2e3503d10dcf4e0

Observation 1040c184-6f17-47fb-8826-4d2ca6522d32 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Wan: Open and Advanced Large-Scale Video Generative Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.471220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.471220Z digest=sha256:998f4996f2184113c3b90c9d610653983388cccb199a9cb46913d4d7d77f0746

Observation 0a175204-78dc-43cf-b59f-16cfa9c37749 · outbound

This paper cites Mixture-of-transformers: A sparse and scalable architecture for multi-modal foundation models.Transactions on Machine Learning Research, 2025.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Mixture-of-transformers: A sparse and scalable architecture for multi-modal foundation models.Transactions on Machine Learning Research, 2025

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.474142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.474142Z digest=sha256:1890b5539222f3d4cda6249dd7d2b9cc807c013322ec09773e7dfd3ce105e0dc

Observation b9ccac01-c547-4ab2-b503-b9602fdd504d · outbound

This paper cites Unimax: Fairer and more effective language sampling for large-scale multilingual pretraining.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Unimax: Fairer and more effective language sampling for large-scale multilingual pretraining

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.476838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.476838Z digest=sha256:ed5b99b5eadd118bbf6d5247c221913b8f1b1dd0445c54976cfe4aee04cedde1

Observation b0eb4d62-8857-46f8-809c-9028f505536d · outbound

This paper cites Flow matching for generative modeling.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Flow matching for generative modeling

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.479147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.479147Z digest=sha256:442a6cbec0cc099af20b073c336e332a5ec542310ca2fcbb7eeec90c1cbba129

Observation 700b5210-90ee-4047-854c-bb9053c7b8b1 · outbound

This paper cites A survey of robot learning from demonstra- tion.Robotics and Autonomous Systems, 57(5):469–483, 2009.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video A survey of robot learning from demonstra- tion.Robotics and Autonomous Systems, 57(5):469–483, 2009

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.481539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.481539Z digest=sha256:b8d72728e33cf6fd4a4057132f72f5efa995e23a78608dbbdf4a57577c6dc57e

Observation df315fcf-07ea-4aab-91b6-32ec55bb714e · outbound

This paper cites Recent advances in robot learning from demonstration.Annual Review of Control, Robotics, and Autonomous Systems, 3:297–330, 2020.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Recent advances in robot learning from demonstration.Annual Review of Control, Robotics, and Autonomous Systems, 3:297–330, 2020

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.484658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.484658Z digest=sha256:006391f5cc94a318d71890456cc8d6dfcbbb6ee3b5306841a65e77aba94a8108

Observation 74048569-48a6-4c74-b4d4-7545e985a833 · outbound

This paper cites Deep imitation learning for complex manipulation tasks from virtual reality teleoperation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Deep imitation learning for complex manipulation tasks from virtual reality teleoperation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.487448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.487448Z digest=sha256:d880ac1e78cfcace55464a85698b230080f45ad63bd790d3a5cbdb05e6f01604

Observation d0b8816f-c6c3-4570-b9ea-8996e34313de · outbound

This paper cites White, De Ru Tsai, Richard Jaepyeong Cha, Jeffrey Jopling, Chelsea Finn, and Axel Krieger.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video White, De Ru Tsai, Richard Jaepyeong Cha, Jeffrey Jopling, Chelsea Finn, and Axel Krieger

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.490077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.490077Z digest=sha256:2291efab6117930f0c3671ae28fefa976aeb1ed1c5c1cee6e3e3201f3d43c668

Observation 6b61211b-101d-4eea-8335-7ab704e100ff · outbound

This paper cites Pomerleau.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Pomerleau

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.492719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.492719Z digest=sha256:c48aa59656408f889bf107e62dbe80426a7cd83dd6a5668f9f69c63c38a1d70f

Observation afa1f508-d928-4466-b0f2-d76c2faf8b7d · outbound

This paper cites DexMV: Imitation learning for dexterous manipulation from human videos.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video DexMV: Imitation learning for dexterous manipulation from human videos

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.495603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.495603Z digest=sha256:7677eb8b5c22e37585878390e221e9fb27c14a22a2f1a90c64fcc7828da64a8c

Observation c7f7a8c9-ec0e-48c9-86bc-532764a32fc2 · outbound

This paper cites DexCap: Scalable and portable mocap data collection system for dexterous manipulation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video DexCap: Scalable and portable mocap data collection system for dexterous manipulation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.498568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.498568Z digest=sha256:de1a400b8e4b58ce608bbe05009e7481ed527fe70c877c7b9134201890b300ef

Observation afb8bc1c-e540-4cb2-94c9-7951c2fd3680 · outbound

This paper cites Time- contrastive networks: Self-supervised learning from video.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Time- contrastive networks: Self-supervised learning from video

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.501559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.501559Z digest=sha256:cef7be8fb12103165080455ede4c3c74d7409d8a65975860d6c4f34f93d3822c

Observation d4c9585e-5891-4164-8631-1553896065e1 · outbound

This paper cites XIRL: Cross-embodiment inverse reinforcement learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video XIRL: Cross-embodiment inverse reinforcement learning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.504183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.504183Z digest=sha256:576ae611ce690c02fb5d43587222ba8bcb0a0c98c18cb6412bea63f4a516feec

Observation 1e33ef4b-0629-4aa6-85b2-fcd1cd9099d5 · outbound

This paper cites Learning latent plans from play.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning latent plans from play

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.506493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.506493Z digest=sha256:ce009d347ceb3102adbc7d8a3f9d99cf3ebfbdecff8892bad7841c5774e72c12

Observation c3c9599d-7029-4486-af54-3c27bdab6c95 · outbound

This paper cites WHIRL: Human-to-robot imitation in the wild.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video WHIRL: Human-to-robot imitation in the wild

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.509891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.509891Z digest=sha256:c9b96b9058f0ea907a572a120db2bff30beea9c9de038086895f659338087f2c

Observation 112d697a-e618-4412-b1e1-f0b5d9fc2ce6 · outbound

This paper cites MimicPlay: Long-horizon imitation learning by watching human play.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video MimicPlay: Long-horizon imitation learning by watching human play

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.512810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.512810Z digest=sha256:8054db254878fc6e3991e7d50dac8e18ff8a7edc3a4c6b8454448d87f3f78b96

Observation 24973500-e395-4d93-b909-25f2071fbfdf · outbound

This paper cites EgoScale: Scaling dexterous manipulation with diverse egocentric human data.arXiv preprint arXiv:2602.16710, 2026.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video EgoScale: Scaling dexterous manipulation with diverse egocentric human data.arXiv preprint arXiv:2602.16710, 2026

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.516026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.516026Z digest=sha256:654974c0892a06f033bf65882312937b031595ce2f08b98e6ae4b4e377b106db

Observation 390b829b-ebf7-4843-af10-4e5c45e3306d · outbound

This paper cites Graphmimic: Graph-to-graphs generative modeling from videos for policy learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Graphmimic: Graph-to-graphs generative modeling from videos for policy learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.518497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.518497Z digest=sha256:748e1f8b30e3b2fb0284ca0a42d1ef5fae2768d14211a23e5f605996f3730b34

Observation 99768708-1335-45ec-81e5-27d75e9e9cc0 · outbound

This paper cites Learning from videos through graph-to-graphs generative modeling for robotic manipulation.IEEE Transactions on Robotics, 2026.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning from videos through graph-to-graphs generative modeling for robotic manipulation.IEEE Transactions on Robotics, 2026

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.520960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.520960Z digest=sha256:2c120fe82c88c92840dac76873d9f60e7bdc7273081787472e72aa2b2dce5322

Observation 3e7e31b8-880f-4587-915b-20ba7637b14e · outbound

This paper cites Unifying latent action and latent state pre-training for policy learning from videos.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Unifying latent action and latent state pre-training for policy learning from videos

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.523405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.523405Z digest=sha256:d52a30717d2c7f7d35b3f697e3879b4ef1f204877fcc0e875db545b7a79629e1

Observation 781cb97d-4995-4ee7-87a4-b320d779f16d · outbound

This paper cites What foundation models can bring for robot learning in manipulation: A survey.The International Journal of Robotics Research, 45(7):1091–1142, 2026.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video What foundation models can bring for robot learning in manipulation: A survey.The International Journal of Robotics Research, 45(7):1091–1142, 2026

Reference 67

Resolution
verified exact
doi, observed 2026-08-01T11:08:33.882244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-01T11:06:14.526382Z digest=sha256:d24a615913c2a54ebb5217649ccf091d6a6499b3e3b13268ba17363809d622a1

Observation 77ab2abc-3199-4f29-bcdb-d2c98badd9c5 · outbound

This paper cites A generalist agent.Transactions on Machine Learning Research, 2022.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video A generalist agent.Transactions on Machine Learning Research, 2022

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.529125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.529125Z digest=sha256:1aa7fbc43052734dd46cbd52314d1c128cda9ce64d1b55285ceafeaa3f46447e

Observation fb0d1a1d-cf45-4f9b-8a78-edb8201e8529 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video OpenVLA: An Open-Source Vision-Language-Action Model

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.531917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.531917Z digest=sha256:568f48dd0c579df4c3906e0fed1c70ab7e82dc7e0d9c53487474617d067fb8d8

Observation ff948d1e-560f-4c63-bf25-c83dfafa71aa · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.535159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.535159Z digest=sha256:84b7980410c17eb80a8eae520dbcdf75d6206b4ab774ce90da612956d39791d6

Observation 87767916-e1b5-43c8-b83d-3dabcddad686 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.538632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.538632Z digest=sha256:9d1e8279dd8f620df89be0b6afcaed0eb408dbe070f37afd70ae5e15fb56f018

Observation 134f733d-6053-4413-a13e-de46d76da460 · outbound

This paper cites DexVLA: Vision-language model with plug-in diffusion expert for general robot control.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video DexVLA: Vision-language model with plug-in diffusion expert for general robot control

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.541377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.541377Z digest=sha256:f9ad50b8e996f9022d7ead18c73044004fb0284bf5e1f9b5f96d1c67dbb28059

Observation 619e6b1c-7c13-4b27-a510-5be54472b18e · outbound

This paper cites World Models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video World Models

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.544302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.544302Z digest=sha256:9f505ebd89af1f57f2734348e121c97ec145636c3612c911650634c6d65752b9

Observation fd73ad17-0a27-4b2e-a7b0-53b3cbdbfd26 · outbound

This paper cites Christensen, Hao Su, Jiajun Wu, and Yunzhu Li.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Christensen, Hao Su, Jiajun Wu, and Yunzhu Li

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.547616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.547616Z digest=sha256:b47d80687af1174cb6a67f93d4b1187ed8cd3942764704c07c67ff55c857be30

Observation 8e5a81c2-8280-457f-955a-7f23b50b82d3 · outbound

This paper cites Dream to control: Learning behaviors by latent imagination.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Dream to control: Learning behaviors by latent imagination

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.550447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.550447Z digest=sha256:e6e131f1ea714e9a61e9906ec2f462efcbbde1361d51cf596bd65c0ff5c3954e

Observation c0ffaba8-ea49-494c-b52a-8cfd9ed2b18a · outbound

This paper cites Mastering atari with discrete world models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Mastering atari with discrete world models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.552879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.552879Z digest=sha256:7262db1ee6d00d270a821f6399bb55f335b151af084cf5ab769684ad009da3c0

Observation e6306f43-0308-4587-98d6-6ca8181dbcdd · outbound

This paper cites Mastering Diverse Domains through World Models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Mastering Diverse Domains through World Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.555885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.555885Z digest=sha256:47fef3d87d76f8e2d389023d8829841ca872765c70d589bdc1201686adda298f

Observation 3c9e6c71-f3cd-4e93-9efd-501502b663c8 · outbound

This paper cites Transformers are sample-efficient world models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Transformers are sample-efficient world models

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.559114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.559114Z digest=sha256:e991de6939185c2dc71407e15da31a70c7d774149bee855a0c77d42f9cadf21d

Observation a5d7df0e-f679-468d-bf9c-ef9570296d81 · outbound

This paper cites Learning latent dynamics for planning from pixels.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning latent dynamics for planning from pixels

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.561523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.561523Z digest=sha256:a390c0d625fae3d9796f74fe97986384ee9ba29af8ffd0e5dd760ee9a6f923bb

Observation 110f68f1-0386-462c-af46-bf45e0b98119 · outbound

This paper cites TD-MPC2: Scalable, robust world models for continuous control.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video TD-MPC2: Scalable, robust world models for continuous control

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.563950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.563950Z digest=sha256:85d7271bdbab9b1ec04cc3501b2124f2c38418adb823d6b2af1a212842eefe9e

Observation 94f8e605-cca3-4b60-9c1d-4c98d62d3452 · outbound

This paper cites Learning universal policies via text-guided video generation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning universal policies via text-guided video generation

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.566431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.566431Z digest=sha256:ed7d95a30f6824ac304fbe7ebf02a8167ddb35e6c9d06c7be831ed2ec548a52e

Observation b7e1df17-e044-4bd8-9a1c-23d5f8d9bc26 · outbound

This paper cites Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.568905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.568905Z digest=sha256:1a506b8b0a1f0ac88a7c5fe6fbcbb29b31c4198c74d3245e2502b9bfaf3a66bc

Observation edce3f62-877b-42f1-9843-747fa3f95dd1 · outbound

This paper cites Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.571661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.571661Z digest=sha256:97a97aba0b1a0847566055e49a5487ae4a9d895e19866d12019f33f14cb4cfd6

Observation 1e470cd9-3bc3-4d45-aa8f-b3fe29a3c7d6 · outbound

This paper cites Dreamitate: Real-world visuomotor policy learning via video generation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Dreamitate: Real-world visuomotor policy learning via video generation

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.574317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.574317Z digest=sha256:9d8300f480e98e3994ee8a93995c3713a6b4ec2cf0f524e0c4026645113ebf8b

Observation 7d410c0f-bfb4-44a8-93d0-6e17055a09c0 · outbound

This paper cites Structured world models from human videos.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Structured world models from human videos

Reference 85

Resolution
verified exact
doi, observed 2026-08-01T11:08:33.664597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-01T11:06:14.579985Z digest=sha256:774cd7f32c309f1d800241790af79ec4e9f6cbe992f5eb769682d44d22f3fa99

Observation de6fb3e7-8aa7-4af0-8004-88e5f269a0a8 · outbound

This paper cites Learning Interactive Real-World Simulators.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning Interactive Real-World Simulators

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.582467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.582467Z digest=sha256:8b3fb3af79913a0e0a9832344c4dc4507b000675fa98b45f7567e3522550dfc4

Observation 4e04411a-ef19-461f-ac1c-3c2cca0e8351 · outbound

This paper cites Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.586156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.586156Z digest=sha256:171754e3c5c434cf0602e3f1846c16cfc59dc2094850981eb8c6804411a65194

Observation 94374568-be18-4fe0-971c-db4add5251ee · outbound

This paper cites GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.589412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.589412Z digest=sha256:170283c93dbf0edfdd8efc4366bad0a9faca7c1e8f999b341e304509b1d12722

Observation 44cb6556-8c36-4438-a42e-1088cef68574 · outbound

This paper cites Causal World Modeling for Robot Control.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Causal World Modeling for Robot Control

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.592112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.592112Z digest=sha256:a8db51fc2acb9dcae42420ae2210aaf71d4c59dc8540c2eb3ff160414d3a8c34

Observation f8598fce-4266-4f2a-bd57-b6f6e29e4015 · outbound

This paper cites Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.594721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.594721Z digest=sha256:a30433fc5943725663b10aaf98af299d45d642e6c9215b9f514830daf970ba16

Observation 300f0d8f-28ff-4f79-a89a-bace91da2b0c · outbound

This paper cites WALL-WM: Carving World Action Modeling at the Event Joints.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video WALL-WM: Carving World Action Modeling at the Event Joints

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.597777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.597777Z digest=sha256:66bda17a9c8803ca2a594cb57f1e01964e5cc8e5a92e712d4258b71858d118ff

Observation 9f9b48b6-6be0-4948-a395-c3d4df44427f · outbound

This paper cites Exploring the limits of vision-language-action manipulation in cross-task generalization.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Exploring the limits of vision-language-action manipulation in cross-task generalization

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.600746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.600746Z digest=sha256:c292efddf8f7dff384896af1ff269806c71174b23bd05bf81eb95dd418bb72ff

Observation f685af6c-0fe3-46cd-821d-980c8586a807 · outbound

This paper cites Fmimic: Foundation models are fine-grained action learners from human videos.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Fmimic: Foundation models are fine-grained action learners from human videos

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.603451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.603451Z digest=sha256:f2555720acf003eeb5865cf3f323426699a9e74dd489a014876dd89b3cca1b09

Observation 0cb56dfd-3f98-4525-a1aa-7f57f7ad67ec · outbound

This paper cites Vlmimic: Vision language models are visual imitation learner for fine-grained actions.Advances in Neural Information Processing Systems, 37:77860–77887, 2024.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Vlmimic: Vision language models are visual imitation learner for fine-grained actions.Advances in Neural Information Processing Systems, 37:77860–77887, 2024

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.605970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.605970Z digest=sha256:44f940e5e1f1d2ab7069596633edda8fa1590b10008814aa16ba1ad99bece096

Observation 93578dbe-4970-45f0-978f-574bf0687ae0 · outbound

This paper cites Coarse-to-fine imitation learning: Robot manipulation from a single demonstration.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Coarse-to-fine imitation learning: Robot manipulation from a single demonstration

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.608524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.608524Z digest=sha256:78a7517f053206d61840a2162acb9e324aea5f7ac3fc86d25053642ffb18c23a

Observation 2ed0a1e3-c3ea-4c8f-96ed-ee567b0706e6 · outbound

This paper cites OKAMI: Teaching humanoid robots manipulation skills through single video imitation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video OKAMI: Teaching humanoid robots manipulation skills through single video imitation

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.610879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.610879Z digest=sha256:ed35f45eaa9e1fb8e48c77040e8362ee01e666dbae9e5d7973078d71c0fded36

Observation c3c9c71f-3c13-4ca7-9a08-079a55691218 · outbound

This paper cites Learning a thousand tasks in a day.Science Robotics, 10(108):eadv7594, 2025.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning a thousand tasks in a day.Science Robotics, 10(108):eadv7594, 2025

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.613386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.613386Z digest=sha256:9f1d8b4eacf498c83f9b9c7207f94f759a51caf5369af35a9deee7bd4487922b

Observation 0ec193ab-17e1-466c-b07c-06420afd6ece · outbound

This paper cites Robot See Robot Do: Imitating Articulated Object Manipulation with Monocular 4D Reconstruction.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Robot See Robot Do: Imitating Articulated Object Manipulation with Monocular 4D Reconstruction

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.615802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.615802Z digest=sha256:524fa41e8327828574477ab47b89a314f470118bd24ed93f4c2b665df1b58b43

Observation 04e4e0ce-0a87-4c93-be22-07bf401752a7 · outbound

This paper cites Zero-shot visual imitation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Zero-shot visual imitation

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.619387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.619387Z digest=sha256:18e63d8ab30bb8f3d04585f37b4038af1f2f9495bf6b4ca970a2689d8b572716

Observation e9d33f21-11f0-477a-a702-ebbe027d99f7 · outbound

This paper cites Transformers for one-shot visual imitation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Transformers for one-shot visual imitation

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.622405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.622405Z digest=sha256:cc614c278aa044c9fc1e99f6c1fb1f2374b8184d5e417ba01723a8b564dfbaf6

Pith citing papers

No inbound Pith citation observations are available.