Pith. sign in

Paper Citation Record · LEDGER

Robots Acquire Manipulation Skills in Seconds from a Single Human Video

As of 9 August 2026, this Paper Citation Record lists 100 of 117 outbound references and 0 inbound Pith citation observations for arXiv:2607.20033.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.20033 v3

Coverage vector

measured 100 of 117 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:06:14.622405Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 117 outbound references displayed

  • verified exact4
  • verified fuzzy0
  • unresolved95
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 94359ae8-8180-4b6a-9a92-35b8e76da95e · outbound

This paper cites Learning agile and dynamic motor skills for legged robots.Science Robotics, 4(26):eaau5872, 2019.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning agile and dynamic motor skills for legged robots.Science Robotics, 4(26):eaau5872, 2019

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.322238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.322238Z digest=sha256:36d8008c9a13176221774b1514da493a34659c30c2ad597e0a15cfa091cb6b59

Observation 31383d2c-fb02-4955-be36-b420bf8072af · outbound

This paper cites Learning quadrupedal locomotion over challenging terrain.Science Robotics, 5(47):eabc5986, 2020.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning quadrupedal locomotion over challenging terrain.Science Robotics, 5(47):eabc5986, 2020

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.326173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.326173Z digest=sha256:0a8fcc7023db048057b709ee9f11b1344b66180092e429a66ad376cc236d4416

Observation e85ff0d9-0444-46cc-b40e-76cfccf17004 · outbound

This paper cites Visual dexterity: In-hand reorientation of novel and complex object shapes.Science Robotics, 8(84):eadc9244, 2023.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Visual dexterity: In-hand reorientation of novel and complex object shapes.Science Robotics, 8(84):eadc9244, 2023

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.329059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.329059Z digest=sha256:309db19b07387d1203c7ee80eb1926ff8c86ddc305b699b735dea046d089d28e

Observation d91eb503-8646-4787-ab3a-267304942f20 · outbound

This paper cites Solving Rubik's Cube with a Robot Hand.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Solving Rubik's Cube with a Robot Hand

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.332530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.332530Z digest=sha256:2af6590ac779b73da85e888d20a4ba99e83559332ec5948d2ff9d60f623831ae

Observation b2be2375-b5ec-447b-86b5-13d25e217a23 · outbound

This paper cites QT-Opt: Scalable deep reinforcement learning for vision-based robotic manipulation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video QT-Opt: Scalable deep reinforcement learning for vision-based robotic manipulation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.336354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.336354Z digest=sha256:add4e3d772c1df8b581c673ec4bbb149e5fb3c4ff14128a466cf3eb663b6ea13

Observation 7a5d49b2-b912-43d9-ba97-5407b2ad3758 · outbound

This paper cites Scaling up multi-task robotic reinforcement learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Scaling up multi-task robotic reinforcement learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.339512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.339512Z digest=sha256:4d5b36ab7e856084c1ed55da384de1d769980662b45025f51a9627539e4c5bb1

Observation 2db30b1a-54e8-4262-91f8-93d5771ac8c7 · outbound

This paper cites A generalized path integral control approach to reinforce- ment learning.Journal of Machine Learning Research, 11:3137–3181, 2010.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video A generalized path integral control approach to reinforce- ment learning.Journal of Machine Learning Research, 11:3137–3181, 2010

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.345777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.345777Z digest=sha256:ae1ae574cffa28b3453670ecf4cdfcd5ebda047627e1ed248b3ec0232b53bfbb

Observation 56be6551-0e6d-4b8e-ac45-478c1fa5d384 · outbound

This paper cites Path integral guided policy search.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Path integral guided policy search

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.349219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.349219Z digest=sha256:a9fc36e9647f9fff5a23e04f3bca875d8204b605702e0732bc5a6bb1a95021e6

Observation 42f478a6-c7ee-4ffb-bcaa-f124f707500e · outbound

This paper cites Learning coordinated badminton skills for legged manipulators.Science Robotics, 10(102):eadu3922, 2025.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning coordinated badminton skills for legged manipulators.Science Robotics, 10(102):eadu3922, 2025

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.355179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.355179Z digest=sha256:f5f4fb7a5a3d48af38c7fe44a9ef12ee79452b9115573d32057734f51ab48a65

Observation ce5ed6d5-3969-4b24-8c5d-287a8c5c444f · outbound

This paper cites Precise and dexterous robotic manipulation via human-in- the-loop reinforcement learning.Science Robotics, 10(105):eads5033, 2025.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Precise and dexterous robotic manipulation via human-in- the-loop reinforcement learning.Science Robotics, 10(105):eads5033, 2025

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.357852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.357852Z digest=sha256:fac17375c62a16e32908fdfd29f107a74fc5ed2de11a0e892ec65a437375ec1a

Observation 58750782-11ed-4e7e-a708-c6d01e743cf2 · outbound

This paper cites Barreiros, Aykut Özgün Önol, Mengchao Zhang, Sam Creasey, Aimee Goncalves, Andrew Beaulieu, Aditya Bhat, Kate M.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Barreiros, Aykut Özgün Önol, Mengchao Zhang, Sam Creasey, Aimee Goncalves, Andrew Beaulieu, Aditya Bhat, Kate M

Reference 11

Resolution
verified exact
doi, observed 2026-08-01T11:08:35.489516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-01T11:06:14.360554Z digest=sha256:69ca6ecc004edebf0dc71aca1a60d00299a21b6e94f5c8390c908582164d597b

Observation ef832a9f-1fad-490c-8c2a-103f579e6ba3 · outbound

This paper cites Is imitation learning the route to humanoid robots?Trends in Cognitive Sciences, 3(6):233–242, 1999.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Is imitation learning the route to humanoid robots?Trends in Cognitive Sciences, 3(6):233–242, 1999

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.363086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.363086Z digest=sha256:51336235e19c9749dae4b5c4f10b251ce25b200e8b06fb2585cdb3ddca54f230

Observation 81d501af-ed7c-425d-b38c-aba62af36d4c · outbound

This paper cites Sanketi, Grecia Salazar, Michael S.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Sanketi, Grecia Salazar, Michael S

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.365672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.365672Z digest=sha256:a77a958649e6c6d62cdc6eb5588c16c28c22b23b7dedfd59a26157929cdad252

Observation e47ab465-17f2-4f00-96d1-ef4b9c8bc63f · outbound

This paper cites In Proceedings of Robotics: Science and Systems, Los Angeles, CA, USA, 6 2025.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video In Proceedings of Robotics: Science and Systems, Los Angeles, CA, USA, 6 2025

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.368337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.368337Z digest=sha256:3d1c23137b87599ec4be620e908177614941c0add559839dd5c1e5f9c0636742

Observation eb283036-c049-4964-b2ed-01884736ce38 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.370745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.370745Z digest=sha256:b6b826895e0ffda1bbdb295df3346b7a9d3803407dce801f3a61d97d06216b70

Observation 1f836854-ada9-4d92-896d-664d4c65a5da · outbound

This paper cites Toward next-generation learned robot manipulation.Science Robotics, 6(54):eabd9461,.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Toward next-generation learned robot manipulation.Science Robotics, 6(54):eabd9461,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.373831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.373831Z digest=sha256:934367d37faa2443f5a6e42b1565abae90ee9fc252fbd97f00d9eaa62d2fd8eb

Observation 7446c1bd-1e0e-4fed-b958-4fa8a336a585 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Diffusion policy: Visuomotor policy learning via action diffusion

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.379606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.379606Z digest=sha256:24b70dd02611d3aace95badde625664d7463e2f8ab0cde9417d18f34fc05d961

Observation 876d813f-f882-4067-886d-d3ac3bec386f · outbound

This paper cites Learning fine-grained bimanual manipulation with low-cost hardware.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning fine-grained bimanual manipulation with low-cost hardware

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.382048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.382048Z digest=sha256:f40b02d12915de93f2d0b1da2687aa4106dc274ba151ed9e186a1232518a0b9f

Observation d1fcdd1e-5b8e-408f-9026-32a24c71b420 · outbound

This paper cites RT-1: Robotics transformer for real-world control at scale.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video RT-1: Robotics transformer for real-world control at scale

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.385066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.385066Z digest=sha256:c8838354eb4b69dac506d87e187e93bded14f615da8c494bbed899cfafbb322c

Observation 8bb048f5-b077-4cc1-b109-24e8d470aeab · outbound

This paper cites Open X-Embodiment: Robotic learning datasets and RT-X models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Open X-Embodiment: Robotic learning datasets and RT-X models

Reference 20

Resolution
malformed identifier
no resolver link, observed 2026-08-01T11:06:14.388416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.388416Z digest=sha256:03a5f04c8fbd84edd9af3e2a1857c6f49875874f6e2b6d9f7fe9a99b443fdccd

Observation 784b7789-1279-423c-b45d-60a55e13a68d · outbound

This paper cites Octo: An open-source generalist robot policy.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Octo: An open-source generalist robot policy

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.391072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.391072Z digest=sha256:7029a6a3cdd11efde36332ea019f405f21b9518c0b01b6011afd8d68676f9aa1

Observation 7274a360-5f3c-4098-acd1-5d1578617171 · outbound

This paper cites From Intention to Execution: Probing the Generalization Boundaries of Vision-Language-Action Models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video From Intention to Execution: Probing the Generalization Boundaries of Vision-Language-Action Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.393672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.393672Z digest=sha256:cbbb7eff205cf14909d45faa63fcb22ae12bb140dd5ec5cf35e6807b63f90e79

Observation ba9bc7c0-aae1-4b8a-8edf-65c29bb08870 · outbound

This paper cites Ricl: Adding in-context adaptability to pre-trained vision-language-action models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Ricl: Adding in-context adaptability to pre-trained vision-language-action models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.396754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.396754Z digest=sha256:bd959b9a2ffd637b9c3798a64c8f8e8dc6b1c4cc7884ab25ce9fd9268a0665ce

Observation 408baf13-066f-49c7-b04e-bd321caf8875 · outbound

This paper cites FLaRe: Achieving masterful and adaptive robot policies with large-scale reinforcement learning fine-tuning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video FLaRe: Achieving masterful and adaptive robot policies with large-scale reinforcement learning fine-tuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.399270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.399270Z digest=sha256:6693a734d19375dbeb5e827dc97cf340753f5e9eff5069c525f6a644ccb11dbd

Observation a1afcf68-1be3-4145-896d-1bdd1d0be6e6 · outbound

This paper cites an unresolved cited work.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.402247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.402247Z digest=sha256:08ab6aa9d77f19b2f9633ddfbaceeab6e5df0463d7c1a336623280a68997a33a

Observation be24fb01-4931-4523-a20d-81b6d89f6d2b · outbound

This paper cites Learning by watching: Extracting reusable task knowledge from visual observation of human performance.IEEE Transactions on Robotics and Automation, 10(6):799–822, 1994.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning by watching: Extracting reusable task knowledge from visual observation of human performance.IEEE Transactions on Robotics and Automation, 10(6):799–822, 1994

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.404884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.404884Z digest=sha256:5e6352d621ee6288eff64422ff04c298df56ec9cabedf2e4c318d03bf5057f95

Observation 73f553e2-3f32-4fca-ab95-9133b8f05908 · outbound

This paper cites One-shot visual imitation learning via meta-learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video One-shot visual imitation learning via meta-learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.407354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.407354Z digest=sha256:22830ac92ce305a755ae7d6897fc3638f242c2603fb88051e48f89a2060545eb

Observation 69b31d01-597f-455d-b3dc-43433d1e98eb · outbound

This paper cites One-shot imitation from observing humans via domain-adaptive meta-learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video One-shot imitation from observing humans via domain-adaptive meta-learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.409817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.409817Z digest=sha256:c6247f8df4b42742d76dbb6b5758a01737a605a7926d833d0fd190fc347dc4d7

Observation 6b99bef7-0c7f-46e8-9acc-c37fe7e00dca · outbound

This paper cites Towards more generalizable one-shot visual imitation learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Towards more generalizable one-shot visual imitation learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.412233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.412233Z digest=sha256:2c4380b19205691444db6ca3939c46e8f145a49a1f2f8e6bd93db4fd6a6bd8fc

Observation 10a83cc3-ad8f-45c6-8c6f-30a7cfeb88dd · outbound

This paper cites Vid2Robot: 23 End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Vid2Robot: 23 End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers

Reference 30

Resolution
verified exact
doi, observed 2026-08-01T11:08:35.235070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-01T11:06:14.414848Z digest=sha256:b0f7fa59b0dc2af684069679d8065be2fdaa6c373a1fa96aa82a3444f59f8b3c

Observation 7cc6af32-645c-4526-8f3f-64c6eddb73f8 · outbound

This paper cites Osvi-wm: One-shot visual imitation for unseen tasks using world-model-guided trajectory generation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Osvi-wm: One-shot visual imitation for unseen tasks using world-model-guided trajectory generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.417781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.417781Z digest=sha256:f5306d3e07e5776a888df242a27220d7dbbe13eff39d808eab4335d943b3e913

Observation b2f1a0b4-c5e9-4fe8-aa35-0436f800d3ae · outbound

This paper cites Remembering the past to imagine the future: the prospective brain.Nature Reviews Neuroscience, 8(9):657–661, 2007.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Remembering the past to imagine the future: the prospective brain.Nature Reviews Neuroscience, 8(9):657–661, 2007

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.420468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.420468Z digest=sha256:44c6154a66333cb0084a138851ca99745fb86be1aeed419eb3014039b48c364b

Observation d8b1e285-b93f-47a8-9890-c6ea725e75df · outbound

This paper cites Infant imitation after a 1-week delay: long-term memory for novel acts and multiple stimuli.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Infant imitation after a 1-week delay: long-term memory for novel acts and multiple stimuli

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.422854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.422854Z digest=sha256:f350f56dbc0821c0a5a41c89c75b6f3254ae4ad1ed51c4ef30542fab63ca79ff

Observation f9c97278-27f3-4a00-877d-01141e6949c0 · outbound

This paper cites Reinforcement learning and episodic memory in humans and animals: an integrative framework.Annual Review of Psychology, 68:101–128, 2017.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Reinforcement learning and episodic memory in humans and animals: an integrative framework.Annual Review of Psychology, 68:101–128, 2017

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.425807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.425807Z digest=sha256:99aa2400d74e7bfbd7cc51f70815a2ce055795eed40b8e8c21d367e6807363b3

Observation 38020a35-ef05-4b6a-9f53-03a01ce2e132 · outbound

This paper cites Neural simulation of action: a unifying mechanism for motor cognition.NeuroImage, 14(1): S103–S109, 2001.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Neural simulation of action: a unifying mechanism for motor cognition.NeuroImage, 14(1): S103–S109, 2001

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.428457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.428457Z digest=sha256:f33d94ea842cb4e232fc9af062f35b7cd41b061d0f65cf433eba4011e24958d1

Observation 8af4d18e-5d39-435e-b4fd-4e21921b9af2 · outbound

This paper cites A unifying computational framework for motor control and social interaction.Philosophical Transactions of the Royal Society B: Biological Sciences, 358(1431):593–602, 2003.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video A unifying computational framework for motor control and social interaction.Philosophical Transactions of the Royal Society B: Biological Sciences, 358(1431):593–602, 2003

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.431672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.431672Z digest=sha256:2c0425e03b9caab7166463e795c3f3d5a0122538491c634e7054a8caa29868bf

Observation 0fca39c9-e272-4c56-85f8-d64780b8154a · outbound

This paper cites Neurophysiological mechanisms underlying the understanding and imitation of action.Nature Reviews Neuroscience, 2(9):661–670, 2001.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Neurophysiological mechanisms underlying the understanding and imitation of action.Nature Reviews Neuroscience, 2(9):661–670, 2001

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.434086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.434086Z digest=sha256:ff150ac2c6f1db41e437523ed43ee0cebd4839892b7f89104ba6aaa73722de81

Observation be67857c-2a1e-49bd-808a-935bd98a4d1a · outbound

This paper cites Neural circuits underlying imitation learning of hand actions: an event-related fmri study.Neuron, 42(2):323–334,.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Neural circuits underlying imitation learning of hand actions: an event-related fmri study.Neuron, 42(2):323–334,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.436692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.436692Z digest=sha256:41a99be41e2cb909bb3931c3dbe25a0fa6b65bde72ac6f28a4ddcb0422e275a5

Observation 12435650-b02d-4012-bee6-23c93cb7b850 · outbound

This paper cites One-shot visual imitation via attributed waypoints and demonstration augmentation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video One-shot visual imitation via attributed waypoints and demonstration augmentation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.443751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.443751Z digest=sha256:1862dd11477f53e4f9b91d666325ad39210425b0a1d48d7b9b12f8173c1bb5d4

Observation 027ea5b4-0fbb-4fbd-bfc5-ced8b2edf762 · outbound

This paper cites Igniting vlms toward the embodied space.arXiv preprint arXiv:2509.11766, 2025.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Igniting vlms toward the embodied space.arXiv preprint arXiv:2509.11766, 2025

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.446310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.446310Z digest=sha256:d18343ca8e557ce73131a1746dee613c386efc44d4e20cb51e9df7ba4de54fc2

Observation ffa0874f-a180-4039-a09f-10a890c3a654 · outbound

This paper cites Fast-WAM: Do World Action Models Need Test-time Future Imagination?.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Fast-WAM: Do World Action Models Need Test-time Future Imagination?

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.449461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.449461Z digest=sha256:a7cf3a58cccb99773299aaaf61cb1b1eba27e2d66d6ce972f76397cf5ba19278

Observation 51a79043-76a0-40aa-87ae-8ca72bf81b8e · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Lora: Low-rank adaptation of large language models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.453040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.453040Z digest=sha256:e1e3b98e827d85db01edb4b59fdd4e4e7368be3693eec6a2230cc41620f93e5f

Observation 18a546c5-4863-4443-b270-8509f0e6d75c · outbound

This paper cites Tenenbaum, and Alberto Rodriguez.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Tenenbaum, and Alberto Rodriguez

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.456412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.456412Z digest=sha256:079a08fbfea97f5b023d846a28a0dd6845a07371613b9898a090c1293d8d3801

Observation 28a84e25-d66a-41f3-9f52-84fc45183b83 · outbound

This paper cites Representation learning via global temporal alignment and cycle-consistency.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Representation learning via global temporal alignment and cycle-consistency

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.461680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.461680Z digest=sha256:576d2094a18b1cada51c8be4e805d9ed9dd6bcdf3ab3473ed5911c482fbc7f63

Observation 5d8f974d-26c1-411b-884b-e7c22281e876 · outbound

This paper cites Temporal cycle- consistency learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Temporal cycle- consistency learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.464919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.464919Z digest=sha256:148d91af6ef861ec213496e6adf9282a952f33dd8dd3519659fdc6024a0a1e5e

Observation a9f95ce4-39e2-4777-9c47-c9c98dc35ef3 · outbound

This paper cites Qwen3-VL Technical Report.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Qwen3-VL Technical Report

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.467761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.467761Z digest=sha256:f12b782321f51c09c201afc38e80ff22aba310467e535b1b5ba9ffab2373984c

Observation 1040c184-6f17-47fb-8826-4d2ca6522d32 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Wan: Open and Advanced Large-Scale Video Generative Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.471220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.471220Z digest=sha256:ef5ba8b2cdade349c8127916d4c9d191ea9be3dc68fd4bce055934b96b7bd1cf

Observation 0a175204-78dc-43cf-b59f-16cfa9c37749 · outbound

This paper cites Mixture-of-transformers: A sparse and scalable architecture for multi-modal foundation models.Transactions on Machine Learning Research, 2025.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Mixture-of-transformers: A sparse and scalable architecture for multi-modal foundation models.Transactions on Machine Learning Research, 2025

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.474142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.474142Z digest=sha256:cbf5c52255e4e9635228101449d55647ffc057271705527cb392da018b6c55a5

Observation b9ccac01-c547-4ab2-b503-b9602fdd504d · outbound

This paper cites Unimax: Fairer and more effective language sampling for large-scale multilingual pretraining.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Unimax: Fairer and more effective language sampling for large-scale multilingual pretraining

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.476838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.476838Z digest=sha256:13d6e793b5091c726f5eea6f3aedb1357dd7a26e6c7d7c4f4510599fd65c239c

Observation b0eb4d62-8857-46f8-809c-9028f505536d · outbound

This paper cites Flow matching for generative modeling.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Flow matching for generative modeling

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.479147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.479147Z digest=sha256:528aded656ec192c3cee913f72739ef99fb56fa1eb1b55c00a6a3087a71183d1

Observation 700b5210-90ee-4047-854c-bb9053c7b8b1 · outbound

This paper cites A survey of robot learning from demonstra- tion.Robotics and Autonomous Systems, 57(5):469–483, 2009.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video A survey of robot learning from demonstra- tion.Robotics and Autonomous Systems, 57(5):469–483, 2009

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.481539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.481539Z digest=sha256:c942599a7a6aade1b552fd5d960f0c18291ea6791e74afd54de48320b30fed61

Observation df315fcf-07ea-4aab-91b6-32ec55bb714e · outbound

This paper cites Recent advances in robot learning from demonstration.Annual Review of Control, Robotics, and Autonomous Systems, 3:297–330, 2020.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Recent advances in robot learning from demonstration.Annual Review of Control, Robotics, and Autonomous Systems, 3:297–330, 2020

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.484658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.484658Z digest=sha256:b66d66827f6bb58a89a39243a3fdf9fc04d93c3ac622a2ab19639804f62ed29e

Observation 74048569-48a6-4c74-b4d4-7545e985a833 · outbound

This paper cites Deep imitation learning for complex manipulation tasks from virtual reality teleoperation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Deep imitation learning for complex manipulation tasks from virtual reality teleoperation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.487448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.487448Z digest=sha256:08af23981ce6e1cbf689e13999640c523b46ba4a709804a652cbcd93da089efc

Observation d0b8816f-c6c3-4570-b9ea-8996e34313de · outbound

This paper cites White, De Ru Tsai, Richard Jaepyeong Cha, Jeffrey Jopling, Chelsea Finn, and Axel Krieger.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video White, De Ru Tsai, Richard Jaepyeong Cha, Jeffrey Jopling, Chelsea Finn, and Axel Krieger

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.490077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.490077Z digest=sha256:cae0051cd6eaed959fdab47dcebfd0b4487933dfa38650f744f689e28b6aa9ff

Observation 6b61211b-101d-4eea-8335-7ab704e100ff · outbound

This paper cites Pomerleau.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Pomerleau

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.492719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.492719Z digest=sha256:a44018b031455be93e70c6280a7fe6cf3e071d948374a3e76abfe35cecadd654

Observation afa1f508-d928-4466-b0f2-d76c2faf8b7d · outbound

This paper cites DexMV: Imitation learning for dexterous manipulation from human videos.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video DexMV: Imitation learning for dexterous manipulation from human videos

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.495603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.495603Z digest=sha256:b6610d79100d37c60226529021ad596e6aad442cc1b4bd012925887fa10cb77f

Observation c7f7a8c9-ec0e-48c9-86bc-532764a32fc2 · outbound

This paper cites DexCap: Scalable and portable mocap data collection system for dexterous manipulation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video DexCap: Scalable and portable mocap data collection system for dexterous manipulation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.498568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.498568Z digest=sha256:8f73fa23f642daaee1096436b21f79abce7fcf9641ccadf6c5410b7e7dea70b3

Observation afb8bc1c-e540-4cb2-94c9-7951c2fd3680 · outbound

This paper cites Time- contrastive networks: Self-supervised learning from video.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Time- contrastive networks: Self-supervised learning from video

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.501559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.501559Z digest=sha256:21ff7db183e81ae5755eedadb3cc9e15b2f9c72dd4e5022c1193fb6d988a83ae

Observation d4c9585e-5891-4164-8631-1553896065e1 · outbound

This paper cites XIRL: Cross-embodiment inverse reinforcement learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video XIRL: Cross-embodiment inverse reinforcement learning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.504183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.504183Z digest=sha256:dc45deb25645a587af8e481a8f3d6f8e90cdf998899bdf08a3dd59a10d6d1c7e

Observation 1e33ef4b-0629-4aa6-85b2-fcd1cd9099d5 · outbound

This paper cites Learning latent plans from play.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning latent plans from play

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.506493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.506493Z digest=sha256:d252567d4bd4c5460fb106ea00c7e1001ab81d588301b2266531454a89ff9f72

Observation c3c9599d-7029-4486-af54-3c27bdab6c95 · outbound

This paper cites WHIRL: Human-to-robot imitation in the wild.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video WHIRL: Human-to-robot imitation in the wild

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.509891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.509891Z digest=sha256:b9d8e7acaf67b7293c53444265b4cceab36868938633a7fd6be85dd7874a83b7

Observation 112d697a-e618-4412-b1e1-f0b5d9fc2ce6 · outbound

This paper cites MimicPlay: Long-horizon imitation learning by watching human play.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video MimicPlay: Long-horizon imitation learning by watching human play

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.512810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.512810Z digest=sha256:c637706a2d9290a32710ec96565114b2fcb86876745acd84dccff0e943c5a55e

Observation 24973500-e395-4d93-b909-25f2071fbfdf · outbound

This paper cites EgoScale: Scaling dexterous manipulation with diverse egocentric human data.arXiv preprint arXiv:2602.16710, 2026.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video EgoScale: Scaling dexterous manipulation with diverse egocentric human data.arXiv preprint arXiv:2602.16710, 2026

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.516026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.516026Z digest=sha256:43e333f7b87be18599678373fc51502dacc013eeb91d38a6be35ff3dd2618b15

Observation 390b829b-ebf7-4843-af10-4e5c45e3306d · outbound

This paper cites Graphmimic: Graph-to-graphs generative modeling from videos for policy learning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Graphmimic: Graph-to-graphs generative modeling from videos for policy learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.518497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.518497Z digest=sha256:d4dc8e03ad264a602fccba6d0ea1ffd201cc09c28e852834ae38068b380c91b9

Observation 99768708-1335-45ec-81e5-27d75e9e9cc0 · outbound

This paper cites Learning from videos through graph-to-graphs generative modeling for robotic manipulation.IEEE Transactions on Robotics, 2026.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning from videos through graph-to-graphs generative modeling for robotic manipulation.IEEE Transactions on Robotics, 2026

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.520960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.520960Z digest=sha256:381b40dadba5feaada8e852e96f5bc5d3df1d6b3aca14f5f3a99b04b34dc9b6e

Observation 3e7e31b8-880f-4587-915b-20ba7637b14e · outbound

This paper cites Unifying latent action and latent state pre-training for policy learning from videos.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Unifying latent action and latent state pre-training for policy learning from videos

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.523405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.523405Z digest=sha256:4160f22537e95dca3725f608ba409be80f57821ac25c0871788c5aad949a7c6f

Observation 781cb97d-4995-4ee7-87a4-b320d779f16d · outbound

This paper cites What foundation models can bring for robot learning in manipulation: A survey.The International Journal of Robotics Research, 45(7):1091–1142, 2026.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video What foundation models can bring for robot learning in manipulation: A survey.The International Journal of Robotics Research, 45(7):1091–1142, 2026

Reference 67

Resolution
verified exact
doi, observed 2026-08-01T11:08:33.882244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-01T11:06:14.526382Z digest=sha256:c371deee5d8d0ae8d5426360609bb825822a5bb1f59cfc7d510de5a8612ad7c2

Observation 77ab2abc-3199-4f29-bcdb-d2c98badd9c5 · outbound

This paper cites A generalist agent.Transactions on Machine Learning Research, 2022.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video A generalist agent.Transactions on Machine Learning Research, 2022

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.529125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.529125Z digest=sha256:407ac6a4c961916028935c6812adfb61286036b0a9238426f90b5346a17664b9

Observation fb0d1a1d-cf45-4f9b-8a78-edb8201e8529 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video OpenVLA: An Open-Source Vision-Language-Action Model

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.531917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.531917Z digest=sha256:d5332622ee13bd61bf416060e69d8fbe68e0a9df502a2bb138136ddcc2cdb6da

Observation ff948d1e-560f-4c63-bf25-c83dfafa71aa · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.535159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.535159Z digest=sha256:c6e9b34b6289bd6e5708a7076125eeda4d1b091fda891bec1212d509cd7f329b

Observation 87767916-e1b5-43c8-b83d-3dabcddad686 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.538632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.538632Z digest=sha256:f2a97eab6c1fedc52a005f9ff8e97dd07098ff318790c8eecb68252848a873cd

Observation 134f733d-6053-4413-a13e-de46d76da460 · outbound

This paper cites DexVLA: Vision-language model with plug-in diffusion expert for general robot control.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video DexVLA: Vision-language model with plug-in diffusion expert for general robot control

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.541377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.541377Z digest=sha256:a18bb460348d5e4850ffd9daed59210eeab9874304e0036ab48a48ed21aa5df5

Observation 619e6b1c-7c13-4b27-a510-5be54472b18e · outbound

This paper cites World Models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video World Models

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.544302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.544302Z digest=sha256:e4d392339b866739cc4cd5af4f24e7b367cef6a49344409b20050cc325526e77

Observation fd73ad17-0a27-4b2e-a7b0-53b3cbdbfd26 · outbound

This paper cites Christensen, Hao Su, Jiajun Wu, and Yunzhu Li.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Christensen, Hao Su, Jiajun Wu, and Yunzhu Li

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.547616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.547616Z digest=sha256:0de64678ed551df344d69f8b0923e81c6ff25cf53c8fa08401c9f7d5c6045471

Observation 8e5a81c2-8280-457f-955a-7f23b50b82d3 · outbound

This paper cites Dream to control: Learning behaviors by latent imagination.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Dream to control: Learning behaviors by latent imagination

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.550447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.550447Z digest=sha256:9ad6953cdfd9c03e34ee7d0886c23eec99a0100f05d0c418574f9323344c6a94

Observation c0ffaba8-ea49-494c-b52a-8cfd9ed2b18a · outbound

This paper cites Mastering atari with discrete world models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Mastering atari with discrete world models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.552879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.552879Z digest=sha256:9569d7743dd082d7e196e73f9346e23be93de370c39a8c70a45cafc0f25f0917

Observation e6306f43-0308-4587-98d6-6ca8181dbcdd · outbound

This paper cites Mastering Diverse Domains through World Models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Mastering Diverse Domains through World Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.555885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.555885Z digest=sha256:1750559eeed7d375f137f1002ccf4ae0db63bf0fc29e7281daca776fcd39c0ec

Observation 3c9e6c71-f3cd-4e93-9efd-501502b663c8 · outbound

This paper cites Transformers are sample-efficient world models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Transformers are sample-efficient world models

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.559114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.559114Z digest=sha256:7d93e0bff67f9759cb218a11263059a9e9d0fd5c16965443f39ec9da2c6079e5

Observation a5d7df0e-f679-468d-bf9c-ef9570296d81 · outbound

This paper cites Learning latent dynamics for planning from pixels.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning latent dynamics for planning from pixels

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.561523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.561523Z digest=sha256:35973ca71c198bb809292207e10adcf09461b3aed8113bcace563ccf71daae1a

Observation 110f68f1-0386-462c-af46-bf45e0b98119 · outbound

This paper cites TD-MPC2: Scalable, robust world models for continuous control.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video TD-MPC2: Scalable, robust world models for continuous control

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.563950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.563950Z digest=sha256:edfcd2e9d0a0f96cf6177214bed41d88f8a2545753ae120f2ee43b4705ff6847

Observation 94f8e605-cca3-4b60-9c1d-4c98d62d3452 · outbound

This paper cites Learning universal policies via text-guided video generation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning universal policies via text-guided video generation

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.566431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.566431Z digest=sha256:c5af07f06eccde226e0768d380dc41ecb5bc553aec1dd9ea04552775f4f3f485

Observation b7e1df17-e044-4bd8-9a1c-23d5f8d9bc26 · outbound

This paper cites Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.568905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.568905Z digest=sha256:ed5228f2fdaca61859c65bff0f13af1bd2209edffda6c8ddffe5b43cbbade4ad

Observation edce3f62-877b-42f1-9843-747fa3f95dd1 · outbound

This paper cites Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.571661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.571661Z digest=sha256:b30e66d6867c76f0d1d63138626a3f3a8ea1657672c8b2340413f028a7f0f601

Observation 1e470cd9-3bc3-4d45-aa8f-b3fe29a3c7d6 · outbound

This paper cites Dreamitate: Real-world visuomotor policy learning via video generation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Dreamitate: Real-world visuomotor policy learning via video generation

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.574317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.574317Z digest=sha256:5b2576e3eba17d0528402814158c2f7b10b5db866752589b361116b1532956fc

Observation 7d410c0f-bfb4-44a8-93d0-6e17055a09c0 · outbound

This paper cites Structured world models from human videos.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Structured world models from human videos

Reference 85

Resolution
verified exact
doi, observed 2026-08-01T11:08:33.664597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-01T11:06:14.579985Z digest=sha256:4138830125a32a1b7df47a7d82e7d7fec6ca1beb6997c3981300353b00eec646

Observation de6fb3e7-8aa7-4af0-8004-88e5f269a0a8 · outbound

This paper cites Learning Interactive Real-World Simulators.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning Interactive Real-World Simulators

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.582467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.582467Z digest=sha256:e062690952f4db897f2003766a9be2afe657131b568f23dfe43ed126be8e9524

Observation 4e04411a-ef19-461f-ac1c-3c2cca0e8351 · outbound

This paper cites Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.586156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.586156Z digest=sha256:bfb1965728ae0fb6c50ada051db14a959c8eafccb362ecfc95e1f54acba37207

Observation 94374568-be18-4fe0-971c-db4add5251ee · outbound

This paper cites GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.589412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.589412Z digest=sha256:2fdbe860e6e85f77bf50123d6fbb5cc455684696b520950c5324533fec6f106b

Observation 44cb6556-8c36-4438-a42e-1088cef68574 · outbound

This paper cites Causal World Modeling for Robot Control.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Causal World Modeling for Robot Control

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.592112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.592112Z digest=sha256:e5c5e1a1c9b23d01d21df25496c46500d8f4019c4670fe8b42d76c715e98c3e8

Observation f8598fce-4266-4f2a-bd57-b6f6e29e4015 · outbound

This paper cites Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.594721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.594721Z digest=sha256:7b2c22646485f8d3b624e65b1e643b35b6df5af084800f6d960dd09164da9172

Observation 300f0d8f-28ff-4f79-a89a-bace91da2b0c · outbound

This paper cites WALL-WM: Carving World Action Modeling at the Event Joints.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video WALL-WM: Carving World Action Modeling at the Event Joints

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.597777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.597777Z digest=sha256:5a3519c0057d093cf2adb3c283a90d96d6f1a36e100bb2bc0ddc71813ff10e59

Observation 9f9b48b6-6be0-4948-a395-c3d4df44427f · outbound

This paper cites Exploring the limits of vision-language-action manipulation in cross-task generalization.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Exploring the limits of vision-language-action manipulation in cross-task generalization

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.600746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.600746Z digest=sha256:c304851b4d165b6032cf6a1b339dbce9be0abe54f4d65acf0a2d58598aaf1fa6

Observation f685af6c-0fe3-46cd-821d-980c8586a807 · outbound

This paper cites Fmimic: Foundation models are fine-grained action learners from human videos.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Fmimic: Foundation models are fine-grained action learners from human videos

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.603451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.603451Z digest=sha256:1b8e6bfdd8c16e5094655e052724ed7d70409b89557769b13a17f3ad23f5fdd5

Observation 0cb56dfd-3f98-4525-a1aa-7f57f7ad67ec · outbound

This paper cites Vlmimic: Vision language models are visual imitation learner for fine-grained actions.Advances in Neural Information Processing Systems, 37:77860–77887, 2024.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Vlmimic: Vision language models are visual imitation learner for fine-grained actions.Advances in Neural Information Processing Systems, 37:77860–77887, 2024

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.605970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.605970Z digest=sha256:e5c8c262a470af3f251d7effa2c7f64f6647061e1567e39ad9f85395003bacd9

Observation 93578dbe-4970-45f0-978f-574bf0687ae0 · outbound

This paper cites Coarse-to-fine imitation learning: Robot manipulation from a single demonstration.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Coarse-to-fine imitation learning: Robot manipulation from a single demonstration

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.608524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.608524Z digest=sha256:3614ffc662c9d2b7b4fa9d986bc34900f85170b9e37719fc52b290b2a1908a64

Observation 2ed0a1e3-c3ea-4c8f-96ed-ee567b0706e6 · outbound

This paper cites OKAMI: Teaching humanoid robots manipulation skills through single video imitation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video OKAMI: Teaching humanoid robots manipulation skills through single video imitation

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.610879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.610879Z digest=sha256:07844affe68729c20b75bb5b828411f6621a9dae260b4536e27e9a7deac093a7

Observation c3c9c71f-3c13-4ca7-9a08-079a55691218 · outbound

This paper cites Learning a thousand tasks in a day.Science Robotics, 10(108):eadv7594, 2025.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Learning a thousand tasks in a day.Science Robotics, 10(108):eadv7594, 2025

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.613386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.613386Z digest=sha256:032ecb661d2c2cca3489040a8398cf066b886fa43a6a2684e2ecd4f955bc6996

Observation 0ec193ab-17e1-466c-b07c-06420afd6ece · outbound

This paper cites Robot See Robot Do: Imitating Articulated Object Manipulation with Monocular 4D Reconstruction.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Robot See Robot Do: Imitating Articulated Object Manipulation with Monocular 4D Reconstruction

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.615802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.615802Z digest=sha256:c95ac9775614bf9b225829db94f68b5dc1dc8c41995d74f1b951c89beec0eabb

Observation 04e4e0ce-0a87-4c93-be22-07bf401752a7 · outbound

This paper cites Zero-shot visual imitation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Zero-shot visual imitation

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.619387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.619387Z digest=sha256:d7aefa1fe875578d94a5cb282f3835c4521ba38f4f265df0112015dc12252140

Observation e9d33f21-11f0-477a-a702-ebbe027d99f7 · outbound

This paper cites Transformers for one-shot visual imitation.

Robots Acquire Manipulation Skills in Seconds from a Single Human Video Transformers for one-shot visual imitation

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-01T11:06:14.622405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:06:14.622405Z digest=sha256:b4017c8badd02f8342192e7351b6e9f2fb518827397af9a642fd20cf0c4150bd

Pith citing papers

No inbound Pith citation observations are available.