Pith. sign in

Paper Citation Record · LEDGER

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

As of 13 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 9 inbound Pith citation observations for arXiv:2509.09769.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.09769 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T18:51:36.127656Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T06:48:14.554799Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T22:16:36.335996Z

Reference resolution

59 of 59 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved59
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 49b4c5c5-1134-42d7-890b-13f411560d03 · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Model-agnostic meta-learning for fast adaptation of deep networks,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.979246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.979246Z digest=sha256:09a46845d4fa6f59d7638e5ae34eac3bb09d16078ceb4a112c70b70be38b65a6

Observation 8bcd6db7-8ff7-4cde-a333-c674cf4f8a72 · outbound

This paper cites One-Shot Imitation from Observing Humans via Domain-Adaptive Meta-Learning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos One-Shot Imitation from Observing Humans via Domain-Adaptive Meta-Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.982472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.982472Z digest=sha256:fde36de63be7ac77c62beb22e46adc4abd76d9b99463e9f9337948698238feb7

Observation f5075199-53e9-4f50-bbbc-2fd451b1e0c7 · outbound

This paper cites Task-embedded control networks for few-shot imitation learning,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Task-embedded control networks for few-shot imitation learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.986558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.986558Z digest=sha256:c2ed2acd24fc8b9d2929869f0c10251c0af1eec312dff87730700bfcf8501b47

Observation 4e79cee5-c384-42e8-8fa6-e3c52dea7609 · outbound

This paper cites Language models are few-shot learners,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Language models are few-shot learners,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.989663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.989663Z digest=sha256:8e8dabd3ddd0cc9c52e84dd3bf0e6a6608ec3c65f6d22bb4f22c69a010ead1f3

Observation b0233dc6-c36b-4f2d-b522-0825b6b43474 · outbound

This paper cites Flamingo: A visual language model for few- shot learning,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Flamingo: A visual language model for few- shot learning,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.992047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.992047Z digest=sha256:d35f5ea69521935a06817baf25cb0a1e27d37c48971fb28d0275794efce3cc62

Observation 724a961b-56b2-4e88-b090-13a0d7358b64 · outbound

This paper cites Prompting decision transformer for few-shot pol- icy generalization,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Prompting decision transformer for few-shot pol- icy generalization,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.995043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.995043Z digest=sha256:6390450e9eb4c3856424ff0428321d9dd390d558588efd80a65dfa7945d70e00

Observation c0b02a88-c2be-414f-b846-d0f8fbb38d4f · outbound

This paper cites AMAGO: Scalable In-Context Reinforcement Learning for Adaptive Agents.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos AMAGO: Scalable In-Context Reinforcement Learning for Adaptive Agents

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:35.998019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:35.998019Z digest=sha256:cec79d706d1560e02ee7734599e3ca8ed92ca3f09de67f9f218702635e70ed16

Observation 49b1f0ed-07a2-448f-97d6-684f2da3d7d8 · outbound

This paper cites In-context Reinforcement Learning with Algorithm Distillation.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos In-context Reinforcement Learning with Algorithm Distillation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.001004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.001004Z digest=sha256:55b3039088cbe4c12724619107cc61a413a8b736de1502e8265c30a636ffe665

Observation d217f832-5e6c-455a-b125-f53f81695552 · outbound

This paper cites REGENT: A Retrieval-Augmented Generalist Agent That Can Act In-Context in New Environments.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos REGENT: A Retrieval-Augmented Generalist Agent That Can Act In-Context in New Environments

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.003802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.003802Z digest=sha256:ea4ddada44c3c31a2631e3f9e7096e4791bb27239fd753c83994ddeb99334784

Observation d1f3c36e-f9f0-42d6-9f23-d465ce58dc57 · outbound

This paper cites Keypoint Action Tokens Enable In-Context Imitation Learning in Robotics.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Keypoint Action Tokens Enable In-Context Imitation Learning in Robotics

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.007414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.007414Z digest=sha256:d9a2a957389ebb39cd03a00eea02702d78d4ccc191112db8bdec79af5999b3d6

Observation e79e46de-9492-40d1-a858-2b00bc01fdc0 · outbound

This paper cites In-Context Imitation Learning via Next-Token Prediction.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos In-Context Imitation Learning via Next-Token Prediction

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.010783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.010783Z digest=sha256:e89bdab5ea536154d270d824a4ea0b0832174f6accd36def2f8f7ac34fda64ac

Observation bc756b72-5d4d-4090-afa6-a22a82fce9ce · outbound

This paper cites Instant Policy: In-Context Imitation Learning via Graph Diffusion.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Instant Policy: In-Context Imitation Learning via Graph Diffusion

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.013303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.013303Z digest=sha256:afcbd4d4636e19f9219a3b51342dd1fd9c06ff34dafdcd9e8e063364f5414fb9

Observation 93d744c2-b19e-4a63-8727-6175ad7260bc · outbound

This paper cites Generalization to New Sequential Decision Making Tasks with In-Context Learning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Generalization to New Sequential Decision Making Tasks with In-Context Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.016697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.016697Z digest=sha256:5104aae596d26d8a8a0b2d97f0176e0fa3af2bb760359478c4ef3570b33a5a99

Observation 54e1324d-dd74-4240-b012-5a1fbc21dcaf · outbound

This paper cites Benchmarking General-Purpose In-Context Learning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Benchmarking General-Purpose In-Context Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.019940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.019940Z digest=sha256:4116c578dcbf72da2bf92d2ca0a08e45505572aff3b1d9cb764530497d13359d

Observation b0fb77fe-8fd0-4cb2-b171-cc76dc4ea49f · outbound

This paper cites General-Purpose In-Context Learning by Meta-Learning Transformers.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos General-Purpose In-Context Learning by Meta-Learning Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.022159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.022159Z digest=sha256:48d97b876ba85733d9eb21f56dddf4604ec476a3da244a61309b18b08106fcd1

Observation 6f9e3fb2-2f4b-4f51-8dde-23275b7f18c2 · outbound

This paper cites RICL: Adding In-Context Adaptability to Pre-Trained Vision-Language-Action Models.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos RICL: Adding In-Context Adaptability to Pre-Trained Vision-Language-Action Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.024764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.024764Z digest=sha256:5bc89bb471c19b9a161f631d9b4566eb99b04d9fec26686e54b0c198b74b1d80

Observation def31ae2-80fe-48f9-96b1-d8309a8ea97c · outbound

This paper cites Roboturk: A crowdsourcing platform for robotic skill learning through imitation,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Roboturk: A crowdsourcing platform for robotic skill learning through imitation,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.027742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.027742Z digest=sha256:c4202cb03a1309f0cf0ca4f6083f42b0873a7ca42c805e1007e42b6f15e7c65f

Observation 07f99e7f-6d36-4972-b129-f6ba770ffeb3 · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.029863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.029863Z digest=sha256:b45296773128e22b38162183e4d9492e6a155b5f92f835a44a2cb63c897c6edf

Observation b74bd45c-f74f-4eaa-84d2-b3ffa5aaa2cd · outbound

This paper cites Droid: A large-scale in-the-wild robot manipulation dataset,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Droid: A large-scale in-the-wild robot manipulation dataset,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.031973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.031973Z digest=sha256:c438fd2018eef1ac3a97f4ab3566fe6a541342c2bf82e9873dc81154eb7fa4e7

Observation 8b1c8003-e080-485b-b741-df5d63ee743f · outbound

This paper cites MimicPlay: Long-Horizon Imitation Learning by Watching Human Play.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos MimicPlay: Long-Horizon Imitation Learning by Watching Human Play

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.033954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.033954Z digest=sha256:2bbe4e76a95c872b16115611d0e5bdf988ee94c01430166982d1b28ce3c92ef4

Observation 37bb429d-0435-4457-96ac-f0202c3c5f22 · outbound

This paper cites Learning latent plans from play,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Learning latent plans from play,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.037439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.037439Z digest=sha256:3636d559614ffea51db47769f9dd1326c6425f85aea978393dd10f3cbf0ccd22

Observation 9bf8e2bb-386b-424e-8cd8-5854ade764b1 · outbound

This paper cites MetaICL: Learning to Learn In Context.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos MetaICL: Learning to Learn In Context

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.039660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.039660Z digest=sha256:aaf81fef3d59aaf64d16619620dad09ac3dbbca5abc37c56fbd6b22198f6cecf

Observation dd95dc28-dced-43dd-8a9d-070730f29998 · outbound

This paper cites Okami: Teaching humanoid robots manipulation skills through single video imitation,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Okami: Teaching humanoid robots manipulation skills through single video imitation,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.042123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.042123Z digest=sha256:1394495c645d90ed18803f49f5c660e85d822162f5393d67a4b715e5b50906f4

Observation 8a51b61e-29aa-4413-aa4d-e71152c903c6 · outbound

This paper cites Humanoid policy˜ human policy,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Humanoid policy˜ human policy,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.044696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.044696Z digest=sha256:f80d88d220086ddc1ff69e14f603a1eca45c1f99ac4f98b68e2e20d505e9afbc

Observation 509fba2d-f547-4d7d-a2b1-a95c9a41648b · outbound

This paper cites Representation and control of the task space in humans and humanoid robots,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Representation and control of the task space in humans and humanoid robots,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.046884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.046884Z digest=sha256:d07ea2de3c753ce858af310ef76e79c05e12c05bc6b03ef0b2641952fbc69cb6

Observation f88b31c8-61c2-44e0-8b32-f35e88c2554f · outbound

This paper cites Masked autoencoders are scalable vision learners,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Masked autoencoders are scalable vision learners,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.049590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.049590Z digest=sha256:3f4f68f73722cfa0160e6e06595cca25b6e4d700230af4bd7046b7e38a1e69ed

Observation 3ec0d3bc-0cb9-44a5-ae09-2088c2739165 · outbound

This paper cites Rethinking Patch Dependence for Masked Autoencoders.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Rethinking Patch Dependence for Masked Autoencoders

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.051840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.051840Z digest=sha256:be67df52f14810073e25581385bde996c8d45f2a890de416765e5dd2eb98cae9

Observation 10acbeeb-755b-47a5-817c-206621baabf8 · outbound

This paper cites Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.054289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.054289Z digest=sha256:2fec0f255c528ec05eacb6bfcfd252333231368e805abd19f312be477ad7ab48

Observation 40cd3cc1-a873-4d66-b080-b5d106956de5 · outbound

This paper cites Zero-Shot Robot Manipulation from Passive Human Videos.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Zero-Shot Robot Manipulation from Passive Human Videos

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.056679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.056679Z digest=sha256:f035ddfe3304c10d33fc16e67bde7d333b1a3ae1db0e323ab52b31f70611ed5c

Observation b60583e3-30e0-459a-849a-ee1c62674b06 · outbound

This paper cites Hand me the data: Fast robot adaptation via hand path retrieval,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Hand me the data: Fast robot adaptation via hand path retrieval,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.058929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.058929Z digest=sha256:cbc89cdea628172d217eef7562aa7e5046f8f0702a2b32104e69111239fa6b8c

Observation 918742f8-31b4-4d7f-ad6f-a71874cdc860 · outbound

This paper cites Robocasa: Large-scale simulation of ev- eryday tasks for generalist robots,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Robocasa: Large-scale simulation of ev- eryday tasks for generalist robots,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.060965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.060965Z digest=sha256:a54af57409f35b8cd82feddc1769930ff986699e7c7b8999bb5831e17eaa329b

Observation 8ebeba5f-2ff0-44c0-b08a-a321f91e036d · outbound

This paper cites robosuite: A Modular Simulation Framework and Benchmark for Robot Learning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos robosuite: A Modular Simulation Framework and Benchmark for Robot Learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.063023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.063023Z digest=sha256:12d1f80fa889d9334a45ea24a4a9c196042bf85513d2638fd4cc7f614fc0208c

Observation 32148449-3222-4240-9882-2ea1883395bd · outbound

This paper cites Evolutionary principles in self-referential learning,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Evolutionary principles in self-referential learning,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.065798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.065798Z digest=sha256:de9dea546464b50692fe4846e66150b0f4ccf5c5f8fb78edf63c84d47b116dbc

Observation aee686bf-a4e2-4d51-80de-505bdfcf9e6b · outbound

This paper cites Meta-neural networks that learn by learning,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Meta-neural networks that learn by learning,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.068210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.068210Z digest=sha256:d1cb94d3f2723421c88dfe8907ca93cb6aad159f12f2fc696c5b5f8e248cf44e

Observation fa7d62ca-4592-4742-9683-69c6a161def8 · outbound

This paper cites Meta-learning with memory-augmented neu- ral networks,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Meta-learning with memory-augmented neu- ral networks,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.070278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.070278Z digest=sha256:d21427dbabff5564482346d6c3d838dfe574bba73bd8e15f543beca6bf73d03f

Observation 02400fbd-9494-42c3-b527-6552e639eb97 · outbound

This paper cites Learning to learn using gradient descent,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Learning to learn using gradient descent,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.072563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.072563Z digest=sha256:c6521d38d44d749566d9025809469b885d6c0b04888d002ffd17562b15b6fa6b

Observation adc6d6fb-deb2-42c8-9d04-e7115d8d3350 · outbound

This paper cites RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.074938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.074938Z digest=sha256:c9e90aa4ebf80df69c7318b4eecaeb89a99965fcfcc6421e543e85b10f539fd8

Observation cdc1fc02-0533-4b80-9b16-a0a1c38eb02e · outbound

This paper cites Attention is all you need,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Attention is all you need,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.077854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.077854Z digest=sha256:5462530e6d606065e7998bcf3540bed08b074f2914e01e8f80a5abee20b6ca46

Observation dca32268-79bc-46cc-8be0-8ab31de29413 · outbound

This paper cites RRL: Resnet as representation for Reinforcement Learning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos RRL: Resnet as representation for Reinforcement Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.079940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.079940Z digest=sha256:25321cd5fc80e846908e3b8ec3dac79eb828ba266fc1e8cc4e2d68eeefb4735a

Observation c69ccdf7-8877-4ac6-ae8a-ee135e90f0c8 · outbound

This paper cites R3M: A Universal Visual Representation for Robot Manipulation.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos R3M: A Universal Visual Representation for Robot Manipulation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.082371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.082371Z digest=sha256:b01aa46a402ff03d032d1f0e3d5faf12d5aea9bdc9f9067df490d8679e490b13

Observation 5d7d30ec-7118-44fe-994b-5689878dff2d · outbound

This paper cites Where are we in the search for an artificial visual cortex for embodied intelligence?.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Where are we in the search for an artificial visual cortex for embodied intelligence?

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.085131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.085131Z digest=sha256:77fa43fba6bdd16481593fb961e592a933b46be48f07a9bd4376942c68d50413

Observation 873910a3-fe4f-49e8-b716-9e1ff0868e50 · outbound

This paper cites Concept2robot: Learning manipulation concepts from instructions and human demonstrations,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Concept2robot: Learning manipulation concepts from instructions and human demonstrations,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.087661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.087661Z digest=sha256:c04ef73085a2026bfdb98458e5a10cea80a545476819d7fb365d06574f35974f

Observation c4a10568-dce8-441f-8e13-4eb6d82b2fa4 · outbound

This paper cites Liv: Language-image representations and rewards for robotic control,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Liv: Language-image representations and rewards for robotic control,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.089935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.089935Z digest=sha256:0ba6aa7e174cfb13486f1e14733fd4876346a9fa2c8e2f0f7445948ff828c343

Observation bb877225-36ee-4d3a-a5ee-bbf5e02b9119 · outbound

This paper cites VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.092177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.092177Z digest=sha256:94cbb44d0e204464416590c970f2cc542980bcd1d81e5497d41c59e3cde26b49

Observation ab8183e0-9cea-4518-8edf-f37c2a33ba02 · outbound

This paper cites Bridging the Human to Robot Dexterity Gap through Object-Oriented Rewards.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Bridging the Human to Robot Dexterity Gap through Object-Oriented Rewards

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.094634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.094634Z digest=sha256:670f78f4147bd2d5931853cd773e91e3eae7a213d9c7b146b883311d4d2ebb39

Observation 2e30f510-7f36-47ca-a90a-90391e87da85 · outbound

This paper cites ScrewMimic: Bimanual Imitation from Human Videos with Screw Space Projection.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos ScrewMimic: Bimanual Imitation from Human Videos with Screw Space Projection

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.096917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.096917Z digest=sha256:8b68ac14b9268a21219352f829bf4d414e461d818c08de121052f8c5adadcfcf

Observation 32091f57-4abb-4acf-95f9-b25915a68215 · outbound

This paper cites EgoMimic: Scaling Imitation Learning via Egocentric Video.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos EgoMimic: Scaling Imitation Learning via Egocentric Video

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.099335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.099335Z digest=sha256:cac27fd744322a4ef5185a8fc2fdf0235ae323bedfb1d2f2c4c3f2a96ee84604

Observation 7e090c30-8b6c-4e10-8bf9-d753907dd415 · outbound

This paper cites an unresolved cited work.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.101720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.101720Z digest=sha256:5f5f8037441100d2490568a9ba422bd226daed7d7eab73b13fd5896ee1bb3f66

Observation 9fde1f81-89cf-4b44-bd30-56569c5c6e50 · outbound

This paper cites Reconstructing hands in 3D with transform- ers,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Reconstructing hands in 3D with transform- ers,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.103866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.103866Z digest=sha256:39a72678754fbedfe293f2b13079a15cfa8a85d13c5503078317ecb8f64a65f8

Observation 95d91406-94d2-4a22-a1e9-38d3cc1a2ba9 · outbound

This paper cites Phantom: Training Robots Without Robots Using Only Human Videos.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Phantom: Training Robots Without Robots Using Only Human Videos

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.106113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.106113Z digest=sha256:0507de7435e0d20167b091afea25449cc7a89aff31f316c128a67c1ac891dc2c

Observation f3066f7f-e8a0-4f0c-8bc1-bb9f7a5c211f · outbound

This paper cites Vision-based Manipulation from Single Human Video with Open-World Object Graphs.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Vision-based Manipulation from Single Human Video with Open-World Object Graphs

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.108784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.108784Z digest=sha256:c11a4c9a77a29090697e160bc0698979332bd5f08cd39de8b74d9f0e18f306d9

Observation b7a1aacb-ddf0-4850-9ced-6162638e0d3f · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos DINOv2: Learning Robust Visual Features without Supervision

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.111093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.111093Z digest=sha256:7bfe6c57f3f9fa2cf05e62bfd306847fa22ba137a646fa15119c1e4bf9ba0d26

Observation c7959343-5e0f-4470-ab35-d99baa8b485a · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.113708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.113708Z digest=sha256:df5bc34e8d603fbed51e03caa56f7db2895f313248b4aef15acd9ea80633a9ae

Observation 629086a4-0c76-4702-99f1-dcef23c9f511 · outbound

This paper cites Robocasa: Large-scale simulation of ev- eryday tasks for generalist robots,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Robocasa: Large-scale simulation of ev- eryday tasks for generalist robots,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.116533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.116533Z digest=sha256:d0ab31eaa4aa787845ded45e0a3439e3c3c365f21c3f46adfe23c1bb0e078bdd

Observation c68f25b1-6ea0-4947-a80c-0f1efd5fdb79 · outbound

This paper cites Legato: Cross-embodiment imitation using a grasping tool,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Legato: Cross-embodiment imitation using a grasping tool,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.118520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.118520Z digest=sha256:05c6e0d98f5284f7731b5c7c0c4baa431519e9107e80ea232030b9305c69cedb

Observation b139d678-ebc3-4869-83b6-b94fe3e0e069 · outbound

This paper cites Self-Distillation Bridges Distribution Gap in Language Model Fine-Tuning.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Self-Distillation Bridges Distribution Gap in Language Model Fine-Tuning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.121101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.121101Z digest=sha256:78adeb8af87fb2ac2d8f04f62a6db1284450941c7e546c65e8c33f53f4c8ea17

Observation 9c882efb-1139-4228-b6eb-e800c10490ce · outbound

This paper cites Does continual learning equally forget all pa- rameters?.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Does continual learning equally forget all pa- rameters?

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.123316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.123316Z digest=sha256:ec12f7455f040ec22bd919f5d595810a2bb62b6dbf251b9ac29747fae1c2161f

Observation eaca2272-ab9b-4419-9cba-fef227e19ff3 · outbound

This paper cites Segment anything,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Segment anything,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.125796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.125796Z digest=sha256:eef00c340cd73096425209dd586188bbc7e7ad35e6cec895a62f0885171f9b03

Observation 21fce94b-14a0-4ddf-a77e-23c89142ee12 · outbound

This paper cites Decoupling human and camera motion from videos in the wild,.

MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos Decoupling human and camera motion from videos in the wild,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T18:51:36.127656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:51:36.127656Z digest=sha256:f921cd4ea9eb66a179ba5869c2628a2c7cd10a368453943f84b020819492c86f

Pith citing papers

Observation 08184c78-6d50-4ee1-bff1-766e366bb783 · inbound

X-Diffusion: Training Diffusion Policies on Cross-Embodiment Human Demonstrations cites this paper.

X-Diffusion: Training Diffusion Policies on Cross-Embodiment Human Demonstrations MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:40:33.752011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-18T00:37:12.170711Z digest=sha256:dbc7d4f90609b88007d08b5b410ffd976715bcbeea340df4319d0f6d62fe804f

Observation 64b637e1-d756-4dcc-8b3d-8a34a03298fa · inbound

Bimanual Robot Manipulation via Multi-Agent In-Context Learning cites this paper.

Bimanual Robot Manipulation via Multi-Agent In-Context Learning MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T00:29:47.506260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T00:25:24.362191Z digest=sha256:4db26dce81087676ab05e845abf16d6e806658d40b34b2a68878b67e7706d284

Observation f1cd0c19-d6c0-4d0f-92a5-81be84f8848f · inbound

HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos cites this paper.

HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:45:48.712142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-30T01:02:58.035784Z digest=sha256:fea0711e44e22f701b47c46228f135501810e6e0da78453f2aebd03745b68d5a

Observation 46752ed2-8080-476f-9419-73bd2bdba9ab · inbound

Learning of Robot Safety Policies via Adversarial Synthetic Scenarios cites this paper.

Learning of Robot Safety Policies via Adversarial Synthetic Scenarios MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:36:58.971265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-28T01:12:25.956015Z digest=sha256:27a45aa206d6e18113390f052a05b8701e28e151dbcdec693393b9d280c0081c

Observation 9afd20f5-db4e-4edf-9feb-0023509c07bc · inbound

In-Context World Modeling for Robotic Control cites this paper.

In-Context World Modeling for Robotic Control MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:00:09.579213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:12:22.513577Z digest=sha256:97366b56a847c9c6bf84f7e922f611fa2de7b363dc1f6737d25ef28edcc2d919

Observation 9da99683-5dd7-469a-8365-3812b0941822 · inbound

In-Context World Modeling for Robotic Control cites this paper.

In-Context World Modeling for Robotic Control MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:29:51.621126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-26T05:11:07.089829Z digest=sha256:280993b5064200837e637d830fb98625f6e13be834bbc25baab7e6d3a62d8e0f

Observation 10c874f3-da37-429e-a71c-642627d80fd2 · inbound

In-Context World Modeling for Robotic Control cites this paper.

In-Context World Modeling for Robotic Control MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T12:05:57.682386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:05:57.682386Z digest=sha256:3e4aeb5632f8a49c14ca7699d44ec66682461f3b0a727a64da02c021297e0df3

Observation 9f163656-9cde-4d67-a8db-f05d86f1c0ee · inbound

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time cites this paper.

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-09T22:16:36.338374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-07-09T22:16:31.529359Z digest=sha256:dec11710801f6a9e90c638b9067d6cfc89094b3ef7f4b7f5163dd5fa11b0dc32

Observation bca86fbd-c713-43c5-b318-7bf3c3740363 · inbound

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time cites this paper.

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T06:48:14.554799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T06:48:14.554799Z digest=sha256:24873c4a32011287ac1c8d26083a7c9fc1b0602b3125bd2ac13aa925a7929c49