Pith. sign in

Paper Citation Record · LEDGER

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment

As of 6 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 1 inbound Pith citation observation for arXiv:2604.10677.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.10677 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T15:36:23.197843Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:54:20.104031Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact7
  • verified fuzzy43
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3941fef2-64f0-4bf0-998a-d21f116c088b · outbound

This paper cites Open x-embodiment: Robotic learning datasets and rt-x models.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Open x-embodiment: Robotic learning datasets and rt-x models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.480489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:3c8e9b79a6039c91c1e5cc309f7a2a96fd4c6c04fe086bd60bc52552e9ec7bc5

Observation 792b1e76-3bab-4a37-bf0e-e6655d513d82 · outbound

This paper cites Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:11:03.890345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:f09f104304902eedd322afb33c61b832bbfa285a927e3fa3e86c28dadd096af4

Observation f755dbfb-8825-42cf-b91a-7be92c29781b · outbound

This paper cites Airexo-2: Scaling up generalizable robotic imitation learning with low-cost exoskeletons.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Airexo-2: Scaling up generalizable robotic imitation learning with low-cost exoskeletons

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.512216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:bf0ef63e6800f61ba895d1fea23160ea0b2742c9baa8dfa94e822f9b6e2d6ffd

Observation 5fe3b5b5-122b-44a9-8128-6e253026f817 · outbound

This paper cites Data scaling laws in imitation learning for robotic manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Data scaling laws in imitation learning for robotic manipulation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.528984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:9adc3943cd9209110b4b8ec29d4cbdfad8b2c24762dea933de7c35032f98f93d

Observation d9980239-9177-44d7-8fb3-20439cd7e729 · outbound

This paper cites π 0: A vision-language-action flow model for general robot control.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment π 0: A vision-language-action flow model for general robot control

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.518325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:3a64d4b76279267951ec3e28b6702d774d05b15235d7550ccbef6802988a77c1

Observation a7850250-86f2-4d2b-9f43-fd4148ace0c9 · outbound

This paper cites Rh20t: A comprehensive robotic dataset for learning diverse skills in one-shot.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Rh20t: A comprehensive robotic dataset for learning diverse skills in one-shot

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.546019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:6cb6923c7c0eab431a2b1f99be627215439e7ac658dd0ad664aabbe3e26c441f

Observation cb228280-0514-4add-83f1-c7a4f20ea4af · outbound

This paper cites Robomind: Benchmark on multi-embodiment intelli- gence normative data for robot manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Robomind: Benchmark on multi-embodiment intelli- gence normative data for robot manipulation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.506554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:871c9fbd4ae62b49e21db599b27569e415d0e1b98dea0ebeb28435ffd4981cf3

Observation 94be8688-51ee-4e70-b8b6-688e3593ae44 · outbound

This paper cites Droid: A large-scale in-the-wild robot manip- ulation dataset.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Droid: A large-scale in-the-wild robot manip- ulation dataset

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.485256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:85195d25704a2a6585f6362aaef9e1dac340edbaf985146ed613145ec36988e6

Observation 2bf4a8b2-11fa-4748-8ee9-c7a1ab4de918 · outbound

This paper cites Taco: Benchmarking generalizable bimanual tool- action-object understanding.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Taco: Benchmarking generalizable bimanual tool- action-object understanding

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.497184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:04826f567dbefb4bdcc0f0a4dece18409d1f5be079f95fcc672222c20ae28810

Observation 4304229c-8689-4f65-8258-96579398830d · outbound

This paper cites Oakink2: A dataset of bimanual hands-object manipulation in complex task completion.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Oakink2: A dataset of bimanual hands-object manipulation in complex task completion

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.475838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:2f06cf2d5daf3598e039615738abca3ecc7621030c16a9352b4ad7f09db7d1e9

Observation 0171e4f3-3380-430e-b78a-ffa8205dccf9 · outbound

This paper cites OakInk: A large-scale knowledge repository for understanding hand-object interaction.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment OakInk: A large-scale knowledge repository for understanding hand-object interaction

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.522463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:653477f5088c6064f26b6cb8438a082969404494cf48189f3abe281464df8225

Observation 11eeed79-4e1b-413c-8e14-6bc0794f5365 · outbound

This paper cites DexYCB: A benchmark for capturing hand grasping of objects.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment DexYCB: A benchmark for capturing hand grasping of objects

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.449811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:43eae40d02c9671c5a4de57dfdc1eb29de560a2b29ef043953b718e7f3b2c669

Observation 8f8b08b9-855b-4629-8805-9cf77a4aa6d6 · outbound

This paper cites Ego4d: Around the world in 3,000 hours of egocentric video.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Ego4d: Around the world in 3,000 hours of egocentric video

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.514877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:46dc21e143ba0050d0ddb8d4808d37ab017a4e75384d2bfbff3bf0065c60bd7c

Observation a9a8b939-9bf8-483a-b04d-bdb494dfd229 · outbound

This paper cites H2r: A human-to-robot data augmentation for robot pre- training from videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment H2r: A human-to-robot data augmentation for robot pre- training from videos

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:11:03.901752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:2d2db5f80096a6f0fdecf969e132b3ea2a75ee513a20e145cb7e1f84504eb482

Observation 416a1d5b-d243-4075-ad91-716ec8917329 · outbound

This paper cites Masquerade: Learning from In-the-wild Human Videos using Data-Editing.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Masquerade: Learning from In-the-wild Human Videos using Data-Editing

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-29T02:04:55.330788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:119a0b4ffa7c6af484f8787610355f9f729ef3388f0fb0228909a80a34925dac

Observation 512a1e3b-bfe6-46eb-81da-fa7bf6f290e4 · outbound

This paper cites Phantom: Training robots without robots using only human videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Phantom: Training robots without robots using only human videos

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.500107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:fa4057d37a472d6f810c3b61859e84d02a77cb346dce6f70e38c8211acf64bca

Observation a850c922-64b2-4363-9a20-9a86d2bb9819 · outbound

This paper cites AR2-D2: training a robot without a robot.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment AR2-D2: training a robot without a robot

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.542656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:1f5119bd04f66db6d9bd33a8d2aaaa2f5504bc2c636057f76dc2254442c6da00

Observation 5d0db213-e176-4815-95a0-571f29016792 · outbound

This paper cites Emergence of human to robot transfer in vision-language-action models.arXiv preprint arXiv:2512.22414.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Emergence of human to robot transfer in vision-language-action models.arXiv preprint arXiv:2512.22414

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:11:03.910466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:5ad1f9609e59215f491e07c400675aba97a02be6fa6f8ef44c90568266571882

Observation 2aee3763-1a2e-42f2-9c15-65bfa3df68b8 · outbound

This paper cites Egomimic: Scaling imitation learning via egocentric video.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Egomimic: Scaling imitation learning via egocentric video

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.478992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:dc355b4e5fa647e2a06f644bd7bb4bae1952076cb379c7b773bfa7dba5d2aec1

Observation 256378bd-56c2-4e95-b54e-0812c74afaee · outbound

This paper cites Humanoid policy ˜ human policy.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Humanoid policy ˜ human policy

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.443304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:c4724b982689a08b63b3122340f7f413a208acc35dcfbdd9b7e65a4fc89f4b7f

Observation 2f849ef2-1c7a-4974-be45-c6cf61e388d2 · outbound

This paper cites EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:32:58.962571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:262f901d8c6ac1350ee093bea4bc2fe16dadf5d50acd5e14aaade923ff65db90

Observation d29a1cae-e256-4c61-bddd-c1cc9eec44c9 · outbound

This paper cites Univla: Learning to act anywhere with task-centric latent actions.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Univla: Learning to act anywhere with task-centric latent actions

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.456459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:a791d3a3bf08676e1f5e9cdc11b9b0f21a353da04f4323135591ebf7fb2b4e1c

Observation afa1b9c1-131f-4115-b3a1-a9e909790b3f · outbound

This paper cites Latent action pretraining from videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Latent action pretraining from videos

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.494175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:0c7ba25c71c5bb2d1b19a8913f6c6e23b60addf7464983e3a9d3fe7b8acc27b9

Observation 29ffa3af-ec0c-4ecb-82ea-0e6b8f377755 · outbound

This paper cites Moto: Latent motion token as the bridging language for learning robot manipulation from videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Moto: Latent motion token as the bridging language for learning robot manipulation from videos

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.531756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:7c7d9656ed4b5f79ee697e01d27048ff724bc3edeaa7c58eaec10d6b1dcc1783

Observation 39e11551-b418-45cd-a60d-932d263b3fd9 · outbound

This paper cites Mimicplay: Long-horizon imitation learning by watching human play.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Mimicplay: Long-horizon imitation learning by watching human play

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.549326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:d77c848f49b0a247c70936376ee9336f3a64ea22ac86eaa1a1da0ab269f1d2ec

Observation 4a09ee17-e49a-4193-95b5-a16e33e2b548 · outbound

This paper cites ViViDex: Learning vision-based dexterous manipulation from human videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment ViViDex: Learning vision-based dexterous manipulation from human videos

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.574822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:0acba8691b4a9b8b1e7ba6bd14b2046ec8f84e364725a9a1971d242aae2d8177

Observation 3b079b06-8263-496f-86de-7eaa3a251f79 · outbound

This paper cites Vidbot: Learning generalizable 3d actions from in-the-wild 2d human videos for zero-shot robotic manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Vidbot: Learning generalizable 3d actions from in-the-wild 2d human videos for zero-shot robotic manipulation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.463957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:f1d4184ecf97383ce771c15b5da4093738ac60532ff1a8ab544ba281caf97cb0

Observation 568091b5-4891-4ec5-8f4b-3b8733d678f2 · outbound

This paper cites Zeromimic: Distilling robotic manipulation skills from web videos.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Zeromimic: Distilling robotic manipulation skills from web videos

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.459542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:06561091e7043411faca33607a975b0e389d8724f1d21fadff1139ad586d1eca

Observation 8f02b45c-42bb-496e-a423-efab5307458a · outbound

This paper cites Affordances from human videos as a versatile representation for robotics.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Affordances from human videos as a versatile representation for robotics

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.567498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:c392c279b1895eaf89a370e0c93fb2a64c8828d8df4975e1be56608a693bae3e

Observation 80f1abf7-ec1b-4e46-a41d-9742046cee8c · outbound

This paper cites DINOv3.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment DINOv3

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T10:11:03.937349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:7b8e4adc880dfafaff4d9c593a264dcfa5dc87084067e89f6d0470dc270e7bf0

Observation 56d46e99-0e1a-42bc-81f9-3f62040197bf · outbound

This paper cites X-diffusion: Training diffusion policies on cross- embodiment human demonstrations.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment X-diffusion: Training diffusion policies on cross- embodiment human demonstrations

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.477914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:53f3458788de39ec8c16a44af25ae626fc4312d1948a99b6a06b1bb452cd26f5

Observation 6669ea4f-a141-4ff5-8fcc-876a461a2e87 · outbound

This paper cites Deep imitation learning for complex manipulation tasks from virtual reality teleoperation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Deep imitation learning for complex manipulation tasks from virtual reality teleoperation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.439811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:03b9040b078ea31e49f14f9dee5a6388c51441a771b6fe6820dcdc35c579c091

Observation 45ace35f-cb12-457e-840d-f2c60307588d · outbound

This paper cites The surprising effectiveness of representation learning for visual imitation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment The surprising effectiveness of representation learning for visual imitation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.453909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:8e052740eb455cd939f561e2d82899a0a7d723dc9ffffa9d137d8840497e55ff

Observation 7da78a45-f978-4693-aca9-0fd5eac66617 · outbound

This paper cites R3M: A universal visual representation for robot manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment R3M: A universal visual representation for robot manipulation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.552058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:19548bab209f44d978d5f7bb0989270a7f44dd7686d08890b53567087bd8df01

Observation 4472963e-0834-42ba-8379-8c938de6694e · outbound

This paper cites LIV: language-image representations and rewards for robotic control.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment LIV: language-image representations and rewards for robotic control

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.482309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:8d576cd52c5bed3cab19c21d904d520fddef173fced72ce5becca93080e60231

Observation a0538adc-f5b6-46dd-87c3-663021737fe6 · outbound

This paper cites VIP: towards universal visual reward and representation via value-implicit pre-training.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment VIP: towards universal visual reward and representation via value-implicit pre-training

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.557468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:d85525fb1a6973a415550eb63926fe982a0167f97e2a8901d59ddb639f0fb0a4

Observation 1f353144-343e-4793-8e93-6a3ff7102fe8 · outbound

This paper cites Real-world robot learning with masked visual pre-training.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Real-world robot learning with masked visual pre-training

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.564248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:f2568a0528814b8eafda7cda4eb9dd8221bef8a67000785c5e21d2f95cc554b1

Observation a9ed8a37-057d-4d90-9249-914e5b34661d · outbound

This paper cites Learning transferable visual models from natural language supervision.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Learning transferable visual models from natural language supervision

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.538428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:f3c148828b9f98da04f1a302857f46f85f3cfe8cc5f4985b939850e60c853675

Observation 1c059970-cac1-4566-865a-7de8613e92e3 · outbound

This paper cites Dinov2: Learning robust visual features without supervision.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Dinov2: Learning robust visual features without supervision

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.560976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:9283b91e1eeefeeb70922f3acafcf54ba932c280e44ab8687c0e26f831366de2

Observation bf77b21c-ffbf-4a65-a6d0-e78aa8beff10 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment High-resolution image synthesis with latent diffusion models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.488396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:6ff5b599b8516cccc077109f39b8d89d61659ff6008ac604b3f2f21c01b3ae50

Observation 09a69b01-ceae-4a76-8a2a-7ea62db034bb · outbound

This paper cites Cage: Causal attention enables data-efficient generalizable robotic manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Cage: Causal attention enables data-efficient generalizable robotic manipulation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.472057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:557ef848c318bcf43a5d9ef68f4500f38eb5a70b20592f6c01075d129054e486

Observation b9c69c93-7e6b-4b96-8598-eb8a880bf777 · outbound

This paper cites Theia: Distilling diverse vision foundation models for robot learning.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Theia: Distilling diverse vision foundation models for robot learning

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.491140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:674bbcd500d9e934005337dd849508b1d03f86c7c27e989a73f57aff67bd86d1

Observation a5fcab8a-e620-4248-ab31-a0e9fc3df029 · outbound

This paper cites Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-09T03:07:11.470849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:6824158f260a88768c3420f94beaa11091da67210303f49ca65bd5182b3159d3

Observation d94e07cb-f2ed-4d53-a322-e97e882a07c1 · outbound

This paper cites Gnfactor: Multi-task real robot learning with general- izable neural feature fields.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Gnfactor: Multi-task real robot learning with general- izable neural feature fields

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.509537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:705b6a29a44fd037eb72ffc15f127be08fc87a39b378ad81cdce228c7f626ab4

Observation 55c47b8f-f327-4d82-8dbd-4ff6f7d979c1 · outbound

This paper cites SAM-E: leveraging visual foundation model with sequence imitation for embodied manipulation.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment SAM-E: leveraging visual foundation model with sequence imitation for embodied manipulation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.453158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:8b7e5bf4f803ac70a4e94f9366e5bce622177b3981f7c096c176cb1681a029f5

Observation 353280cd-a49c-4955-a82b-72ccf2ed27ba · outbound

This paper cites Spawnnet: Learning generalizable visuomotor skills from pre-trained network.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Spawnnet: Learning generalizable visuomotor skills from pre-trained network

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.571372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:64101dd20addb17cb3e6b86a31d9380542e095ed7bdbdeb6e9291578fe048ac1

Observation 8ac2a8b8-7954-4970-899a-8655c09f5502 · outbound

This paper cites Recasting generic pretrained vision transformers as object-centric scene encoders for manipulation policies.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Recasting generic pretrained vision transformers as object-centric scene encoders for manipulation policies

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.462435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:7dc7a22f2017c5b3367fadb9326b0083af5bd8fcac8983d3502bfc836903d6e3

Observation f283615d-bf94-49ff-9622-d4e4bbf12879 · outbound

This paper cites Emerging properties in self-supervised vision transformers.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Emerging properties in self-supervised vision transformers

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.526227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:17aa9a2ab0fcdb4b7fd0f0d65441e0f4196568d20eb20d81dc649469be9aeb36

Observation 7d6bc277-e28e-48f5-b72b-4e0f35ddae03 · outbound

This paper cites Propainter: Improving propagation and transformer for video inpainting.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Propainter: Improving propagation and transformer for video inpainting

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.535049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:e5bc4d004507b76664fae9ffa3cebb3a5423825287dc6068bb558b66eaafeb72

Observation b5284360-4acd-42ed-ae5c-d73e26a84faf · outbound

This paper cites Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:11:03.882518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:ffd155669e4cfa1c2cd2b47e8719ba188837d468cd1e69d6d6c58926c256ef56

Observation 9d340648-6b4d-4af0-b528-e43b0d83f2e9 · outbound

This paper cites Multi-view hand reconstruction with a point- embedded transformer.

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment Multi-view hand reconstruction with a point- embedded transformer

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T19:42:05.503815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:36:23.197843Z digest=sha256:593d0f8cd5b5318a714ee8f873b768cec480e3d6188f4e5d287cbc99027b9a83

Pith citing papers

Observation 67af477c-5b1b-4692-b6b5-3501432580d5 · inbound

EgoRecovery: Acquiring Failure Recovery Ability Through Human Recovery Demonstration cites this paper.

EgoRecovery: Acquiring Failure Recovery Ability Through Human Recovery Demonstration LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T11:54:20.104031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:54:20.104031Z digest=sha256:cc637ec0df00e2e53c4cd73d7916c34ae69b1c6a2d3d33db39f9aa510c45692d