Pith. sign in

Paper Citation Record · LEDGER

ICLR: In-Context Imitation Learning with Visual Reasoning

As of 21 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 2 inbound Pith citation observations for arXiv:2603.07530.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.07530 v2

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T18:38:32.683746Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T19:23:45.402022Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T04:09:35.083408Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d80c348b-5981-4ee2-9662-c6063109bd82 · outbound

This paper cites Good old-fashioned engineering can close the 100,000- year “data gap.

ICLR: In-Context Imitation Learning with Visual Reasoning Good old-fashioned engineering can close the 100,000- year “data gap

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:27.085805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:27.085805Z digest=sha256:28c6662fa9114e6e01ee77e7a1a7ac98f7b685e1361f5a3ac20abaa1c3d15519

Observation 9ffcc999-6676-4d50-904b-21c27005fde1 · outbound

This paper cites Droid: A large-scale in-the-wild robot manipulation dataset,.

ICLR: In-Context Imitation Learning with Visual Reasoning Droid: A large-scale in-the-wild robot manipulation dataset,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:27.165280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:27.165280Z digest=sha256:eb4372ffc49ecc8feb16c462e855e038ef5eb770ddca615a346513a99d13f6a0

Observation 78b35951-be7a-44b8-9a07-68051b20f17f · outbound

This paper cites Bc-z: Zero-shot task generalization with robotic imitation learning,.

ICLR: In-Context Imitation Learning with Visual Reasoning Bc-z: Zero-shot task generalization with robotic imitation learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:27.311037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:27.311037Z digest=sha256:163af5cb33029cd8478d38c9b44605d254fc6c21bef942412ce0c983ade92928

Observation 3eac1579-ba3d-4dab-adca-b916b374d69d · outbound

This paper cites Rovi-aug: Robot and viewpoint augmentation for cross-embodiment robot learning,.

ICLR: In-Context Imitation Learning with Visual Reasoning Rovi-aug: Robot and viewpoint augmentation for cross-embodiment robot learning,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:27.425122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:27.425122Z digest=sha256:1681cc9f3024360918cc5059ffaf1e5a70422d6aa610eeb57b2e4ba3d5a7b980

Observation 4311fc8c-d289-46fb-bca1-cf379e6ebfb8 · outbound

This paper cites Icrt: In-context imitation learning via next-token prediction,.

ICLR: In-Context Imitation Learning with Visual Reasoning Icrt: In-context imitation learning via next-token prediction,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:27.565031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:27.565031Z digest=sha256:7aae593a57b9df0c9332d4694e2c51bd1e6a505f094922d7bb430a8cae0ccdff

Observation 96b0dea1-4da1-44df-911d-59bfaf99ef60 · outbound

This paper cites Novel demonstration generation with gaussian splatting enables ro- bust one-shot manipulation,.

ICLR: In-Context Imitation Learning with Visual Reasoning Novel demonstration generation with gaussian splatting enables ro- bust one-shot manipulation,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:27.813897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:27.813897Z digest=sha256:7848730f55380e2a7e2b476d7e526d65462e40eb496175bd5f203f21afc6abfd

Observation ade2ccc9-72c9-4867-8d39-a29f175c9d67 · outbound

This paper cites World Action Models are Zero-shot Policies.

ICLR: In-Context Imitation Learning with Visual Reasoning World Action Models are Zero-shot Policies

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:28.003247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:28.003247Z digest=sha256:c0cbf63ea629d84ce930b324807bcbe19041d5669c3447df0668cdb0b670cc73

Observation 8d127a3f-d17d-4ce2-b003-6b5fb5ccd416 · outbound

This paper cites A systematic study of data modalities and strategies for co-training large behavior models for robot manipulation,.

ICLR: In-Context Imitation Learning with Visual Reasoning A systematic study of data modalities and strategies for co-training large behavior models for robot manipulation,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:28.142854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:28.142854Z digest=sha256:6628bf575862cfd766c130196d412d29ea578fc4fd5e1c5566a91b4245fd11e7

Observation 13bc154e-1da5-4d4f-b7a2-15be5b94a6b6 · outbound

This paper cites Keypoint action tokens enable in-context imitation learning in robotics,.

ICLR: In-Context Imitation Learning with Visual Reasoning Keypoint action tokens enable in-context imitation learning in robotics,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:28.259751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:28.259751Z digest=sha256:b78b5ccc35907e90663cec494b820bb2e6e5c0c21b081bfdbdcacf9eafdddc9b

Observation cf2654fc-c237-4ef8-902b-82b9d9e9a1e9 · outbound

This paper cites Instant policy: In-context imitation learning via graph diffusion,.

ICLR: In-Context Imitation Learning with Visual Reasoning Instant policy: In-context imitation learning via graph diffusion,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:28.453787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:28.453787Z digest=sha256:5f5cc6928b05c5510ef8bbcbfee3a64418af311f7efa7fef9b9389890e42425b

Observation ba46dd64-f7b8-4409-8b67-e45bfb2cb271 · outbound

This paper cites Mimicdroid: In-context learning for humanoid manipulation from human play videos,.

ICLR: In-Context Imitation Learning with Visual Reasoning Mimicdroid: In-context learning for humanoid manipulation from human play videos,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:28.624749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:28.624749Z digest=sha256:ed627f6f9ba0468c2e3cfce1bb071e622342a0a04b60f18a66cce0367d389ef5

Observation 3c52c04e-4d7b-4e2f-9b8b-c0de83b56db8 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models,.

ICLR: In-Context Imitation Learning with Visual Reasoning Chain-of-thought prompting elicits reasoning in large language models,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:28.713775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:28.713775Z digest=sha256:bea081c6e686eaee655842479e886824a509fc1c1639c6e28ec6ea0080dd0792

Observation 19c3dd14-e035-4a5f-8076-4af38194a9f5 · outbound

This paper cites Ddcot: Duty-distinct chain-of-thought prompting for multimodal reasoning in language models,.

ICLR: In-Context Imitation Learning with Visual Reasoning Ddcot: Duty-distinct chain-of-thought prompting for multimodal reasoning in language models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:28.803936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:28.803936Z digest=sha256:3777d93a1cac3a1a7d2525e5d04507a38d21bcf782025c1a30135b18ce76f7af

Observation 3c96c0ec-34ef-4a44-9228-74ef9ac3487a · outbound

This paper cites Multimodal chain-of-thought reasoning in language models,.

ICLR: In-Context Imitation Learning with Visual Reasoning Multimodal chain-of-thought reasoning in language models,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:28.896572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:28.896572Z digest=sha256:2dd531f0f13e9cd35438d792f87589fe8190f7508ac7479798180a0a7a4a85ce

Observation b1bc496a-41a2-4659-a2ce-b2fc38baf200 · outbound

This paper cites Few-shot in-context imitation learning via implicit graph alignment,.

ICLR: In-Context Imitation Learning with Visual Reasoning Few-shot in-context imitation learning via implicit graph alignment,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:28.978623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:28.978623Z digest=sha256:2e903bd37bdc6537f792b9706097fc50516947e95ac3cc49c34608943d4e77f0

Observation 59ff90f1-920d-44c4-bfbe-05fa3a7a0a05 · outbound

This paper cites Vid2robot: End-to- end video-conditioned policy learning with cross-attention transform- ers,.

ICLR: In-Context Imitation Learning with Visual Reasoning Vid2robot: End-to- end video-conditioned policy learning with cross-attention transform- ers,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:29.104355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:29.104355Z digest=sha256:8ddaa5b54fe5f9c600b53dff34f05f6fb450a4d9e411a0e7df1a053bc4e6cdab

Observation bf90aeba-e38a-4c74-ab3a-963cabe41a06 · outbound

This paper cites Ricl: Adding in- context adaptability to pre-trained vision-language-action models,.

ICLR: In-Context Imitation Learning with Visual Reasoning Ricl: Adding in- context adaptability to pre-trained vision-language-action models,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:29.213358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:29.213358Z digest=sha256:e722ba601ad2059c500efb651a0e83ba37e36264120bba0d7fc780d8664541f1

Observation 4186bb43-a20f-42b5-950c-1035caa8c1be · outbound

This paper cites Thinkact: Vision-language-action reasoning via reinforced visual la- tent planning,.

ICLR: In-Context Imitation Learning with Visual Reasoning Thinkact: Vision-language-action reasoning via reinforced visual la- tent planning,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:29.363559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:29.363559Z digest=sha256:b4ab484dc8026e12fda119984604bee0ecbbbeaba9b180259e63804d78863b32

Observation f6151216-eac9-41d5-ae41-388eb3614db0 · outbound

This paper cites Robotic control via embodied chain-of-thought reasoning,.

ICLR: In-Context Imitation Learning with Visual Reasoning Robotic control via embodied chain-of-thought reasoning,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:29.495047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:29.495047Z digest=sha256:edcdf41473dc35c8703fd80fb4414fec146d35323f879184578e285f1e87ceeb

Observation c7819539-e49e-4782-a170-6ae5ce59e204 · outbound

This paper cites Cot-vla: Visual chain-of-thought reasoning for vision-language-action models,.

ICLR: In-Context Imitation Learning with Visual Reasoning Cot-vla: Visual chain-of-thought reasoning for vision-language-action models,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:29.587167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:29.587167Z digest=sha256:712d08b3cc2c2e80187af4056a706f1af040c338b7d7d5634af30f85d27896fa

Observation 9311bbc7-0518-4675-8896-261d9659116f · outbound

This paper cites Molmoact: Action reasoning models that can reason in space,.

ICLR: In-Context Imitation Learning with Visual Reasoning Molmoact: Action reasoning models that can reason in space,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:29.742935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:29.742935Z digest=sha256:dc467978f7eb7697d104999a9af776548fc41a60903331583695cbf62b587d2b

Observation dcafd521-371d-4187-9fd5-0c9125342106 · outbound

This paper cites Gemini Robotics: Bringing AI into the Physical World.

ICLR: In-Context Imitation Learning with Visual Reasoning Gemini Robotics: Bringing AI into the Physical World

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:29.811396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:29.811396Z digest=sha256:28d86fcb975e2444de2564797311b59875b13bfc8ff7f5b2bf868597621b34d4

Observation 405bbe2e-f41b-47b2-914d-1a1945b1ee92 · outbound

This paper cites pi0.5: a vision-language-action model with open-world generalization,.

ICLR: In-Context Imitation Learning with Visual Reasoning pi0.5: a vision-language-action model with open-world generalization,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:29.910279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:29.910279Z digest=sha256:0234a8ccba73d691c664329a1152c05cf72714bddc96d104428b3e1bf98ebb1b

Observation 42e832b9-6ad8-47c6-8b4b-62824c7af6d5 · outbound

This paper cites Hamster: Hierarchical action models for open-world robot manipulation,.

ICLR: In-Context Imitation Learning with Visual Reasoning Hamster: Hierarchical action models for open-world robot manipulation,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:30.061677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:30.061677Z digest=sha256:3e4b01b70be583d7fcbb8bff59efce942a2544a4c97f4aef89b99e8e17d1d502

Observation 48b0227c-cf2a-45a6-afa7-9bf659dd65b6 · outbound

This paper cites Coa-vla: Improving vision-language-action models via visual-text chain-of-affordance,.

ICLR: In-Context Imitation Learning with Visual Reasoning Coa-vla: Improving vision-language-action models via visual-text chain-of-affordance,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:30.177967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:30.177967Z digest=sha256:b439138283cd1409a6f1af432342e229402421dd06c747db3da4c2b63b7a3a8e

Observation 056963f4-df48-44f5-90b1-5f54ae9ea9e4 · outbound

This paper cites Peek: Guiding and minimal image representations for zero-shot generalization of robot manipulation policies,.

ICLR: In-Context Imitation Learning with Visual Reasoning Peek: Guiding and minimal image representations for zero-shot generalization of robot manipulation policies,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:30.335213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:30.335213Z digest=sha256:8279ac2a9eb7905a1ea27cb5329294fd7606207e9add3c806d4e8d0c1e43373f

Observation 5a8ef74c-9ec3-484f-a600-36815b438e8e · outbound

This paper cites Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding.

ICLR: In-Context Imitation Learning with Visual Reasoning Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:30.443103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:30.443103Z digest=sha256:fde5d0b6a417155b0484bf524a5afde5fdecd9211b2e9cd23629778b2ce43f6b

Observation 8562aa9d-4d8f-40f9-b2bf-b0b8b38853a6 · outbound

This paper cites Gemini Robotics 1.5: Pushing the Frontier of Generalist Robots with Advanced Embodied Reasoning, Thinking, and Motion Transfer.

ICLR: In-Context Imitation Learning with Visual Reasoning Gemini Robotics 1.5: Pushing the Frontier of Generalist Robots with Advanced Embodied Reasoning, Thinking, and Motion Transfer

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:30.577678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:30.577678Z digest=sha256:5a8f5151d34d51309b699afd348a7166154723035b57cab50dd8a34c3446f6b2

Observation 2a75cda4-351e-4679-bacf-f1388c6c5b10 · outbound

This paper cites Qwen3-VL Technical Report.

ICLR: In-Context Imitation Learning with Visual Reasoning Qwen3-VL Technical Report

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:30.711742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:30.711742Z digest=sha256:30cc2d08a0159b05fea2b476b7dd06c26453ab172aad98eebc275ad6124f8806

Observation 754b6d9c-30e8-473d-ae3c-02e53bdbea57 · outbound

This paper cites Sam 2: Segment anything in images and videos,.

ICLR: In-Context Imitation Learning with Visual Reasoning Sam 2: Segment anything in images and videos,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:30.815742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:30.815742Z digest=sha256:e147ef8af08d20b3ee09308ddeadd8da019c94d64ca0ba0f24b9dd6eca608543

Observation 333daf4f-5ecc-4c47-87cb-02c260040d20 · outbound

This paper cites DINOv3.

ICLR: In-Context Imitation Learning with Visual Reasoning DINOv3

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:30.929430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:30.929430Z digest=sha256:04c83f14708f13b8dde300a989f1967d24d8bf6c17bac1b996d34bf6239ac347

Observation 8c85f343-f28f-4e95-89a7-cbc00d6d3ad1 · outbound

This paper cites End-to-end object detection with transformers,.

ICLR: In-Context Imitation Learning with Visual Reasoning End-to-end object detection with transformers,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:31.070097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:31.070097Z digest=sha256:f8298147d20f780103e757f15c7eac2515ea79026ca28efb68d98360b1692141

Observation de69b2ad-38eb-4722-8a3b-4b62cbe9bdea · outbound

This paper cites Scaling open-vocabulary object detection,.

ICLR: In-Context Imitation Learning with Visual Reasoning Scaling open-vocabulary object detection,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:31.243983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:31.243983Z digest=sha256:6fd4613e155f952dffb7c9ec623e0c7986efd88a5f50af26805144e8da5d66ad

Observation 5b4df95b-2733-4fd6-8e23-e6a0648451eb · outbound

This paper cites Hand me the data: Fast robot adaptation via hand path retrieval,.

ICLR: In-Context Imitation Learning with Visual Reasoning Hand me the data: Fast robot adaptation via hand path retrieval,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:31.358590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:31.358590Z digest=sha256:90e78b0bc8bb3f77c246a52f469770fa2434f86f39f5fcf36b75a370094d59df

Observation 73b1832e-b6f5-403c-b3ea-3b81b066e968 · outbound

This paper cites Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control.

ICLR: In-Context Imitation Learning with Visual Reasoning Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:31.486289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:31.486289Z digest=sha256:4e66c75cebd64cb19cd01bb7219fe061c5330704c8a8c92419b5e2258c0f85fc

Observation 7e9376bf-7e01-455f-81a2-ea50aa33768e · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

ICLR: In-Context Imitation Learning with Visual Reasoning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:31.571799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:31.571799Z digest=sha256:2c9e41e9b30b78e05b2dc20b758f3b652716f4d08099df29ebc1c50cac487bdb

Observation 9fa11f34-4612-4028-998d-fb43d758e218 · outbound

This paper cites Set transformer: A framework for attention-based permutation-invariant neural networks,.

ICLR: In-Context Imitation Learning with Visual Reasoning Set transformer: A framework for attention-based permutation-invariant neural networks,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:31.747133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:31.747133Z digest=sha256:1ee5527adea4bfd3432cba3b2a2d3d98652cf0422b9559cd052da9145dbf08a9

Observation 599ea514-6f5a-4bf0-91c8-8d45cbc0f6b9 · outbound

This paper cites Learning fine-grained bimanual manipulation with low-cost hardware,.

ICLR: In-Context Imitation Learning with Visual Reasoning Learning fine-grained bimanual manipulation with low-cost hardware,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:31.927027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:31.927027Z digest=sha256:b72973c9e932c9c32290f4611091d1fc27fff43615b8753c98484fb321c21492

Observation 13b1dcc8-dbb7-41a9-86d7-f4d507547a59 · outbound

This paper cites Training strategies for efficient embodied reasoning,.

ICLR: In-Context Imitation Learning with Visual Reasoning Training strategies for efficient embodied reasoning,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:32.022643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:32.022643Z digest=sha256:e856c79f5cf13dfe48ce93805a92278022b0ebf4f2c89509c4988a6446f5a7ff

Observation 91ddcbc1-cfbd-400b-8665-2f125be4b877 · outbound

This paper cites Libero: Benchmarking knowledge transfer for lifelong robot learn- ing,.

ICLR: In-Context Imitation Learning with Visual Reasoning Libero: Benchmarking knowledge transfer for lifelong robot learn- ing,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:32.112890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:32.112890Z digest=sha256:5c48b9db2c65ca0b8c530d7d47fff67117be8a76a10458637e27d0acafe92f99

Observation 21aa1972-2f98-4fa4-83ec-bc111325ed86 · outbound

This paper cites Openvla: An open-source vision-language-action model,.

ICLR: In-Context Imitation Learning with Visual Reasoning Openvla: An open-source vision-language-action model,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:32.228681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:32.228681Z digest=sha256:ed801b4f6b608b9628898af95b64455443eb59aa4f8ee6de740257ba56398d77

Observation 6ad5e4be-21bd-457a-86cc-e18f711bd331 · outbound

This paper cites Universal manipulation interface: In-the-wild robot teaching without in-the-wild robots,.

ICLR: In-Context Imitation Learning with Visual Reasoning Universal manipulation interface: In-the-wild robot teaching without in-the-wild robots,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:32.354144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:32.354144Z digest=sha256:9209f7c5091afcc967609ca45d90bc0c482cecb4511a6c1d11ff10112187692f

Observation c8977d95-6cc2-40b2-ab05-5793e07c9826 · outbound

This paper cites Gello: A general, low- cost, and intuitive teleoperation framework for robot manipulators,.

ICLR: In-Context Imitation Learning with Visual Reasoning Gello: A general, low- cost, and intuitive teleoperation framework for robot manipulators,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:32.486598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:32.486598Z digest=sha256:f743d3bf3503407e67f1a0d86cedcce4e31f4810c5dce38944355e3f0ef33db0

Observation 28197a0e-f769-4046-a997-65462d5e87b6 · outbound

This paper cites Lmact: A benchmark for in-context imitation learning with long multimodal demonstrations,.

ICLR: In-Context Imitation Learning with Visual Reasoning Lmact: A benchmark for in-context imitation learning with long multimodal demonstrations,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:32.598356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:32.598356Z digest=sha256:3158ddcb635338a012a877a165eac2eb8791a9b7afdcb7dce4e70727e8cf039a

Observation 0408feb0-c3c0-449c-bc0a-bcf73051b30e · outbound

This paper cites Language models are few-shot learners,.

ICLR: In-Context Imitation Learning with Visual Reasoning Language models are few-shot learners,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T18:38:32.683746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:38:32.683746Z digest=sha256:6026d7710f43334d4f7e4fa4603f1634a1845bff4e82426e9d776a4d69afe946

Pith citing papers

Observation 6dd66034-f8ad-4778-9593-c0c2b758e45a · inbound

SynthICL: Scalable In-context Imitation Learning with Synthetic Data cites this paper.

SynthICL: Scalable In-context Imitation Learning with Synthetic Data ICLR: In-Context Imitation Learning with Visual Reasoning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-21T01:19:54.117323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T19:23:45.402022Z digest=sha256:525d52881134ac8c73086c8e86ebde8c6dfedcaaf3c4436fe59d24f8407be327

Observation 9c2511c8-82be-4f77-9a93-24c4d2d67409 · inbound

World Action Models: A Survey cites this paper.

World Action Models: A Survey ICLR: In-Context Imitation Learning with Visual Reasoning

Reference 124

Resolution
verified exact
arxiv_id, observed 2026-07-21T01:19:54.117323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T17:11:12.686936Z digest=sha256:f1ac9fbf901915dd121791285d78bf2596de1c039a8a57887e398adc41f5a0a1