Pith. sign in

Paper Citation Record · LEDGER

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models

As of 9 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2606.19965.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.19965 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-26T18:37:27.171215Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch10

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 00a5e44c-a6f2-4af0-a8dd-0203bca8404f · outbound

This paper cites Scaling Learning Algorithms Towards.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Scaling Learning Algorithms Towards

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:0c66f1f79d196df4ab05a713a593fbd1af88f01040cc0f24349aac44847b76dd

Observation f489f4eb-1908-4b17-bbba-6e2aa937d5d8 · outbound

This paper cites and Osindero, Simon and Teh, Yee Whye , journal =.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models and Osindero, Simon and Teh, Yee Whye , journal =

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:8696d1c0316284460a4f7f91ea71fe164c3ddf1c0188bea6ccd55e04012b7dca

Observation 48484566-af6c-421b-93f7-942935b969e0 · outbound

This paper cites 2016 , publisher=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models 2016 , publisher=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:86505a30fa574819a1984fdf4db441788abed142e713ff2d3d73c01e38174d7d

Observation 13ae1fbe-5a4f-42d5-822e-9d67d7944e0d · outbound

This paper cites GPT-4 Technical Report.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models GPT-4 Technical Report

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T02:59:26.142060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:f819721041d24b42145142296876a55aeeff8deabb0298a63ac7065d533c6cd7

Observation 6cd26660-e967-4366-b90d-212dbb007261 · outbound

This paper cites GPT-4o System Card.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models GPT-4o System Card

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T02:59:26.139173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:e299d6158bf61501fe133c8d8ad2985a07525f3b617714cc7efb346b716ee101

Observation 6aeabeec-95e1-4f19-95ca-297c0d43249c · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Gemini: A Family of Highly Capable Multimodal Models

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T02:59:26.148947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:be1033cec13d0bff2ffd43d58682cd0b4bb884f93a5949f24d1447bcb903d7b5

Observation 541f6e75-b5a7-4eca-acfe-a76fdf005bbc · outbound

This paper cites Qwen3-VL Technical Report.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Qwen3-VL Technical Report

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T02:59:26.146079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:11558f78fe1647a9da435b1ea3db2d4e4e750be8f9cb91e4c87a9c06994a70a3

Observation 33edcc4f-f5ed-492c-b6b8-ae8bfcfae151 · outbound

This paper cites Qwen3.5-Omni Technical Report.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Qwen3.5-Omni Technical Report

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T02:59:26.115155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:9acc5b69da300d9c726407542a9df2a95230764121717ae2afc8eca7eff0d245

Observation 99a99f63-bb8a-4d72-954e-203c5251c624 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:61ee515ab57fc7b1f8568ac1cdd004500928d0bcb672a386697730e66f51c2d0

Observation 37c1dbc0-d586-40b6-80ba-95d3792ed1d1 · outbound

This paper cites Ref-Adv: Exploring.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Ref-Adv: Exploring

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:7e624f850a896a365c64d27751af22a2604623487892fc196d2fdf44e64eb31d

Observation 459e4827-4e11-4608-9914-519eddf067c4 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:866f51e1bfd1af71147466983cafd1f4c5f936d21be5e9526f47ec03f8c414f9

Observation 16232f52-2aeb-4a6e-847d-087b8e54fefd · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:f39156c6ba93535068ce8814d06705387af00df163658dda6e5f93c1e795aca1

Observation 09d3fe6c-c8e3-424a-a7d6-fe6156626684 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:c67a3edfaf44e379acb4a8bf8572faff58724f6729e80d98f27a5e340a1736c3

Observation 14324db1-861b-4814-b1f6-71ab12fffc2d · outbound

This paper cites arXiv preprint arXiv:2602.10138 , year=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models arXiv preprint arXiv:2602.10138 , year=

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:59:26.152642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:9ad16d4d6349dd32dbce102a55177760eafcc8eec37c05dc3376d1c60d2b0adb

Observation beea45e5-9950-4604-85ab-ecef0a1c5104 · outbound

This paper cites Sciegqa: A dataset for scientific evidence-grounded question answering and reasoning.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Sciegqa: A dataset for scientific evidence-grounded question answering and reasoning

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T02:59:26.119016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:e37cbc4be6cabe6b5851873b9cb653f80c60bd14817a1990d7fc8a25a0a82f63

Observation 54dd0633-c3b9-4643-987f-21dab3b0b556 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:e79c2fc1ba7834ce6642704d5d33f636e8a61c51c64e7fcca92252efb8fcae16

Observation 6ef5f344-8e23-48b5-90ab-ccb01d311514 · outbound

This paper cites The Fourteenth International Conference on Learning Representations , year=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models The Fourteenth International Conference on Learning Representations , year=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:c45555946b1f5aa57be6cae473ac558b7899c5ffcf8a3c917d11e8d9d61758a6

Observation 0374ffd3-b362-4d24-a011-013580d8b215 · outbound

This paper cites an unresolved cited work.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:79b91f3a31324c4f06a683a991509739a393e0872c1c3fd50e5b06bf64baef2e

Observation 66259507-f9a2-4ee7-97d9-1f253f60a7b3 · outbound

This paper cites arXiv preprint arXiv:2602.12196 , year=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models arXiv preprint arXiv:2602.12196 , year=

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:59:26.136100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:8adfe989abfa85247567e6a5106a67bae94471dd535d7b3944fe206fb7e2cd26

Observation e0cf176e-6dfe-4f12-9e58-40c373e7ba5c · outbound

This paper cites Rynnec: Bringing mllms into embodied world.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Rynnec: Bringing mllms into embodied world

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T02:59:26.132963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:f11a37e0145da297c02d8a99dbac03aaed0d9dc9ee9682ed48f4c18253b24b84

Observation d7c0fdce-c9be-4260-a415-ab877f2325f3 · outbound

This paper cites Rynnbrain: Open embodied foundation models.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Rynnbrain: Open embodied foundation models

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T02:59:26.125037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:ea7fedfec0b22606fc3dfd8a70265673787705c2e7af02caa60d98d14dfe3831

Observation 7d0539b8-8054-4c36-ab9f-a08e8015c0b4 · outbound

This paper cites Proceedings of The 8th Conference on Robot Learning , pages =.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Proceedings of The 8th Conference on Robot Learning , pages =

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:f8c3a5d4867c89a238feb0864e3a5f3457f61f0922294a9a24f5b572155d2a50

Observation 75c5e932-c306-4d84-b1a6-a3d298fbe0ff · outbound

This paper cites 2026 , url=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models 2026 , url=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:bcaa15c36cb011a81146cc2e8e470c9173614bbed28fab3134b10f819a9a6334

Observation d8a293ce-3b10-4b86-b8ae-688252c7c2af · outbound

This paper cites The Fourteenth International Conference on Learning Representations , year=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models The Fourteenth International Conference on Learning Representations , year=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:d4d183ae047828263762f6a5aa14848a50df2502b07d8584f5c5d05483690353

Observation 0995f466-4d24-47ff-af0d-1aaba54f1a19 · outbound

This paper cites The Fourteenth International Conference on Learning Representations , year=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models The Fourteenth International Conference on Learning Representations , year=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:25822d2f574f7d2a763e32f2a4e9287703d2874db8be1b4283f4356048312f43

Observation 0114579a-b248-47e7-a0ab-997e05354579 · outbound

This paper cites VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T02:59:26.128499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:590abc8e74bfa009da72cb6a22a84e8b1d91df5c779c57e7e7608bf703188240

Observation 078e324d-f897-4ff0-aedc-e057c488368f · outbound

This paper cites Forty-second International Conference on Machine Learning , year=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Forty-second International Conference on Machine Learning , year=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:29c56ff7a6cfd6ea254af29955390c3a2ba3fd4fef3517b600632ff7c935bfb8

Observation 54523f20-8be8-4fc0-aa50-f19dc2badf5a · outbound

This paper cites 2026 , url=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models 2026 , url=

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:91a728e43efabfba81c15873af858a4df1109bb9729c903f7d7bf3e14b3223f4

Observation 45673753-fe77-4d0f-aea5-2a54474f5eed · outbound

This paper cites 2024 , editor =.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models 2024 , editor =

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:4977ab6018d9c65445af83de79fe87ee1916ed45bac3cdbd92e7774a98bb9012

Observation c4e0ade0-0914-414c-b049-aab12f81d487 · outbound

This paper cites Proceedings of CVPR , year=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Proceedings of CVPR , year=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:fe453d9ebd7e36acb66642031d9cbf9950065159ec8a47cc113b491dc7aaddbd

Observation e8b01a17-4b02-4434-9b64-4d0fa48f0a47 · outbound

This paper cites and Ma, Wei-Chiu and Krishna, Ranjay.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models and Ma, Wei-Chiu and Krishna, Ranjay

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:b5c0b841abbea9114e27a18d6854a46adbbb6df9ca989414275c334eb1b6a429

Observation a7ba8f5d-f34e-48a8-aa32-eb1ca9be5226 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , month =.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , month =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:f9cabdf7dba29cbe5748070fe0651d22ed7fbb7a0cc2f452f3e818aed424da7e

Observation 695de61f-fee8-4c04-b671-77943ec5e307 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:c06cbb5ff6388307950290f935035e2d28f916b1c7c71a9ea58c6644a7835b87

Observation cb42289b-853e-4eec-b5b2-cf6c36860f02 · outbound

This paper cites ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T02:59:26.121974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:e499d030a8138b4edad04af25a8bad884fb4b2b792acaf1835f72fb104f88c50

Observation cef3d664-722f-4214-8ba4-3cdc91fde2cd · outbound

This paper cites CODIS : Benchmarking Context-dependent Visual Comprehension for Multimodal Large Language Models.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models CODIS : Benchmarking Context-dependent Visual Comprehension for Multimodal Large Language Models

Reference 35

Resolution
verified exact
doi, observed 2026-06-26T18:39:43.074903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:2a795253746bafac538112464c16327b763ccc6027bb75f08d84b180258e05ca

Observation 82d1fd7e-1174-469c-88af-c7b2dcc62d11 · outbound

This paper cites International Conference on Learning Representations , volume=.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models International Conference on Learning Representations , volume=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-26T18:37:27.171215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:b754b4fff3309a94b8583a480254e6392aefc595eae7427a77edbedf1ab46eff

Pith citing papers

No inbound Pith citation observations are available.