Pith. sign in

Paper Citation Record · LEDGER

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning?

As of 19 August 2026, this Paper Citation Record lists 82 of 82 outbound references and 1 inbound Pith citation observation for arXiv:2606.07872.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.07872 v1

Coverage vector

measured 82 of 82 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T21:56:25.350827Z

measured 83 of 83 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-31T04:55:07.057727Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

82 of 82 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved71
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch8

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 73d332f2-dd32-4ebd-8c8b-ca19da6703a4 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:be3ed4ee84f11543d3a82b80a85af41861518165e9cd06f5ba1cdfb7934c6138

Observation a52275f8-ce0b-4e57-984b-863aaad56572 · outbound

This paper cites 2024 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2024 , url=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:3de2e55067f3d5c3e491a23e7078a9f69fc2392704f4615c0981d3773fbef85c

Observation c522db48-6a0c-4649-8630-dcf2aa7b92e1 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:4f88d9420c2001fc96b31057219c40ab8dc5d4934459a0a271b6558e9a262ebf

Observation 9f841798-f745-407f-96a1-f3d102b2ec40 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:5175e0a8f3063973b71f86486b2f7ee3403da06fd0305d37862b5a447b2a8785

Observation 584627d5-9c19-4e5f-8730-95e127daefce · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:dfb9e9b0f30b9a768aabbd06cf662f326c60f086563100f18125ea22ac51c6fd

Observation ed969ea0-d0a7-4e81-abbe-9aa6c63d10a2 · outbound

This paper cites RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T17:37:14.729587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:28583bf72aefdfee9f7ea7124e46dd09c13e97a4d83abcf6640643df4ea953b1

Observation aa705132-3b4a-4c41-b429-e03bf5258b0b · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:bec7e5821291b352a6fafce1a1ce08d1e71aa32b3e3ca1f9f8eb27c06f1aee79

Observation 4e32d530-4987-4fc1-8ead-180363bf7690 · outbound

This paper cites IEEE Information Theory Workshop (ITW) , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? IEEE Information Theory Workshop (ITW) , pages=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:724340f1c6bf5f862b061cb0e468fd9e6b4c0224e409bb21e17e1891fcaff115

Observation bfccbbb3-18a4-4560-86c7-fb6726d4f782 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:254d0c9f6980fffca244ed5bef5723a3f60729a5500702180e7725f7caeba30a

Observation f68fbdae-83a5-4502-9977-1b1dd0d008f6 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:14a5b3254304bc5fa38cd3db1bc14bbe3e729692953854a40bed5c9698db4c55

Observation eb2d31a5-67b1-4c90-9f65-08baaf04bb18 · outbound

This paper cites Measuring multimodal mathematical reasoning with.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Measuring multimodal mathematical reasoning with

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:dd713070db6b638c50c5e4df6a763f9034e7be3660984682c91a37a86615f551

Observation 2419d13a-5789-440b-9d2e-7ea94d4ad529 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:0df09b547736516e224627bc3eb46db1266cee6dcde53185561af3d92feb325a

Observation e206a514-3f72-4cde-8e04-4cad33095c87 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? SAM 2: Segment Anything in Images and Videos

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T17:37:14.734860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:1a2a8f62957f023b3fdf5c3c3674ad1dfa5cf33fc35524de1b37fffcb841f6de

Observation 699f6b30-2ae7-48ea-bf9d-a2785004f3c5 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:1e30e0e184951699dc665226787d91c3fd52afd70a7212aa2f1a5c57a9b07ecc

Observation 1ca26c88-5f18-4b4e-8447-41e13d3204ce · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:7b7fd7756684b17b54a8c2322b4bc5f19694e336a4bd9fe30532b89eceb83cef

Observation 750c2649-8d1a-4be4-9ee6-c1a0006dafd3 · outbound

This paper cites Reinforced.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Reinforced

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:53fdcbea761be1d00cc4358be149d9e7d8bbbc9dc22fa39f743108a8ff08c93d

Observation 508d172c-b047-4f6a-9dad-28b2652407c4 · outbound

This paper cites BLINK: Multimodal Large Language Models Can See but Not Perceive.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? BLINK: Multimodal Large Language Models Can See but Not Perceive

Reference 17

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T17:37:14.727108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:90ca6749cb6a2cd16a796d0b1742206ca22709056f6de36db8c9fa3c1b790b37

Observation 976aaab2-d4e6-4fbc-b22c-dbcc39fe1de8 · outbound

This paper cites 2025 , note=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2025 , note=

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:6ea36d4e42e47b516fffa3f42701a338bfcb15976fa4478ca60a6be83abf609f

Observation ebec02da-68f1-482d-a1e7-6828d15c6b56 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:e9fb7f7a6792c4af281997614b3415663aade64990685ff24579fba68d751362

Observation e00d5c6a-ceaf-4432-9991-ad6e2ec92a2c · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:35dff3c026a890fc5cbe3e8bea912f56a7cdf3fc13b62657676a20ea08835b67

Observation 539e1c46-e0d1-48d8-8c9d-c248c76c2df7 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:f57459e761b06971276a21da19f9a48543355904cfb659d059e8e10749e9c367

Observation b6ffef6b-5e09-4a35-b845-84956410bf8c · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:2a4e7ecb79a4fb8a18772f066c0a8975bfb9e90a9114941891c16f029f849035

Observation 0a3a244f-1a18-46bf-941f-6012d2f73680 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:44171313f70b2ffbb9c89602b8619193c18952aad26f6a7d09df86b15f21852a

Observation a4b256ae-17c0-4c8f-b230-baf1cae5d030 · outbound

This paper cites Conference on Computer Vision and Pattern Recognition (CVPR) , year=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Conference on Computer Vision and Pattern Recognition (CVPR) , year=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:887104e51fd602dc6fa183709efabd0ba856aab00c4d5c467ebfa41b7c1abae4

Observation 674ecc18-d9f2-4006-87c0-b9d1b71c5b2f · outbound

This paper cites Generative RLHF-V: Learning Principles from Multi-modal Human Preference.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Generative RLHF-V: Learning Principles from Multi-modal Human Preference

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T17:37:14.732384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:27ef53ff4770b9f4bff08a138d1af8edd7b67fce7b97607aac9add02b8784055

Observation 37bfa628-aa4b-4262-a186-6ad2a2cd260e · outbound

This paper cites Findings of the Association for Computational Linguistics (ACL) , year=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Findings of the Association for Computational Linguistics (ACL) , year=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:d534f62cde0d851876e26dc2b0d1f11e850d96b1df1a0e927152dba40412c2a0

Observation d281e681-3834-4a2d-be98-f3f61969143c · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:db05829e8b72043661960e50046658dba1054c29b15b92d8a41142f9444f358e

Observation f5029e8d-4ec8-40e4-94fd-c42b2569496d · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:0dffefafcfa1f97289f8edc01ec77c57c89cfc220b0b6a03ab509474788d50da

Observation 96b8e16f-62ac-4ffa-a858-d00726f1f16e · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) , year=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) , year=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:40e8a0ef00998a8f74d4e4c1a145a850ae2c882d3df804105710e68f337fda06

Observation a884c1e3-21cf-4939-99e6-08f84be51f59 · outbound

This paper cites Core Knowledge Deficits in Multi-Modal Language Models.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Core Knowledge Deficits in Multi-Modal Language Models

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T17:37:14.721403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:c1dc104f8a1e440ad16c212ef76e8e1491880d125f93131bedb7478b185fa43d

Observation ea4ca770-7292-4377-9394-ccc6f4431cf0 · outbound

This paper cites arXiv preprint , year=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? arXiv preprint , year=

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:53443c0aa53a65fa903dbb479f776f49541a14d66b8d6aaa9adafc3340be0c4f

Observation cd85c497-6722-4a59-aed0-5c9a3f6d0387 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:5051eeb0f68bf833bcda5c48533507837c9233244fcc5019ba9af6a7fb0e46ed

Observation 3ec7e628-920d-4307-9122-714f9dea74b2 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:bd202528507c495b6d516685ceb2c5a24e3c0786ff5e92f225248a0222e33637

Observation 9dc6c818-da0d-4feb-92d2-297293be630d · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:290d20db57e153bf0af6d8409f05cb97f550b79b4721f70005160f256288e82b

Observation b21ff58f-20be-4076-8f5d-4f5ecc3937e7 · outbound

This paper cites Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pages=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:ded46336ceaffad2c5fe6d72dca10a5c12aa1bb6fd8ffd769fed3133c0d66e4f

Observation 5c269a40-0fb2-44d7-901d-cb10952c1ac7 · outbound

This paper cites Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pages=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:0db217bf90e36b2a3bbf893bf6365b2afa943440ea5228388864dea55270dfd4

Observation b3705f13-1356-47d4-852d-ea201afaa355 · outbound

This paper cites International Conference on Learning Representations (ICLR) , year=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? International Conference on Learning Representations (ICLR) , year=

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:c3acfb9094c05c5a6118886c55f84a6f35bc8912c00a725a9f5d9cca7cd30977

Observation fbfcf12d-4722-445c-b05e-451d43bf033b · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages=

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:70d119baf49ef279bf26d5378b6d872e6805341938f6d5249c05f23ae4cd43d7

Observation 6aa2ce82-e81d-4a17-bb77-3c47471e4914 · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence , volume=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? IEEE Transactions on Pattern Analysis and Machine Intelligence , volume=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:4c4d30fac5f633c871babfd36050427a4468ecaf4c3581ff7250039f9737d343

Observation ef45899a-dcaa-47d1-b134-020f1c0992ab · outbound

This paper cites European Conference on Computer Vision (ECCV) , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? European Conference on Computer Vision (ECCV) , pages=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:a8f52022038f376fafe30517aaf33ebd4fba6fa6f99e319530047cb8fcd73ae0

Observation 69b4caa4-90b1-432d-9c63-5a2ce0b6c407 · outbound

This paper cites Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pages=

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:0835ee7c33afd3514f37cd73fa112554e6cec7cd28745788b9a2c3f693b42905

Observation 91f3899c-ae8a-4719-a049-be286d84c82a · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages=

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:9b827dabfb423bae896e7cc4df8c509e3db62d07a3477478b132a4ca318759b1

Observation eed56838-c6c1-4f9e-8424-f50c3474869b · outbound

This paper cites Let's Verify Step by Step.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Let's Verify Step by Step

Reference 43

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T17:37:14.725267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:62e70c93d3e3cc796a81f749381ce6ba992c0ee398d026de84f934b5dbeb1e1b

Observation db57f91b-7320-4612-aa8f-87d1eebf0ef2 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:30e685725924fe44155985af0c71c26793895b27813c0217a97eb0874f69acdc

Observation 69a92182-45fd-42d5-926f-0f8b5ad90819 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:5f601838202120f2ab195cd1ef2544461dee38fe857b9aeb41b0542449c54e2b

Observation 1a021567-02a6-4bba-a061-d0f84799495c · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:77b7d2b43e5607e9c1931b4c86266f579b11f38a1027f21d501a15cb47154b0c

Observation 68242788-f916-4326-b8da-b0c2d94bc4e8 · outbound

This paper cites 2025 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2025 , url=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:4859550ceb3189785ced5edad1827443694f08638659d456f048d94bf164dfbf

Observation 92c39ce5-4898-481d-9cf0-e03db47a9db2 · outbound

This paper cites 2024 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2024 , url=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:27b7594a06bc8352089ca24fca6aac4cb333ac72f584e43f6efd7cc0add04ad4

Observation 84a91cd8-3d53-4bca-a6d3-9eaf288bc9df · outbound

This paper cites 2025 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2025 , url=

Reference 49

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:81a401e75235d2061b7669ad0f5cd5134e09db1b26220656921e72aa629f1193

Observation 05b4fc9b-4bdf-4f7b-b7ba-8888138b9647 · outbound

This paper cites 2025 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2025 , url=

Reference 50

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:ee473c510918140ab27b5a1c1a4148dc80fa5112e70536a5e37907182b69374f

Observation 2241200c-d04b-49d9-b013-21908ef09004 · outbound

This paper cites 2026 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2026 , url=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:679bedc08d715f08e6f8b9ae31757cf4888a020875607aff94fd7dcaa06d3431

Observation 4bf1b99d-f33c-48f3-a206-b57ecd9ff80f · outbound

This paper cites 2026 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2026 , url=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:85850bee7f975622d218168febe50c1b6465a9e3eae0f056a7434d78d22d9c19

Observation a55010a0-6845-45bb-b3a9-770fa78cc475 · outbound

This paper cites 2025 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2025 , url=

Reference 53

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:c90e727b5efb706301aa386306c09c8400d261908b59645f40247215fe1fdf99

Observation 8460547d-a7b8-4bca-953d-0b1576980976 · outbound

This paper cites 2026 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2026 , url=

Reference 54

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:6095687d96c0c2dacd2c929b1f97e6d974fd8615314bf149da5ccd7c9e720535

Observation deee305c-093c-487d-a376-c031f9929978 · outbound

This paper cites 2026 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2026 , url=

Reference 55

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:e4d006652e61205c27e6370325955be47400bcad4d896722f043d3b0ee7db743

Observation 0fcd8c7d-6dbf-4022-90f1-db1a7d94d6f9 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:f15ebbe3bf4f0d47006fea131a04d0fccf5ea5ed6ec773281d8671eab41742f6

Observation 4057ae2b-4d23-4318-bd3e-a928551dc6b7 · outbound

This paper cites Kimi-VL Technical Report.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Kimi-VL Technical Report

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-07-02T17:37:14.714347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:fde9c263f3c21ed8a0d9c396e1016a43e93ffb43d1beaaecf1ccacda28a90edb

Observation 8a551125-e0c7-4011-a07f-f9247e39ee9d · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:f642b27931f409c601c6de3325c1061b4e68b68ff4b4b14a96c8212629206d72

Observation b694cb91-920e-4045-bd6b-fc90f482f2f0 · outbound

This paper cites 2025 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2025 , url=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:492f7e7aa3536345a6edb0c810eef5e9b9cb87ae83cfd65b3268dd18fdb48edf

Observation 24b0317b-0336-4106-bcd1-92ccf13a45ee · outbound

This paper cites 2026 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2026 , url=

Reference 60

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:127ce263894585fe07f293c0cf50bac9b15ee94c5b4890797f2a60bca8411d07

Observation 320a22fc-03c9-4f1a-a0cb-79d3c3271705 · outbound

This paper cites 2026 , url=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2026 , url=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:ed61292f707128549fe766eaaa4833bbe31f4a1af7078db24140e85078d97636

Observation 2df3f418-31bf-4016-8370-910982a0b763 · outbound

This paper cites Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning

Reference 62

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T17:37:14.727936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:2d4512f6bb8641ec577c9dad9ecb8b04a7feea34dfb801aeeb5ae8bbaaaf249e

Observation 95799d3a-ca43-49ee-a022-c85b85859acc · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 63

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:240238ac1701867d0c83925029ac5a5529facb6969ec0097ea164e7855dca970

Observation c647b412-f0cc-4186-93d6-3e12c22f9b28 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:3b5fce4ecab2c6a82d3e39a6db03dbe01feacb0ab350b3897ea29a5875a32042

Observation a49e15ae-2a37-4475-a17c-eff5da1417c9 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:cf788001aae1bb21281a51899dbd6aa820ecc02a8b1f17d6d5a25cb105e12636

Observation 60c5df1e-a74b-4fa3-bb79-0af9a8785e50 · outbound

This paper cites Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs

Reference 66

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T17:37:14.738340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:cfcb44a52db1d30264f38b111040c63a1389f343b1899597badcca80a53cb5ee

Observation 9844d28c-797a-4717-b509-dd1f78f179a6 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 67

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:c2f023e2596e196b4a73f993ad8b1b1f02316171fb85fd416026ea4bafb42ed5

Observation d4e57d00-14a2-4460-aaf7-9f1c18c0d5a7 · outbound

This paper cites International Conference on Learning Representations (ICLR) , year=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? International Conference on Learning Representations (ICLR) , year=

Reference 68

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:bc08a32656cee44a358aa7d9801ea5090209b12f5f955253b2ff6b72be6ce443

Observation d7bb54e5-cdb5-480c-a27b-a3268d127168 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:340d361b94591cbae9d1196df6fb2d90f49f6c0548ff3f2c317a896bc2d32771

Observation 03a88d1d-cdbf-46e3-b22e-7c53649505ee · outbound

This paper cites Qwen3-VL Technical Report.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Qwen3-VL Technical Report

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-07-02T17:37:14.735843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:02d8035d1e5d22e45d0e7093cbf700df188de3bb5247b3792b90a92a11b711ae

Observation 952f409a-169a-4664-9c65-c5e043ef424e · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:b611518f184dd84bb2f2d90c742e8b46820ea58b278c4f9b6ef50e91312352b5

Observation 7a4a9ffe-31e3-4f13-8e72-14acad64ba3f · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Advances in Neural Information Processing Systems , volume=

Reference 72

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:4df02270314cdaf77b235614110be2485d91175635497fa3fa0244a9f87706ba

Observation 6e2017ee-ae65-422b-b30e-b3f5af86b2f3 · outbound

This paper cites 2024 , organization=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? 2024 , organization=

Reference 73

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:20cf50548f4d16c25c271b185cf64b866e396d8fd744bbe4f28ff1618442b936

Observation 14a2a29d-54bb-4749-9bc6-efca804c7d47 · outbound

This paper cites Conference on Computer Vision and Pattern Recognition (CVPR) , year=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Conference on Computer Vision and Pattern Recognition (CVPR) , year=

Reference 74

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:b81a4e05399ba9f87518ec2449f81d2ad02967fdd74172a18f5a30e55e149e1a

Observation 05bfeb6d-75c0-463d-91d5-68b08a3d8c06 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 75

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:413a89d043cb7e7e428b6d22d97f906eecd60cdc8f1325d12abce909a52829ad

Observation 0dd96faa-e6c9-4b3d-a9ab-1874a1344665 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 76

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:2e01d95ebecd84826b552eedfc0e7bcb3307a9c14c19f524bd7fc663943314ba

Observation d33f68f7-a0b5-4534-b9de-562ca6c29ec6 · outbound

This paper cites arXiv preprint arXiv:2509.25851 , year=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? arXiv preprint arXiv:2509.25851 , year=

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:37:14.733611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:13706eaf28bfb68c14a6dc27fe825882d2c3ad39ed4c9c86f0ef1885654d8b6e

Observation e8f5446e-4a58-4bd0-977c-1adf954666e0 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 78

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:c07df3d50474a8c8974592901d9e71100a10d990f3fa841c6a8591950bf87873

Observation 6755ed33-205c-47fb-8950-6b0299a6d07e · outbound

This paper cites International Conference on Learning Representations , year=.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? International Conference on Learning Representations , year=

Reference 79

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:b8e5b9a01e5851897525e21b39ac1981f3a1fb7cbf1f7a26a8eca89eaa969342

Observation bab4ed2c-c313-4db3-ad57-99397397d831 · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 80

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:d1986fa30e67b00c7c4adfc21231c95103bc48fc98874210978705bfca591311

Observation a6ba0b98-c64f-415b-86a7-e437f09fa01e · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 81

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:97337bd4ef1af7fd688e5d108dbde555c650f39f9b5e8d0a40bb344b37d1eb51

Observation 2701cfae-4f31-4a90-957b-dea0f319637b · outbound

This paper cites an unresolved cited work.

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning? Unresolved cited work

Reference 82

Resolution
unresolved
no resolver link, observed 2026-06-27T21:56:25.350827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T21:56:25.350827Z digest=sha256:87728f9034cefb85ee6e63c34d42e6039f886789d5199135d03969d5f9acf767

Pith citing papers

Observation 71278e46-56f6-4aa5-b899-07d7839f5e1d · inbound

Beyond Frame Selection: Generative Latent Evidence Aggregation for Long-Video Understanding cites this paper.

Beyond Frame Selection: Generative Latent Evidence Aggregation for Long-Video Understanding VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning?

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-31T04:55:07.057727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T04:55:07.057727Z digest=sha256:e05c42b9b3ad0366d868604d16f7a110cd7d5de543ab4e103694ba98df0c53a0