Pith. sign in

Paper Citation Record · LEDGER

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning

As of 18 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2509.22746.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.22746 v2

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-18T13:33:12.508639Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact38
  • verified fuzzy23
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eae4980e-6bef-4e83-8138-d448e7efe1f3 · outbound

This paper cites write newline.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning write newline

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.750252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:612340682d346be99e65c4cd384d10659ea83d007a08b985ef1117a1bbaa2c8b

Observation c8ac538d-0065-4f3b-a0db-1d6e6af1e94c · outbound

This paper cites Qwen2.5-VL Technical Report.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Qwen2.5-VL Technical Report

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.232148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:b59324d440e46dcbc2c1987a5eadcffc52a79396b911c10440d6a96bde8da49b

Observation 17c759d6-33ce-4c0c-81a8-af07493f16c6 · outbound

This paper cites Graph of thoughts: Solving elaborate problems with large language models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Graph of thoughts: Solving elaborate problems with large language models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.743189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:f13d8cb95cb49e4e5c67b62dd57b92577e60f962ec62286836993f43194c901f

Observation 6d1a59b4-9003-4387-9eaa-8ccc8389a82b · outbound

This paper cites Ground- r1: Incentivizing grounded visual reasoning via reinforcement learning.arXiv preprint arXiv:2505.20272.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Ground- r1: Incentivizing grounded visual reasoning via reinforcement learning.arXiv preprint arXiv:2505.20272

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.149784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:34a1cba327f395d006290d5281d25ea8385d7983e9da1ede6c56b6c788d4fce3

Observation d4b59fc0-9878-4ddd-876b-91f6ccbf08e0 · outbound

This paper cites SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.155140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:afd21e5a78627c713a568a55b8a205685739df80e92d96abfe1b65e8109e5190

Observation f194457a-7a47-4c76-b4ce-a86eb070ed54 · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.302737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:4f3bcaa9352646a90ce0a5329a47a3a3095545da6715542e492089fc83feb767

Observation 6ef69faa-0bc5-496b-8918-bf34aa4904a2 · outbound

This paper cites Are we on the right way for evaluating large vision-language models? Advances in Neural Information Processing Systems, 37: 0 27056--27087, 2024 a.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Are we on the right way for evaluating large vision-language models? Advances in Neural Information Processing Systems, 37: 0 27056--27087, 2024 a

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.739617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:5e92c72a9a1c039d47f4f888edc4eaa3a95c572918946f22aa676e9b3ad7ca1e

Observation e43bead3-ca1c-4d9d-9253-4a27eda87d19 · outbound

This paper cites How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.682307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:35e9fd3eeb8bac02b5303ecf2db1ec8ab05ffb3b8630e3bc24bba75510f5484b

Observation 3e86c49b-ab93-4d7c-8b77-484c4265308e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.236412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:ff672723bbbd944cb96f6ad6913bd287aff30bfe9a6b6c66d176c8ed8f33b880

Observation 8bbbb5f5-c80c-480a-8c5d-981d46d98f93 · outbound

This paper cites Insight-v: Exploring long-chain visual reasoning with multimodal large language models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Insight-v: Exploring long-chain visual reasoning with multimodal large language models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.746641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:12dcc1765252a83ef4bc0bedeb4a6aa0c7d8c5347c405269aaa2f07c6e36ee82

Observation d21b6467-f76a-4848-8bce-95109d40f558 · outbound

This paper cites Virgo: A Preliminary Exploration on Reproducing o1-like MLLM.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Virgo: A Preliminary Exploration on Reproducing o1-like MLLM

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.184297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:d0a95885ca8cda136e1971146ac6af816cbc9455036dffaacd6fe6e89b15f249

Observation 9c9dc810-3210-4559-aafd-a0dafc2aa787 · outbound

This paper cites GRIT: Teaching MLLMs to Think with Images.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning GRIT: Teaching MLLMs to Think with Images

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.109996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:5b1a8e1a558a99f9a5763a0b9b70186b33934ef488ec83aff6e97f5ba9d31c9f

Observation 3f8f19ae-d2cf-40d4-ab00-d60404a94583 · outbound

This paper cites G-llava: Solving geometric problem with multi-modal large language model.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning G-llava: Solving geometric problem with multi-modal large language model

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.730466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:f048a68b11f1e492c0daefee2c8f3a29d96fd9df585d98bdb5055c237433811b

Observation 0aed07cb-5226-4026-b352-fb2322e0c3bb · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.213344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:221a477ea24b5c82676084778e553d89f1dcf135c12240459e8d446db9d7f8d1

Observation 7fa6fbd7-27a6-4211-ae90-4efd5d9c972e · outbound

This paper cites GPT-4o System Card.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning GPT-4o System Card

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.105758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:7be09be32d901fde76968db846922b813af119e8c75f396e3d7dfc6812b87c91

Observation e46a6aa9-4bfd-4f67-82dd-d0f155e4dc36 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.170186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:78a5eb7bcd4c01e28766f5a0a93c79c40f4db856e3a7bd2e715d13d31f6d7853

Observation 8973cd9b-1d6a-4b9c-b762-eb876efbe796 · outbound

This paper cites Large language models are zero-shot reasoners.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Large language models are zero-shot reasoners

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.723765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:7403f9495fbb5da43b155cb57e467f7b0e48798f58b23dc9afabbc9ba3e01dc7

Observation bb16ecd9-18b3-442d-aff4-02166106f159 · outbound

This paper cites Hypertree proof search for neural theorem proving.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Hypertree proof search for neural theorem proving

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.727128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:592230bad4eb46fee74caec78ca9811146ccbe06797b259811288c73e07d00fa

Observation 596e9e5d-faa1-4a5c-8635-1f1fdec22f2d · outbound

This paper cites Scaffolding coordinates to promote vision-language coordination in large multi-modal models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Scaffolding coordinates to promote vision-language coordination in large multi-modal models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.736572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:d1041ba34868a236b12a8bf9b80b5a82350325c1ff3f3a5c2b11c81458af326c

Observation e21b9d30-ad94-4c8a-89a7-2b52dcf8ce34 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning LLaVA-OneVision: Easy Visual Task Transfer

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.193876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:6de3e66d4eef29637f6d82740ba4c342171e4fac2da8aa185e11460f6216cf71

Observation 6dd36938-4398-4374-b00a-c05e0a2b47a0 · outbound

This paper cites Numinamath.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Numinamath

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.733370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:1a2bca6888eb91492167ebb87680c57f874bc403a20c8b9ccc140612e93bc3fa

Observation 6d9d66cb-4387-4d7b-ba27-34cfb4bc259c · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Evaluating Object Hallucination in Large Vision-Language Models

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.174850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:a9242131ed2444f587353a1c0cbf76fb1da6d425cef729b359fe98c40c041613

Observation 494333cc-0eec-4838-ba19-b8d7b4637278 · outbound

This paper cites VoCoT: Unleashing Visually Grounded Multi-Step Reasoning in Large Multi-Modal Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning VoCoT: Unleashing Visually Grounded Multi-Step Reasoning in Large Multi-Modal Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.203478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:c8b15e84626e629069cb9f6c20935d33a7d7edd75411775e4e69252daddb0281

Observation 55ab3a8a-ffba-4052-b05a-8e2206cdbf30 · outbound

This paper cites Visual-RFT: Visual Reinforcement Fine-Tuning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Visual-RFT: Visual Reinforcement Fine-Tuning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.298265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:a9c0b4625e00b86173b336cd6e835643e39508ca957a1fa03de09a6e968e0706

Observation 86a720e7-bcee-4441-a8da-013578a46e74 · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.143460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:184bee3bdf720e372a787ae27d7c5aea7e8e208c0cb079e31249655ba1d4893f

Observation 489fd87b-cca6-4c56-a5fd-ae0c8460da40 · outbound

This paper cites One RL to See Them All: Visual Triple Unified Reinforcement Learning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning One RL to See Them All: Visual Triple Unified Reinforcement Learning

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.286858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:2dfa242a1a59c8c4219acc88b5465e096ffa45ecf97f45c3038c07d9f0833b66

Observation eadc6229-3e3b-4bc8-923f-6196c9449831 · outbound

This paper cites Self-refine: Iterative refinement with self-feedback.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Self-refine: Iterative refinement with self-feedback

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.671092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:ead1cd9f00503b41530eb7000ed9282f4cb95a5f8e716380d7707dc36ecf3d95

Observation 9f1c250c-57b9-4e58-8e2c-4a66ec98a420 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.130294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:e12dc46da14b18d61bb90dcd25b8db24a59549dd0bc354f283ec4ebd0b95a50a

Observation 613234ac-408c-4fae-919e-eb474ca6ee8f · outbound

This paper cites Compositional chain-of-thought prompting for large multimodal models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Compositional chain-of-thought prompting for large multimodal models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.663341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:f6eafac62693068ef0e1a7138c3b9223817711aa0c9112d2cc18d2258f5488ef

Observation f9f65693-3908-4838-91d7-c64f85198f84 · outbound

This paper cites Omnicount: Multi-label object counting with semantic-geometric priors.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Omnicount: Multi-label object counting with semantic-geometric priors

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.656261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:677b860782a493428147f9a6c07f3eb0b5fcc2c6a9b2ac36e360e7a83873df5b

Observation 671b24a3-b707-41dc-bda2-00e89fc29b94 · outbound

This paper cites Thinking with images.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Thinking with images

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.667756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:d89eff788cb03fc56bc3208fe216815c17795eb329a7918f40a7feb3b1b1b507

Observation 54bc6d70-db0e-4b97-9749-2feecf785e46 · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.114501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:088688fc47da34805a866bbdf0713ff8fac8ff836ab5fcef62c3bb0ccab4eab7

Observation c6303bcc-b9d2-4b01-a448-de45481f3088 · outbound

This paper cites We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.124560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:0c03155d93adb78faebd6d72ed7d3366778863f1b60715c14da1540dfc7bf54c

Observation 0e8882d0-2c75-484b-93d9-0e90df5dc9bd · outbound

This paper cites QwQ-32B : Embracing the power of reinforcement learning, March 2025.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning QwQ-32B : Embracing the power of reinforcement learning, March 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.674868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:540a0614dbac4bea2d2ac32641ee1250ae9e47aba0fbd3681c3eddbda732104e

Observation eb758a0e-2145-4ff9-afca-dceab286a150 · outbound

This paper cites Grounded Reinforcement Learning for Visual Reasoning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Grounded Reinforcement Learning for Visual Reasoning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.189444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:1bf978ea2101218bdab555b7c05e0eff99b97e41e02e189a48dc89dc84828bcf

Observation 2aa5a4b8-039d-4a8d-936d-578fd348553f · outbound

This paper cites Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.165701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:fabbd87d2e6bf292e221b82828a812879872e6b78ff8a79e751e8afddf67280c

Observation d55c139d-dd28-4bd1-965c-94d5930f9d24 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.198881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:81b75edd146e09d1ef56d9caa8f2a19e4c236485094d6a4a527c23b8c5a9e9c6

Observation 6b4bfc76-4bee-4a75-ac32-64f54fd4d0b2 · outbound

This paper cites ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.241161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:a6db74f7f806c88635cc47a7932548f0fdd05b631d59576de9dea9f33481badf

Observation df381330-8b9c-4ad3-9593-fbc176df5458 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Reflexion: Language agents with verbal reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.678876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:4a39849ef15619a3d69c113a864b7e25ccfad6d403069e117dfb3979360c47a6

Observation 848dac7b-ef8a-4f42-a83e-ab8cb60f7440 · outbound

This paper cites LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T13:36:25.179472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:f13beee97da4434e40810f45bae3f459b603696fe48a9323af3685fcc37533a5

Observation 34c7873a-49d0-4992-aa6d-4e9c1cc6747c · outbound

This paper cites Toward self-improvement of llms via imagination, searching, and criticizing.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Toward self-improvement of llms via imagination, searching, and criticizing

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.772917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:af4358a34d793258ae415d98c9f7359105819070408b1d90c51bc85e3cea9d37

Observation 96f4f656-afad-4a21-94d7-8d4589147942 · outbound

This paper cites Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.223356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:8ca6753a29483d5fab06f2d7515bb17eeb2d6fa846b375bc5bc9e8a993340b25

Observation 58d32fa8-b084-48b7-a943-96adcba97d2a · outbound

This paper cites Self-consistency improves chain of thought reasoning in language models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Self-consistency improves chain of thought reasoning in language models

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.764897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:364c1645073f20dc1e1e9eaa9a198da6725f0b22232890a380f526c97632f4f7

Observation 567823d3-0618-4525-af76-ca165c7c2c7a · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Chain-of-thought prompting elicits reasoning in large language models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.769023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:c07529fd1c83d7f5de02134b2651824b2571df04d52233e8d716a92506936d32

Observation 0706187b-201c-4eee-8fa5-6dc413ea7157 · outbound

This paper cites Open vision reasoner: Transferring linguistic cognitive behavior for visual reasoning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Open vision reasoner: Transferring linguistic cognitive behavior for visual reasoning

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.136959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:ac81ba96aa9995a7490a7befc773a2a1b4948e22f1572fabe61ab7db47e791a3

Observation f669f1a8-8023-4c2a-8d36-def6a0b6de81 · outbound

This paper cites SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.264926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:3133c8f1e2218fdef18d645ce28767d8a5b5e5494fc864514874eb53b7235c1f

Observation e64e0d55-fe91-403e-8dc5-24610c1afd0b · outbound

This paper cites V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.246023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:92cd03c619ee347a77e6a9d565028ea0f6bbfd3089e0b25bbfd2eea342723a7d

Observation 17b940df-2077-4fa0-ac82-5b632675a6f1 · outbound

This paper cites Grounded Chain-of-Thought for Multimodal Large Language Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.208637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:0dfad7920cda87a89c4f0c1f6a485b3f9a91e5b921a4c45accfffdd395c8f728

Observation 09915245-a471-48c9-899f-0195c2f18fa8 · outbound

This paper cites Self-evaluation guided beam search for reasoning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Self-evaluation guided beam search for reasoning

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.776685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:effc9bfbc9ae7ba873745cad00d130882536a771c832f7dbcc9f3093f27b7f73

Observation ef47cf23-1767-48a0-8c37-d3d5bfeceadf · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.280777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:2d151da56d7030db12a947b13ba23cabf51e19b32e5b7d694659cd5f40e87b80

Observation c2ff1905-bcb4-4ad8-b597-6d20402c7b01 · outbound

This paper cites GeoSense: Evaluating Identification and Application of Geometric Principles in Multimodal Reasoning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning GeoSense: Evaluating Identification and Application of Geometric Principles in Multimodal Reasoning

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.255208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:8193f77321f1c289f8d7e1e9659e50d6e9f9bc8def3d9e18bfa02ff01b0104b2

Observation d0084653-492f-4f57-86f7-7dbd5c26c31e · outbound

This paper cites Set-of-mark prompting unleashes extraordinary visual grounding in gpt-4v.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Set-of-mark prompting unleashes extraordinary visual grounding in gpt-4v

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.761087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:b11f78c0f3bc45dde9ea1e55261d7d5dcf6f743d268e7ac27dc7ba35ca51a75f

Observation ada1b670-4925-4ab9-8335-318ae0fea32d · outbound

This paper cites R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.260535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:1d6a4d9d225cf4372f8d5fb629e8f0c1c609465308792ea4c1f824b89c4d128a

Observation 1901e20f-0587-4281-904a-fc08c4dbeab0 · outbound

This paper cites Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.218707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:ff149841136fad1593ddaae6ce00796a5bd54a08dc39af78b3830cd57944acfc

Observation b1cb55ff-7440-413f-84dc-0803f795c833 · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Tree of thoughts: Deliberate problem solving with large language models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.757761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:132a7bcbbe4b7ad18a754aefeb89a699c786ff6b22409ea87452dc1303692ca4

Observation c9be5a1e-b53a-4b62-b4db-e4ca1eede66e · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.307894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:d75abf2878df4ec52509921281e5e962000d73bb92f2a23a764b1567c81c420b

Observation 64f8fb67-8516-4565-a997-430f917769b6 · outbound

This paper cites MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.269615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:c48566750f80a164ecb18211cbd6d1817f54eed68267410bc7990af6f5ddfd7d

Observation 1c72cbf4-188f-4bfd-9167-af8f49785e6a · outbound

This paper cites Improve Vision Language Model Chain-of-thought Reasoning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Improve Vision Language Model Chain-of-thought Reasoning

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T13:36:25.275632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:a38f8e3216dce908d700194489a5e63c0eec74641488a4ef7f472d9743cc4ac6

Observation b35e14bb-6c22-4050-877b-00556a8a467d · outbound

This paper cites Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.119251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:ac3b7e4a58d7de7cc163c3e62e8794eed2a8c2e71191ebdd01f8003e671702be

Observation e14bc0e6-da42-4368-9d66-5d4696bbf492 · outbound

This paper cites DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.292168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:e6cf43ac9bf488daddbcee9d0304fcb37178a9b440d55928500f68b07cedcbd4

Observation 0cf59d23-0df2-4584-a084-056771c44046 · outbound

This paper cites Image-of-Thought Prompting for Visual Reasoning Refinement in Multimodal Large Language Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Image-of-Thought Prompting for Visual Reasoning Refinement in Multimodal Large Language Models

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.228065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:5f2ac9912bb4c059576d93edf835881de6bbb5c59cbd46a0aacf289514239f0a

Observation f714d8db-b659-4c1e-b422-80eb5910e66a · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.250615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:e38936bdb2b079de1396c1225252878581860ec5f3949fc4c24fcd8733dcff73

Observation fe667d02-bb25-4f8e-b3e8-a8bef73ff84b · outbound

This paper cites @esa (Ref.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning @esa (Ref

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.652566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:6c00a84bdc302e18828e3c0aead994b51873c35165d64a614b7ee6671e037702

Observation 3acfbb19-209a-46af-ba66-2e891a7e98ee · outbound

This paper cites an unresolved cited work.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-05-18T13:36:26.754015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:196a7214d64761db7c29a75907d9eec7bd5e2a1842fe4751b75d2bfc065332e0

Observation dddb4ee2-b4b7-4bc8-b4bf-9f7c464dc0d2 · outbound

This paper cites (QGT+  o/߸ ;fQ Zt鐒gvZxG*J Y ȮY! dZs (HE E 2 n=#R.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning (QGT+  o/߸ ;fQ Zt鐒gvZxG*J Y ȮY! dZs (HE E 2 n=#R

Reference 66

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T13:36:25.160912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:8239ff8a2f675f53d7b761e5d7421ad3171547e758a6f21fba78b4bd3932f0c9

Pith citing papers

No inbound Pith citation observations are available.