Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:47:01.785737Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 14 inbound Pith citation observations for arXiv:2505.23558.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:47:01.785737Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T18:35:57.326931Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T20:56:13.710075Z
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 60f2861c-ab4c-4645-bc90-50a223680fb1 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fcef8b2-045e-4190-9575-d16262bf1bc3 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4d0c8dd-2a5e-4fbf-ab85-a3e4745f2f97 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Step-level Value Preference Optimization for Mathematical Reasoning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bc66e3e-f9b7-42f2-92bd-0c0baef14c5f · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Are We on the Right Way for Evaluating Large Vision-Language Models?
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1673187c-8713-4f58-bd18-612a207dcaf0 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 919b809c-8c4f-4fe3-956a-f558d02f8ecf · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaee9fe0-d871-4513-a911-4c97683c6145 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Domaino1s: Guiding LLM Reasoning for Explainable Answers in High-Stakes Domains
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a513fc2-8e07-4d06-960f-ea9839f087af · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Multi-modal hallucination control by visual information grounding
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fd15205e-b644-4b87-925c-6436a4cf584c · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 352f77d5-3375-454d-9d7d-b45883502b41 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Interleaved-modal chain-of-thought
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 20f902ef-7921-45f8-b284-8687f67e0516 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e68b1931-f95f-453c-b956-09a05213bcd1 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d5df679-2071-4885-a2e0-647e16a7b1c4 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection-allocation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 033b129e-89f6-41dd-848c-a5e69abb4537 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 741f2efc-eef0-49d3-8ff2-1ee6e5acdc7c · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information OpenAI o1 System Card
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f55dcc76-586f-4dfe-9f97-b2bc8b2fdf61 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Biomistral: A collection of open-source pretrained large language models for medical domains
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b5158816-4f1d-4986-b535-6059ce02f7fb · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Evaluating object hallucination in large vision-language models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a6b90a7f-732c-47cb-b351-cabaff831831 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models via Visual Information Steering
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be3c8ace-1240-4e8f-96df-d91da30c9868 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information SceMQA: A Scientific College Entrance Level Multimodal Question Answering Benchmark
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 123bfd87-ab2b-4802-8549-14d80ab4d50e · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information LongPerceptualThoughts: Distilling System-2 Reasoning for System-1 Perception
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d2a8a9f-486d-4d1f-9cc6-99fa2dd92d98 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Microsoft coco: Common objects in context
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 763e2e7c-a068-4cbe-aa2f-ab0415f2af7c · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information A Survey on Hallucination in Large Vision-Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45410499-e16f-4cab-9deb-be28ce4a0e86 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Paying more attention to image: A training-free method for alleviating hallucination in lvlms
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 51a56be6-df60-46ad-9151-b97942bd1319 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Mmbench: Is your multi-modal model an all-around player? In European Conference on Computer Vision, 2023
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 558c957c-870a-4438-88c3-8e79bc8119a1 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Visual-RFT: Visual Reinforcement Fine-Tuning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd78694e-34c7-4efb-bfb5-dac2d468a5de · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Inter-gps: Interpretable geometry problem solving with formal language and symbolic reasoning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6358699e-f146-4869-8c2f-51d4bc6c52f8 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Self-refine: Iterative refinement with self-feedback
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5e35968-f5ed-4daa-adde-f75dff7d9c8d · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Llama 3.2: Revolutionizing edge ai and vision with open, customizable models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a7418947-df2c-4e2b-a983-b4eeec6425af · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Training language models to follow instructions with human feedback
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ae12a27c-aacd-4056-9629-3b1e1ff09aa6 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92a6b55c-ab2b-4c49-9a12-cdc211b13e87 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Object hallucination in image captioning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0725dd7e-172c-4309-87d9-727c3714796b · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70cfd12d-6bb8-4076-b6aa-c696ae98a268 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb4c7f28-f036-42de-bae9-f07e8590b54e · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Aligning Large Multimodal Models with Factually Augmented RLHF
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 829e5e99-5495-429a-ba3e-38a27a30850a · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 583ae5aa-a7bf-4e70-918a-c47143a2f432 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Kimi-VL Technical Report
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c511529b-98b2-4b79-a228-5d26dff8fecf · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Qwq-32b: Embracing the power of reinforcement learning, March 2025
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffaf6595-cd68-45fc-8ce8-5698ee2aa02a · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Is a picture worth a thousand words? delving into spatial reasoning for vision language models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b4f6cf26-2359-498a-9f6a-86ef2d58c6c1 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Measuring multimodal mathematical reasoning with math-vision dataset
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d746bc4b-215f-45be-bd21-171ad9a37e00 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Measuring multimodal mathematical reasoning with math-vision dataset
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a418903f-ed73-4b48-b6e9-4361ec7457ec · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67f8666e-ca3d-4af8-a749-fc20a58d4965 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba9b954b-098b-449f-86d4-2609cd100792 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Effectively Controlling Reasoning Models through Thinking Intervention
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6545da88-f4b4-4137-8dcf-8ef097ed4269 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Deepseek-vl2: Mixture-of-experts vision-language models for advanced multimodal understanding, 2024
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97cfb2fa-c047-4b8a-b8a0-dad610865a94 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Monte carlo tree search boosts reasoning via iterative preference learning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7260feb5-fffd-4693-b521-b44416107a60 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information LLaVA-CoT: Let Vision Language Models Reason Step-by-Step
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1eeb526c-3259-4c60-8d94-077fd5debf1a · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15fad10e-f36d-4452-91f7-f6b04cb1e2f2 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Introducing Visual Perception Token into Multimodal Large Language Model
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dd40c86-4953-4b08-9e51-f7ac614f600a · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Advancing llm reasoning generalists with preference trees
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2168333b-3154-4a8b-85c0-796d7a797f7d · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c732349c-ed7c-48c5-951d-accacbfb2eda · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5aba6c03-03ed-4c42-b360-c96b929b6138 · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Self-contrast: Better reflection through inconsistent solving perspectives
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 20243db5-940e-4204-80eb-b655940ea41f · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information Analyzing and Mitigating Object Hallucination in Large Vision-Language Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7c07630-ab3d-4dc9-8703-8c114c61759b · outbound
Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information correct" or
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f2029779-be6f-4f46-9899-fc776e715073 · inbound
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 249
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation abc59b88-cb27-4942-91d2-6eec3e75b2fd · inbound
A Survey of Reinforcement Learning for Large Reasoning Models Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 06c7f784-ff5a-481c-af43-884de631629c · inbound
Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bfc4a53-e113-4053-b49f-f3a8bfc9871b · inbound
VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f082c79b-181f-43b6-9177-cebbd494ba4d · inbound
Reflection Anchors for Propagation-Aware Visual Retention in Long-Chain Multimodal Reasoning Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2b9c0382-4228-465e-8d9e-8b8b7bad606c · inbound
Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 802207f2-0def-40a0-a74c-838777043278 · inbound
PDCR: Perception-Decomposed Confidence Reward for Vision-Language Reasoning Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4b0eca96-75f6-4ed6-97a1-39d6de2693a8 · inbound
Vision Inference Former: Sustaining Visual Consistency in Multimodal Large Language Models Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a657bc6d-0f89-468e-ad1a-a62fc1c75155 · inbound
Vision Inference Former: Sustaining Visual Consistency in Multimodal Large Language Models Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3bc2bbb6-ff74-4f4e-904b-77d8cf1c76ab · inbound
DeepLatent: Think with Images via Parallel Latent Visual Reasoning Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 82634a68-c29c-4b80-af0a-e17b2ce07de2 · inbound
Trust Region On-Policy Distillation Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5bc3f27d-8948-4577-8e54-56eb41bc179f · inbound
Recompute or Reuse? Diagnosing and Mitigating Textual Shortcuts in VLM Self-Reflection Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03ffff5b-05c1-40d8-8dda-06e9bffd6418 · inbound
ReGround: Restoring Visual Grounding in Multi-Step Reasoning through Self-Diagnosis and Visual Re-Examination Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 814def5c-f9f9-4f37-b6d8-ae11523efc6c · inbound
FOCUS: Decoupling Expert Personas in LLMs to Enhance Domain Expert Capabilities Qwen Look Again: Guiding Vision-Language Reasoning Models to Re-attention Visual Information
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.