Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T10:47:41.183211Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 3 inbound Pith citation observations for arXiv:2606.11576.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T10:47:41.183211Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T22:09:38.842587Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T00:31:39.939203Z
86 of 86 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c1d10113-6a0c-4d28-a8b4-b2d7739a5272 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Qwen2.5-VL Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bd52fa1e-df50-4ad0-8d31-20f3e5bccc3d · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models LLaV A-OneVision: Easy Visual Task Transfer.Transactions on Machine Learning Research, 2025
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec2a6cad-d544-4e94-895d-a2fbaeb50613 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 27228014-78f0-4134-ac44-79659ba643a6 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 425fd672-8555-450b-8ae0-4d147f4bda38 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cae7c323-968f-49b6-9fbd-5210bdf1ed4c · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51a2785a-af28-4695-919e-22675b4baf02 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Ziegler, Ryan Lowe, Chelsea V oss, Alec Radford, Dario Amodei, and Paul Christiano
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c6c65d2-ccb7-4d8f-98b7-b40acae309ee · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Evaluation of Best-of-N Sampling Strategies for Language Model Alignment.Transactions on Machine Learning Research, 2025
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01a4a787-68f3-4871-b3a9-54fde78796d6 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ee07fd5-6180-4424-8c6d-503e2ba45f88 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38844e25-3d3f-4444-854e-b25c4fb62241 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models LLaV A-PruMerge: Adaptive Token Reduction for Efficient Large Multimodal Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e16d7e38-1593-43ee-a354-a52ae2db81c0 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Limits and Gains of Test-Time Scaling in Vision-Language Reasoning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 087cc637-eeed-4ff3-bfbd-2f3860f7e51a · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Visual Instruction Tuning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3e747b1-894c-4718-afff-6ed93a32062f · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models BLIP-2: Bootstrapping Language-Image Pre- training with Frozen Image Encoders and Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 905779dc-3e25-470e-988e-11643f7f7c4c · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models LLaV A- NeXT: Improved Reasoning, OCR, and World Knowledge, January 2024
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36731078-b9d7-4c09-808d-d70e45670558 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models LLaV A-CoT: Let Vision Language Models Reason Step-by-Step
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72064742-5105-426b-8ecd-cdc7f2488570 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f3b7609-3c84-4d53-aa03-14f915f7ad40 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2aed1795-662e-4db5-a588-7bd1f7bf7e9f · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Puzzle Curriculum GRPO for Vision-Centric Reasoning.arXiv preprint arXiv:2512.14944, 2025
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6e42a8d8-6041-4a40-a37e-c7654626e1af · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b3eb7964-fa77-4a78-b4dc-b970b3884e84 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 742ba8f3-c0f9-4678-bfab-e399f5d2bb6a · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models When Does RL Help Medical VLMs? Disentangling Vision, SFT, and RL Gains.arXiv preprint arXiv:2603.01301, 2026
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e38bf516-a527-4225-8fd5-2168b2346526 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Evaluating Large Language Models Trained on Code
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ab321964-d0aa-45d7-bc3f-d5d8a82875fd · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 299826ec-235d-4284-9d1e-6b7e704518ff · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Majority of the Bests: Improving Best-of-N via Bootstrapping
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c327420-b881-45b2-a2bd-b0f7136fda44 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Self-Refine: Iterative Refinement with Self-Feedback
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d96d1cb-d588-49b6-8129-821e39e7d313 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Let’s Verify Step by Step
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61c055f9-9f92-4601-92d0-aa93e4818f0e · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55e7a0d6-a857-4bea-98eb-601dd7cc944a · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models s1: Simple Test-Time Scaling
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c16a5dc-dadb-4fb5-8ea1-f9021b8ff470 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models SelfBudgeter: Adaptive Token Allocation for Efficient LLM Reasoning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3e831cad-a8a7-4145-a1a7-e22b4c8de4be · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ed61dc1a-580c-4023-916f-dbf868dd6c80 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Scaling LLM Test-Time Compute Opti- mally can be More Effective than Scaling Model Parameters
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0b93ebd-32e2-45cb-a59a-9438605e49f2 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Learning How Hard to Think: Input-Adaptive Allocation of LM Computation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bec2ef43-3501-4268-ade0-ea680fe68fe0 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models CyberV: Cybernetics for Test-time Scaling in Video Understanding
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bb97b15e-69c1-411b-b3b6-05a219fbbe90 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Vista: Mitigating semantic inertia in video-llms via training-free dynamic chain-of-thought routing
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 37009677-e164-4331-94ee-6427dc504079 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Learning Adaptive Reasoning Paths for Efficient Visual Reasoning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9dbd312c-ab38-4353-abba-14e53e86babe · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models ARES: Multimodal Adaptive Reasoning via Difficulty-Aware Token-Level Entropy Shaping
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e91a4022-da5f-4c1c-b022-f8fb5c3f7a72 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Efficient Test-Time Scaling for Small Vision-Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1a728f2-27eb-4f0e-9a5c-adbf4f408ff1 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Improve Vision Language Model Chain-of-thought Reasoning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2f37f97-6c6d-47ff-860a-3f66a96f996e · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Aha moment revisited: Are vlms truly capable of self verification in inference- time scaling?arXiv preprint arXiv:2506.17417, 2025a
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a1e3b4b9-1721-4539-9413-73f061516819 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Similarity-Aware Token Pruning: Your VLM but Faster
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cf165c29-1646-4073-9ebc-4ad66361df70 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72d7ac93-7673-4491-9f3a-fc9f37fad972 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abf55d21-b095-4827-a0d7-7738819772b1 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models VisionZip: Longer is Better but Not Necessary in Vision Language Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06901dcc-af4c-4161-92f4-676f8d42fed4 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models PruneVid: Visual Token Pruning for Efficient Video Large Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c4a636f-f22a-4396-bebb-15e0992cb8b1 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Token Merging: Your ViT But Faster
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5059fa4-e6ac-45b5-b1be-7c25f572cfca · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd85ac8e-8961-41b5-b746-a30b97f927d8 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Beyond Attention or Similarity: Maximizing Conditional Diversity for Token Pruning in MLLMs
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f38d9df-4fc2-4966-b07f-4e99f125a18c · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a39d5062-6d91-4620-963a-475933721c25 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models TopV: Compatible Token Pruning with Inference Time Optimization for Fast and Low-Memory Multimodal Vision Language Model
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63bc77e3-8231-4112-a387-473ba560c1df · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models CATP: Contextu- ally Adaptive Token Pruning for Efficient and Enhanced Multimodal In-Context Learning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3204da28-5f07-4cd9-9c40-87eec7f451f0 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models FlashAttention: Fast and Memory- Efficient Exact Attention with IO-Awareness
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 194aa806-0e9e-41e0-ae4e-fd6a7db41eba · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models VScan: Rethinking Visual Token Reduction for Efficient Large Vision- Language Models.Transactions on Machine Learning Research, 2026
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d9986cf-33b2-4062-aafa-6d9daa60909a · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models KeyDiff: Key Similarity-Based KV Cache Eviction for Long-Context LLM Inference in Resource-Constrained Environments
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31df5339-a59d-47c3-afc2-3584911f23d8 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0abf0c17-b79d-42be-9b60-6ff67d19e93c · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d5a536d-7141-40b1-bc52-b4c520926909 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? InEuropean Conference on Computer Vision, 2024
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c586b138-e026-4719-9f58-780e3cd94b57 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset.Advances in Neural Information Processing Systems (NeurIPS), 37:95095–95169, 2024
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dcd9a28-6285-40c7-ab2c-6ee507bdc2e1 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models DocVQA: A Dataset for VQA on Document Images
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bcb69df-7035-4877-bc02-b85a5c0027c7 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c373c94d-748c-47b3-9de4-a5b9312d4020 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b58cd25-3106-42f1-bd92-3a042c2a9c8b · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Are We on the Right Way for Evaluating Large Vision-Language Models? InAdvances in Neural Information Processing Systems (NeurIPS), 2024
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4102869c-37cb-482c-94e8-bedbbb2cf7ea · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models MMBench: Is Your Multi-modal Model an All-around Player? InEuropean Conference on Computer Vision, 2024
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93d8c24e-56df-432a-80cd-d16399b2fed4 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65262a67-54f0-4467-bfa2-bcbf099eb45e · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Evaluating Object Hallucination in Large Vision-Language Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e8e5a5b-8fa7-4e66-8029-d71d64dd13a9 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Smith, Wei-Chiu Ma, and Ranjay Krishna
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ffeb757-ee6c-42c6-845b-eb4e9dc242ba · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Traceable Evidence Enhanced Visual Grounded Reasoning: Evaluation and Methodology
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 349039f5-68e8-4699-bd8d-c0d466c4a2c7 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fef393f5-4bfb-48dc-a064-20bb0c0b9bc2 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models TempCompass: Do Video LLMs Really Understand Videos? InFindings of the Association for Computational Linguistics: ACL 2024, pages 8731–8772, 2024
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b3bce9f-06b3-445b-b472-df869a9fbc17 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Towards Video Thinking Test: A Holistic Benchmark for Advanced Video Reasoning and Understanding
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb99e5a9-a171-4409-9dc0-eb817966d233 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models MVBench: A Comprehensive Multi-modal Video Understanding Benchmark
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58316217-3316-4300-bc9b-2a9080b7e29d · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Q-Bench-Video: Benchmark the Video Quality Understanding of LMMs
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 477122f4-1492-47ec-b288-ce6ac919f074 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cdf5bef9-e05c-4e55-a5ba-be5961c06ac5 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Gonzalez, Hao Zhang, and Ion Stoica
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1b05748-a5e5-4f6a-b77b-ce0da3055a87 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0781584f-13ad-4afd-8480-cde6783bf8a0 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Perceiver: General perception with iterative attention
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80c56837-1c06-4d64-87bd-666e635e9582 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models LLaV A-Video: Video Instruction Tuning With Synthetic Data.Transactions on Machine Learning Research, 2025
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c76ad6c-71e5-4935-b2f9-16bd3eaa3c37 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models A- OKVQA: A Benchmark for Visual Question Answering using World Knowledge
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8324f67-9fb9-4970-adc8-9d5e3a03ce92 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models A Diagram Is Worth A Dozen Images
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4730b2f8-e46a-4d3f-96de-10976aa182b1 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Are You Smarter Than a Sixth Grader? Textbook Question Answering for Multimodal Machine Comprehension
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f1b29ca-cfde-4359-9103-f79fd6369aeb · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6aae95e4-c377-4639-8cdb-bc4ade857cd2 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering.Advances in Neural Information Processing Systems, 35:2507–2521, 2022
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba7e04dc-8421-4f85-b33f-42c64933907a · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed2d3af1-df27-4b17-8d27-b92274981a9d · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0973a0c-117a-4aa3-81ee-87c8776e8ae8 · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Are We on the Right Way for Evaluating Large Vision-Language Models?
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 300c80a8-6768-4b64-8bad-cb125859067d · outbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Guidelines: • The answer [N/A] means that the paper does not involve crowdsourcing nor research with human subjects
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94581dcf-32f0-4a7d-a91d-e740d1b659df · inbound
It's the Decoding Format, Not the Perturbation: Auditing Consistency-Based Selection for Vision-Language Test-Time Scaling AVIS: Adaptive Test-Time Scaling for Vision-Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 100e795c-0d12-452d-b2fe-f0ce1c4abafa · inbound
It's the Decoding Format, Not the Perturbation: Auditing Consistency-Based Selection for Vision-Language Test-Time Scaling AVIS: Adaptive Test-Time Scaling for Vision-Language Models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e7fd8d1-21cb-4921-a874-0d840380841b · inbound
Not All Visual Tokens Are Equally Safe to Remove:Consequence-Sensitive Visual Token Compression AVIS: Adaptive Test-Time Scaling for Vision-Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.