Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:40:40.290572Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 2 inbound Pith citation observations for arXiv:2507.10202.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:40:40.290572Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T20:30:58.091365Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-19T14:27:24.213919Z
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 851cbd77-5559-4208-8134-c10a629ff47c · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9606cab-eb93-46b2-b7b9-3a150b16005e · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Lan- guage models are few-shot learners
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6b6550d-3f4b-4a44-b865-a89967a88b5e · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Emerg- ing properties in self-supervised vision transformers
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c6bd09c-3ad6-419c-84ac-49eb1fbe995f · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5fcf4a7d-c972-46ca-980c-90db114f6b68 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62070ed8-9b1d-41c4-bde5-dcaef039002a · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Palm: Scaling language modeling with pathways
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8519e9a4-b332-4971-bb49-87a39693aade · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Llava-uhd: an lmm perceiving any aspect ratio and high- resolution images, 2024
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e48fe762-3355-4422-ad46-75249e10b127 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images GPT-4o System Card
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 499b9c21-a768-4a9a-ad3c-35bb8e1c4779 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Screenspot-pro: Gui grounding for professional high- resolution computer use, 2025
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c92f28e4-58b9-4c27-b487-cc5733b7bae1 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Visual instruction tuning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b02df333-5030-4d93-a675-c96a714e8558 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images InfiMM-HD: A Leap Forward in High-Resolution Multimodal Understanding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04f92453-e9e4-4bf9-8fd9-dafccf4c7bdc · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66d49876-07ef-42ef-8be7-b7c6ef2ba733 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images INF-LLaVA: Dual-perspective Perception for High-Resolution Multimodal Large Language Model
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 683aaed2-2f3e-4711-80e6-a086e56633b6 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Learning transferable visual models from natural language supervi- sion
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bd70a414-2818-4b9c-a681-e26a7def89d4 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Divide and conquer: High-resolution industrial anomaly de- tection via memory efficient tiled ensemble
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 127f03cd-3421-4825-a581-56623e3d8213 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a15456b-100f-4342-8ffe-776f179f7449 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Fixing the train-test resolution discrepancy.Advances in neural information processing systems, 32, 2019
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 40b55d28-a80c-4212-86e1-64cb93d305b8 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images LLaMA: Open and Efficient Foundation Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54d99456-7761-4120-999a-633afa533ff2 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd65673c-bd1a-4940-8b85-dd4bc318fbda · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 260882ba-87f2-4e7b-a439-4db25da6134c · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images OS-ATLAS: A Foundation Action Model for Generalist GUI Agents
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa3342fe-f379-4b64-8168-773eb6a7445d · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Omg-llava: Bridging image-level, object-level, pixel-level reasoning and understanding
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cde3fcfd-ef63-43be-8f42-583d70769997 · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images Large Language Models Are Not Robust Multiple Choice Selectors
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 350fa9db-b1dc-409a-8387-d220a0c7aefb · outbound
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images MROVSeg: Breaking the Resolution Curse of Vision-Language Models in Open-Vocabulary Image Segmentation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e58942a-f5e4-4b74-93c2-ab7e7c0e8702 · inbound
LookWise: Knowing When and Where to Look for Fine-Grained Visual Reasoning in Multimodal Large Language Models A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 058f0829-90fb-46e2-aaf3-6ba1586c3262 · inbound
DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.