Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:58:06.130584Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 1 inbound Pith citation observation for arXiv:2505.08468.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:58:06.130584Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:02:45.290988Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T12:02:48.057736Z
46 of 46 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 911d7246-3d83-439d-bbed-4e24168cf271 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5310a3c-1ac4-44d4-b42c-e48ce9ee47be · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7ac8ac5-2587-44d8-9baf-4058724cd976 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Qwen2.5-VL Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69f8eb1b-f8fa-4510-834f-20c747f4c4a8 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? PaliGemma: A versatile 3B VLM for transfer
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 882e82a6-4310-4bbd-a09b-6de2c6f1d089 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d728a538-1e8a-4dd6-84bd-1e43c385483c · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91ea524b-07b0-4780-a018-f30c03f490f2 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be86b068-3cbe-4d36-8fce-ceaf4b3abb3c · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? A Survey on LLM-as-a-Judge
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21fd0795-0030-4042-95fd-8b3ba8e44139 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? ChartLlama: A Multimodal LLM for Chart Understanding and Generation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cdcd40c-403d-4b0a-98c6-57d4ce1f8550 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Hoque and M
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ffb56156-cb08-417c-85c7-bf991837908b · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 44c7f213-3f1e-4290-b67c-a900d10dcc67 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a36e0e2-ad8a-4ae4-95b6-fa5e1d0950b5 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Mistral 7B
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b153f563-57fa-4131-b90b-41326f814989 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80b714d6-bf15-4dde-91fe-d98e1739be0f · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? FigureQA: An Annotated Figure Dataset for Visual Reasoning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c87e1fe9-e02b-449f-bbe2-52655b0fa87c · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation edf98761-dc47-4ea9-a969-a1ec0f721605 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c78bcd9-93ab-4a71-8524-85ed308c8067 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b32b164e-97b8-4be4-9a96-83d50a814fb7 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01b8dcde-723f-4ef2-b8dd-fe2312499b51 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation df86cadb-78b7-4efe-90e4-e53eb8598e5a · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1e2fc81-1c6f-41c5-a5a2-4c78e6d2336a · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24891392-0bbb-4fa0-beb1-b2f55301937d · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f50d615-9781-47bb-a4d8-fce889e197a4 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72fefe5d-a095-496f-af1e-9c1b2ca1c032 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 436205ab-49df-48ac-89c6-7ca65dd5cd4d · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3bd1e347-76a4-45f8-aecf-2390577300df · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d1bc2630-7b9e-439b-b684-6a22cc082512 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 27859357-9ebd-4609-b14b-ccd9e942291a · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? GPT-4 Technical Report
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 860e3d6c-1ab6-46c8-b985-3922da4d0dc1 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Tahmid Rahman Laskar, Md
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8507f25a-4403-4ba6-a37d-679f67fd02ea · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cc6bca43-bac0-4e0f-8457-a8e607d929b4 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Tang, Angie Boggust, and Arvind Satyanarayan
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 18b58b9c-f4e2-45db-a138-fae3e206dc21 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Gemini: A Family of Highly Capable Multimodal Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ecd3fbc-bed7-46a1-8f08-6a08785fafba · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3deb986-72d4-4420-86cb-f567bde32c2a · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0381991f-33b4-4d98-b74a-379a32734184 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? LLaVA-Critic: Learning to Evaluate Multimodal Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba27bc36-7c99-4308-afdd-6a82dec2b5b7 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7a8651a-0138-4793-90e0-6830be302215 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f841f93-ca19-418a-8f90-a1e425039be0 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52f49ded-e6fa-4408-a823-d99b8446e625 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? TinyChart: Efficient Chart Understanding with Visual Token Merging and Program-of-Thoughts Learning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd943d81-dbec-4632-9b13-ba85c9954024 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? A Survey of Large Language Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 348f4967-1323-4c34-8cd2-19619d73d37d · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc318073-01fe-472f-b1a5-d9194d15ef0a · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? URL: " 'urlintro :=
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26aa9106-fbab-4363-9da1-a37dc869d716 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? write newline
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afc087d4-7c45-4c60-8921-77849caef84d · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? online" 'onlinestring :=
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30ff7b4f-832e-4220-a2b0-a7f133e46cc1 · outbound
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? write newline
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a2ef83a-53bf-4b90-80a8-e7d056c11299 · inbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning?
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.