Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2405.18415.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:01:14.376214Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-22T18:26:55.252818Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d7c1eef8-82cb-4789-8adb-af3f9e23165c · inbound
De-biased Multimodal Electrocardiogram Analysis Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fb2b9bb-2758-4d91-bdc8-85307ab7a811 · inbound
VisGraphVar: A Benchmark Generator for Assessing Variability in Graph Analysis Using Large Vision-Language Models Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eafa2640-0500-469a-9c8d-5b60b542520c · inbound
DuetML: Human-LLM Collaborative Machine Learning Framework for Non-Expert Users Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a1675ab-2c43-475f-9b3b-d6bc147de72c · inbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 109
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 879d9d47-ff82-46aa-9669-be9f00d0af6c · inbound
Exploring Compositional Generalization of Multimodal LLMs for Medical Imaging Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 124
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7048e534-36ba-4009-97b0-65f362a3a0f4 · inbound
Enhancing Multimodal In-Context Learning for Image Classification through Coreset Optimization Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a00e4c4e-a295-454d-ac90-bf1040b28d75 · inbound
Benchmarking Large Vision-Language Models on Fine-Grained Image Tasks: A Comprehensive Evaluation Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c38dcdf1-237c-439e-94e0-91c601db2e2e · inbound
Single Domain Generalization for Few-Shot Counting via Universal Representation Matching Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22352f41-f8c3-4b2c-851f-73ded0b38e90 · inbound
GeoVision Labeler: Zero-Shot Geospatial Classification with Vision and Language Models Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4abf9151-01b7-4781-ac18-c573087a4c8d · inbound
AIGI-Holmes: Towards Explainable and Generalizable AI-Generated Image Detection via Multimodal Large Language Models Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 103
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f4deb52-f0f0-4b76-8427-1491df391ee8 · inbound
Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b47dfc0-6f42-40af-9456-f23975f858c7 · inbound
Filter-And-Refine: A MLLM Based Cascade System for Industrial-Scale Video Content Moderation Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9df4ce5f-c840-4476-becc-a370ef38184e · inbound
Decomposing Visual Classification: Assessing Tree-Based Reasoning in VLMs Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62803457-e2fe-4e8e-8a31-9fc12368c114 · inbound
VisualOverload: Probing Visual Understanding of VLMs in Really Dense Scenes Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 118e2245-8dfb-4c81-b2d5-eeebd7a5a3b7 · inbound
Unpacking Hateful Memes: Presupposed Context and False Claims Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f71b06a0-b4fc-496a-a091-da1324cc9c3b · inbound
Foundation Models for Astrophysics Why are Visually-Grounded Language Models Bad at Image Classification?
Reference 146
Source-reported events for the cited work
Unavailable: canonical work link unavailable.