Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T19:24:21.549089Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2509.09311.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T19:24:21.549089Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9883c8b3-ce2f-4d7d-ba90-b22152ec868c · outbound
Image Recognition with Vision and Language Embeddings of VLMs Getting ViT in Shape: Scaling Laws for Compute-Optimal Model Design
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79b454bc-263c-4864-bfb1-dc73f3bc623d · outbound
Image Recognition with Vision and Language Embeddings of VLMs Are we done with ImageNet?
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a130b777-9e34-4895-93cc-42df65db3df9 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Imagenet: A large-scale hierarchical image database
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a158f1a2-b443-42da-a6ab-6d59617aef43 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Eva-02: A visual representation for neon genesis
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f82c2987-ceb4-48f8-b189-b0354759ffde · outbound
Image Recognition with Vision and Language Embeddings of VLMs Renovating names in open-vocabulary segmentation benchmarks
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd80e3a1-1bf5-4e67-acb5-6911634527ee · outbound
Image Recognition with Vision and Language Embeddings of VLMs RADIOv2.5: Improved Baselines for Agglomerative Vision Foundation Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61a2475a-a115-4991-afc0-1744b5160b9a · outbound
Image Recognition with Vision and Language Embeddings of VLMs Openclip, July 2021
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03b350dc-7bd6-48b1-b7df-0cdd057b1163 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Berg, Wan-Yen Lo, Piotr Dollar, and Ross Girshick
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f61acd8f-110f-4bc2-83f9-bf00489db786 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Flaws of imagenet, computer vision’s favorite dataset
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ddf585f-62c6-434f-822c-ccb3333993be · outbound
Image Recognition with Vision and Language Embeddings of VLMs Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0070f8cf-d175-4caa-934d-1c5f9598f731 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Visual instruction tuning,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1c85b76-e1e6-4e7f-b84f-3a570739fec6 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Pervasive Label Errors in Test Sets Destabilize Machine Learning Benchmarks
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a15fc33b-2c4b-441a-abc2-ff9225f78a20 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Chatgpt: Optimizing language models for dialogue, 2022
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0fcf758-83ae-4563-850a-aaa7e5619bee · outbound
Image Recognition with Vision and Language Embeddings of VLMs DINOv2: Learning Robust Visual Features without Supervision
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 540d8450-2e48-4db0-9e36-8ad336aebf26 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Learning to name classes for vision and language models, 2023
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 878ad9ef-f2ab-48e1-8663-6300a301619d · outbound
Image Recognition with Vision and Language Embeddings of VLMs Learning Transferable Visual Models From Natural Language Supervision
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d8e7bdd-e8a1-4ef9-a9a0-4f0ddc0315b9 · outbound
Image Recognition with Vision and Language Embeddings of VLMs ImageNet Large Scale Visual Recognition Challenge
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cadf1b4e-088b-4dcc-a538-116ebe94679b · outbound
Image Recognition with Vision and Language Embeddings of VLMs Evaluating machine accuracy on ImageNet
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4258173-5baa-4ad9-9d49-88d1b5c7885f · outbound
Image Recognition with Vision and Language Embeddings of VLMs PaliGemma 2: A Family of Versatile VLMs for Transfer
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e3e8dc7-5b40-4a0c-8a6a-fb6ce18526d0 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb45ebac-444f-4468-9565-8b2156c2d286 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4966a22b-7840-4dbc-ab81-3d99222383a4 · outbound
Image Recognition with Vision and Language Embeddings of VLMs SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ccb2a77-e206-49d0-898d-6cd2b0d04462 · outbound
Image Recognition with Vision and Language Embeddings of VLMs From ImageNet to Image Classification: Contextualizing Progress on Benchmarks
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98fc8d27-3340-4fa4-b60e-9b4b440f9092 · outbound
Image Recognition with Vision and Language Embeddings of VLMs When does dough become a bagel? Analyzing the remaining mistakes on ImageNet
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c6f9b95-61e4-42ee-ba3e-65f622c2b323 · outbound
Image Recognition with Vision and Language Embeddings of VLMs ConvNet vs Transformer, Supervised vs CLIP: Beyond ImageNet Accuracy
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e414068e-e0e8-4937-abf9-1501886f0e84 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Pytorch image models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11799de6-8e27-453a-92f4-d13f1bc23230 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Self-training with Noisy Student improves ImageNet classification
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e96df1d-d9e6-4e7f-9663-d8f563d65d6d · outbound
Image Recognition with Vision and Language Embeddings of VLMs Leveraging cross-modal neigh- bor representation for improved clip classification
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11bd0161-0444-46c4-a228-89f9118114e2 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Sigmoid loss for language image pre-training, 2023
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6fc6248-da4c-4349-a485-ec227cf8cc84 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Vision-language models for vision tasks: A survey
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6e85d05-2b34-4bf3-95a5-50a12d6426be · outbound
Image Recognition with Vision and Language Embeddings of VLMs Tip-adapter: Training-free adaption of clip for few-shot classification
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b041f9d-4f58-4ac6-bf92-e96c1a93d2fe · outbound
Image Recognition with Vision and Language Embeddings of VLMs Conditional Prompt Learning for Vision-Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9172bc4a-4b87-486d-8d06-370a2c1e901a · outbound
Image Recognition with Vision and Language Embeddings of VLMs {class name}
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e4496f4-70f1-4cd2-ab1c-2ca998b2b319 · outbound
Image Recognition with Vision and Language Embeddings of VLMs EfficientNetV2: Smaller Models and Faster Training
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c372c09-a47f-49c3-aefa-25cedcfdfc32 · outbound
Image Recognition with Vision and Language Embeddings of VLMs Visual Instruction Tuning
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0db88a8-b69d-4bea-9f53-723bc2abef3f · outbound
Image Recognition with Vision and Language Embeddings of VLMs URL https://proceedings.mlr.press/ v119/shankar20c.html
Reference 8644
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.