Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T12:42:01.307694Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2412.13947.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T12:42:01.307694Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
30 of 30 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 57f7fd7a-a03f-4465-8301-1340777590ff · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Food-101 – mining discriminative components with random forests
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 5785fb7b-5673-467b-91cf-9c21af468b69 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Crossvit: Cross-attention multi-scale vision transformer for image classification, 2021
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 71b3f697-bf71-48ab-9d68-97a77c8f1807 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Ovarnet: Towards open- vocabulary object attribute recognition
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1ec9753c-8f5d-43ff-bfe4-c423ccd0c27d · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Multi- modal classifiers for open-vocabulary object detection
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 5f26d611-452d-4b0a-a489-45fc7211787c · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Novel dataset for fine-grained image categorization: Stanford dogs
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a32dff49-1dc5-4eac-8f82-d7d0d9f189b8 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition 3d object representations for fine-grained categorization
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6b7d32d4-d92f-4e24-bcf4-8a35e5c7d59c · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Descriptor and Word Soups: Overcoming the Parameter Efficiency Accuracy Tradeoff for Out-of-Distribution Few-shot Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 49d266cc-7619-4fd0-b5b7-67c81d3d18cb · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Feature pyramid networks for object detection, 2017
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 64612c1f-0577-43dd-be12-ab3b7620beba · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Visual Classification via Description from Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e1d83d8-e508-4672-ba79-8f9f2c40b04a · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Automated flower classification over a large number of classes
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 90582daa-7a69-415f-815c-9bb25566ef98 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Parkhi, Andrea Vedaldi, Andrew Zisserman, and C
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2822b140-c3bd-4417-ae51-b7e881f74c26 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition What does a platypus look like? generating customized prompts for zero-shot image classification
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ca6f0a4b-d1be-42e8-88f5-e1070f37daf3 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition What is the limitation of multimodal llms? a deeper look into multimodal llms through prompt prob- ing
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 044e5b52-e0b7-455d-9909-db804f121917 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Learning transferable visual models from natural language supervi- sion
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43d5c935-1826-40df-a087-c8ff82670836 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Paco: Parts and attributes of common objects, 2023
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 55b7f611-7b0d-4623-afe3-5f3d02067bb5 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Imagenet-21k pretraining for the masses
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 11c97817-e1b7-414f-8fd6-81dc73e2ebff · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition U-net: Convolutional networks for biomedical image segmentation,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a140945d-003c-4216-99e9-edbaa9c998b0 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Waffling around for performance: Visual classification with random words and broad concepts
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation cce4af3c-f0bc-4f1f-9cba-31b08cb50ec8 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Conceptual captions: A cleaned, hypernymed, im- age alt-text dataset for automatic image captioning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8c46673c-c832-4e37-93dd-c0fa975f4fdc · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition When do we not need larger vision models?, 2024
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2d682432-b4a2-4a7c-ba9b-9acc6e281ca4 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Alpha-CLIP: A CLIP Model Focusing on Wherever You Want
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d9f8e97-8c4d-443b-8e2a-4e7701721f1f · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition ArGue: Attribute-Guided Prompt Tuning for Vision-Language Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 97b0b32b-9eaa-4840-96f6-0807da876d59 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Efficient object localization using convolutional networks, 2015
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a569b525-0438-41e8-9a15-979427efe460 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0b8e864f-4e75-4d24-b9b1-3ca99974aef7 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Learning concise and descriptive attributes for visual recognition
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5059604-3cba-49d4-84af-d05be2df3d68 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Focal self-attention for local-global interactions in vision transformers, 2021
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 60c42623-3596-4e84-b6e0-e0f2cd0384e3 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Language in a bottle: Language model guided concept bottlenecks for interpretable image classification
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 70ceeb75-2be3-491a-8560-29ff5d6a44c1 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition When and why vision- language models behave like bags-of-words, and what to do about it? In The Eleventh International Conference on Learning Representations, 2022
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 47fb1570-e149-4915-960e-25642ac14be4 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Prompt, generate, then cache: Cascade of foundation models makes strong few-shot learners
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation cf47ab0c-65a8-45c7-9986-e6f037f5e101 · outbound
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition Vl- checklist: Evaluating pre-trained vision-language models with objects, attributes and relations, 2023
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
No inbound Pith citation observations are available.