Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-19T11:12:41.130806Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 100 of 158 outbound references and 1 inbound Pith citation observation for arXiv:2506.09082.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-19T11:12:41.130806Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-14T22:19:38.973091Z
A source-named dated measurement, never combined with another source.
Source: cited_works
100 of 158 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 437a8071-a879-40b2-94a4-b357927457f6 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b7803521-163e-40d4-9ffa-3c0623ce07ba · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Foundation models defining a new era in vision: a survey and outlook
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8286d9f7-720f-4b99-9349-4e1d4a0063a6 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Eureka: Evaluating and Understanding Large Foundation Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 08a5b9b2-fb20-4a66-9bb4-73927304f280 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Scene text visual question answering
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 68c006e2-0e69-42f3-909e-e0f73ecc247d · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models On the Opportunities and Risks of Foundation Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6ce07b22-b15f-4c6d-85eb-bba5f39b38c4 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Emerging properties in self-supervised vision transformers
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3999625d-bc2c-457b-adac-093c94514c8e · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Decomposing complex visual comprehension into atomic visual skills for vision language models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 92cc5132-7175-4c94-a15a-9ce0adad6381 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Sharegpt4v: Improving large multi-modal models with better captions
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 279d5f37-ea74-49f7-8507-b45140b68e69 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Segment Anything Model (SAM) Enhanced Pseudo Labels for Weakly Supervised Semantic Segmentation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 720ed980-043f-4937-95ef-8c99458adeaa · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6fd232ec-337c-4e50-9796-f460cee33860 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models On Domain-Adaptive Post-Training for Multimodal Large Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dd796da5-a3e7-4af0-afcd-154ddd1da6c1 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models MegaCOIN: Enhancing Medium-Grained Color Perception for Vision-Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 348c2c32-8dab-46db-9e03-282ad73bc4f9 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Palm: Scaling language modeling with pathways
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e04d3eea-4b88-4095-b940-aa224aff8e83 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Prompt-CAM: Making Vision Transformers Interpretable for Fine-Grained Analysis
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4c968fa0-f649-4219-90fb-859a301ebe34 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Cimpoi, S
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 114f337c-b6ee-4faf-9bbd-9834a1597742 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Humanvlm: Foundation for human- scene vision-language model
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0f0ea91a-202e-4131-8681-5cf6fd4cb4dd · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 28701936-19fc-490e-a422-603c0e74bd02 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Shape and texture recognition in large vision-language models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4edbc07f-e1d9-452f-b859-a0d647a0a8a2 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models There is no SAMantics! Exploring SAM as a Backbone for Visual Understanding Tasks
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 14ee273e-135e-47aa-9a9f-7a009c3d6e88 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models MLLM-SUL: Multimodal Large Language Model for Semantic Scene Understanding and Localization in Traffic Scenarios
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8f3a1ec8-383f-4a72-92bd-bacd9751eb5e · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Multimodal Autoregressive Pre-training of Large Vision Encoders
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c7e1e9ca-418f-4a91-ae97-22de8b9f90ac · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Blink: Multimodal large language models can see but not perceive
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6b402099-349e-4aec-86e3-ab2dafd7eae3 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Can We Talk Models Into Seeing the World Differently?
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1349b0e5-2b4d-4536-acce-60f3a83c11c7 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Vision meets robotics: The kitti dataset
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 84c79353-0db4-4b05-8878-4e830072408b · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Battle of the backbones: A large- scale comparison of pretrained models across computer vision tasks
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 71a977c4-dd80-4fb7-a8ec-e2e7c1c92d02 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Making the v in vqa matter: Elevating the role of image understanding in visual question answering
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dab1989f-8231-43f9-901a-70790b83010a · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Making the V in VQA matter: Elevating the role of image understanding in Visual Question Answering
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 761f0e06-25f8-4d87-8ab4-5b96ba627dcf · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models The Llama 3 Herd of Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a23c15d7-6a06-4f51-919b-64c884ff4908 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Bioclip 2: Emergent properties from scaling hierarchical contrastive learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 938290bf-92e8-4df5-b001-a3d7ad78e1af · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Lvis: A dataset for large vocabulary instance segmentation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 10c56288-30be-4fef-95c6-3fbda059ac18 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models A survey on vision transformer
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 581289eb-d9d6-4aba-9543-f16e2f48ab58 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models RADIOv2.5: Improved Baselines for Agglomerative Vision Foundation Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 587cdc6e-f774-46de-b132-eaef90308a44 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Parameter-efficient transfer learning for nlp
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 12c3d566-fafb-456e-9f18-c08253b40c1e · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Drone-based object counting by spatially regularized regional proposal network
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ec97031c-663a-4160-872b-faab2438c3e8 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Lora: Low-rank adaptation of large language models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f09d3e57-92b7-4da0-9496-38c4d707aeac · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models A Survey on Evaluation of Multimodal Large Language Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2f3652a3-a196-47f5-8dfe-c0797ad7d866 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models T2i-compbench: A comprehensive benchmark for open-world compositional text-to-image generation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3b18747c-e5c2-4e08-b29d-f0194bff28b1 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0adf2cfe-40d1-4484-a6cd-278204a07da3 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Position: The platonic representation hypothesis
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f9d7873b-8f6f-47f2-8402-09bbd1246c7f · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Is 'Right' Right? Enhancing Object Orientation Understanding in Multimodal Large Language Models through Egocentric Instruction Tuning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c088ac66-75f3-42bb-8ec2-5d2aeabd789f · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Transformers in vision: A survey
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 90af026b-3dbf-4772-bc01-fbf76c449e15 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Mllm-compbench: A comparative reasoning benchmark for multimodal llms
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d60c2d93-3d20-4f62-b0e5-6a5ccfc3a5f6 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Finer: Investigating and enhancing fine-grained visual concept recognition in large vision language models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c4205e3f-d4cf-4898-a504-05a1c8e043fc · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Segment anything
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2780cd61-b304-4fe8-b521-a2eb3f05ae84 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Diagnostics- llava: A visual language model for domain-specific diagnostics of equipment
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 36929f7f-0e8c-4f36-93ac-32430d2e1063 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Kylberg texture dataset v
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 91724a07-6f19-46b8-862a-18de7a7caa48 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models LLMCount: Enhancing Stationary mmWave Detection with Multimodal-LLM
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2f785da9-4afd-42d3-9533-0324a7c94836 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models EALD-MLLM: Emotion Analysis in Long-sequential and De-identity videos with Multi-modal Large Language Model
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4d6fc00e-e34a-432f-ad90-da15443ca303 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Video crowd localization with multifocus gaussian neighborhood attention and a large-scale benchmark
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 92e62d28-3d9d-4cbc-816d-387002fdbab4 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Object detection in optical remote sensing images: A survey and a new benchmark
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bee65464-8792-4404-b280-dd0755252bec · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Reliable crowdsourcing and deep locality-preserving learning for un- constrained facial expression recognition
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d4b948b0-1000-47c2-baf9-0a7d7665ae36 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models OpenVision: A Fully-Open, Cost-Effective Family of Advanced Vision Encoders for Multimodal Learning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 78c8ddaa-47de-4843-8a1b-958b2fea1532 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Visual Large Language Models for Generalized and Specialized Applications
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 235fe6d8-1aa0-4bd6-a9f3-78d964c14141 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Expression analysis based on face regions in real-world conditions
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2c457a82-0bc7-4ae8-a8da-7d3184341a90 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models A comprehensive survey and guide to multimodal large language models in vision-language tasks
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5634bf49-ca72-46da-807e-9fb3bc46e823 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Visual instruction tuning
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation eaa0e79a-a998-4364-b3c5-60d2f61f302b · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Ocrbench: on the hidden mystery of ocr in large multimodal models
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 629a158d-0e98-4440-aa15-752af39e5a47 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Cvpr 2020 continual learning in computer vision competition: Approaches, results, current challenges and future directions
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7959ce49-47a4-46f1-92a9-9abd5f30d7b4 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models The development of the cie 2000 colour-difference formula: Ciede2000
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9d991c58-b310-4351-8e5e-b96c42ec8568 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Fine-tuning is fine, if calibrated
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 90165c19-b3b1-419c-a4fb-96263d7bc45a · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Batch-level Experience Replay with Review for Continual Learning
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bc806bbd-a42e-45c5-8bd5-0929314470f8 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Online continual learning in image classification: An empirical survey
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 63a3c159-cb68-4ac3-9f4d-fd6dc5045147 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Supervised contrastive replay: Revisiting the nearest class mean classifier in online class-incremental continual learning
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 258241d7-3370-43ee-a30f-ed82d10371bc · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Lessons and insights from a unifying study of parameter-efficient fine-tuning (peft) in visual recognition
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3e61a3dc-c11f-49cd-b0f1-a526db319c9c · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Fine-Grained Visual Classification of Aircraft
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 58126c4a-9a57-4fff-acee-edffef7b7dbe · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models The kth-tips2 database
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c52d5607-8014-40b7-9097-c57445bb0ab4 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Hierarchical interpretable vision reasoning driven through a multi-modal large language model for depth estimation
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c47b2d68-de1e-4a29-9b69-d2fcb5abfb89 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Mishra, K
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 81a9079f-32f9-43b1-9692-5c1966ee7e75 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Moments in time dataset: one million videos for event understanding
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f407c0bc-cd9f-4746-b490-ae7dfeeaf049 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Intriguing properties of vision transformers
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2bda31ff-adde-4223-ae84-e3c5287a7ff7 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models DINOv2: Learning Robust Visual Features without Supervision
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2d451229-ee15-4655-8d1c-ddbf50c64bb7 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models A simple interpretable transformer for fine-grained image classification and analysis
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 288c74a6-29f8-467e-bbfc-8561dc9f974f · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Learning transferable visual models from natural language supervision
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c5a373b1-63b1-49f8-8f15-9c4fe9ceb839 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 61e4a709-cd34-4325-abbb-3a485ace0b03 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Learning to count everything
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ee6569a7-5989-4dbb-8918-d7ea671eb1f5 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Am-radio: Agglomerative vision foundation model reduce all domains into one
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0eb6ed2e-cad1-428b-ba80-9e926668f3ba · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Generalized intersection over union: A metric and a loss for bounding box regression
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 529d4ca2-5bc1-4922-8da1-c1ca9a9287dd · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Object detection with multimodal large vision-language models: An in-depth review
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0e2a7e22-46de-41bf-8050-c79d6cd5a24d · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Objects365: A large-scale, high-quality dataset for object detection
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5fc250af-a27a-4e9c-aa01-363cd7d71db9 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Online class- incremental continual learning with adversarial shapley value
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a837bc62-ca33-4516-92ff-35b87f2b68ec · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Indoor segmentation and support inference from rgbd images
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bac7d0bb-d7cf-4d9a-91cd-326c1898d884 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Multi-modal large language models are effective vision learners
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a4c18339-ab8b-43aa-830f-6c9bcba4d52a · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Grounding multimodal large language models in actions
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d36839aa-c845-46ea-a942-ba82fe55cdfc · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Temel, J
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f88cff89-3b75-46f4-8f65-0845fa473fc7 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Kth-tips dataset
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dc319ec0-c93b-4a95-9bd0-8a77737afcc8 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Semantic segmentation using vision transformers: A survey
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation df647130-248f-4f48-8cac-d6ddab8075af · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Cambrian-1: A fully open, vision-centric exploration of multimodal llms
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ac95146e-21d8-4074-b945-b0c5c5d71d14 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Eyes wide shut? exploring the visual shortcomings of multimodal llms
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d3ed9224-d83b-4860-9309-d5ac4fec49e8 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 334a8d07-268e-4315-b5cc-672a93d8d441 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Holistic transfer: Towards non-disruptive fine-tuning with partial target data
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 96ca5726-1e9e-46b6-a194-c8a2ec7c782b · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Visual query tuning: Towards effective usage of intermediate representations for parameter and memory efficient transfer learning
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e962a679-52ab-4177-a821-94f2673b8beb · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Benchmarking representation learning for natural world image collections
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2c669783-fe6d-4e56-ba56-abe50041330b · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models COCO-Text: Dataset and Benchmark for Text Detection and Recognition in Natural Images
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 12569ffd-69b6-444f-a391-b50780b989e7 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models The caltech-ucsd birds-200-2011 dataset
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 28a70ba5-ed18-4be6-b4e3-910923c938c0 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Harnessing multi-modal large language models for measuring and interpreting color differences
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3cad97b3-0415-4738-a46f-b2c110d9613b · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d5c2b3b6-725a-4bcb-b5d1-1a95e099416d · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Visionllm v2: An end-to-end generalist multimodal large language model for hundreds of vision-language tasks
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ff3c0eb4-d383-4fb5-86d5-06f5b357602a · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Multimodal large language models: A survey
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8a9d4413-ca5d-4590-968c-e60a853897f6 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models Visual Compositional Tuning
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 61dff8f3-c356-46eb-9b8e-a662c59e2651 · outbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models ConceptMix: A Compositional Image Generation Benchmark with Controllable Difficulty
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5e2c7784-6a4f-44c5-9b7c-acb1a08a2ebb · inbound
Revisiting Model Stitching In the Foundation Model Era AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.