Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T20:08:46.326666Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 8 inbound Pith citation observations for arXiv:2501.09425.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T20:08:46.326666Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:34:58.344234Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-14T00:18:29.517494Z
57 of 57 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 876019cf-4e11-484d-93b7-b82286f09e37 · outbound
Vision-Language Models Do Not Understand Negation Aug- mented reality meets computer vision: Efficient data gen- eration for urban driving scenes
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 76626a12-b8c4-4540-861e-c24ddca9336b · outbound
Vision-Language Models Do Not Understand Negation Effective conditioned and composed im- age retrieval combining clip-based features
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5d08414e-b6b6-492a-ac86-dbac81d7a1fe · outbound
Vision-Language Models Do Not Understand Negation FitCLIP: Refining large- scale pretrained image-text models for zero-shot video un- derstanding tasks
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation adc40a9b-4ad3-44b7-96a0-b98ec2b7ff5d · outbound
Vision-Language Models Do Not Understand Negation Conceptual 12M: Pushing web-scale image-text pre-training to recognize long-tail visual concepts
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19db9dd0-2961-4ed7-ad64-5bae5646be84 · outbound
Vision-Language Models Do Not Understand Negation Learning semantic segmentation from synthetic data: A geo- metrically guided input-output adaptation approach
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3e0cc5c6-38ab-4ba6-855b-8a80c4030f3d · outbound
Vision-Language Models Do Not Understand Negation The Llama 3 Herd of Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97d27b26-b7c2-45d5-ab5c-dbe351a3529a · outbound
Vision-Language Models Do Not Understand Negation The pascal visual object classes (voc) challenge
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 11613278-2c26-4b96-9ae0-c64f4c515b2a · outbound
Vision-Language Models Do Not Understand Negation Multimodal Autoregressive Pre-training of Large Vision Encoders
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43de77e5-ed52-44ba-8ed4-68d1bedef7ab · outbound
Vision-Language Models Do Not Understand Negation Datacomp: In search of the next generation of multimodal datasets
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e1117c67-0df2-4e42-a899-fb9d4b5c7bcd · outbound
Vision-Language Models Do Not Understand Negation This is not a dataset: A large negation benchmark to challenge large language mod- els
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 504dc068-fbb3-4026-b1d4-87f3c4e9d131 · outbound
Vision-Language Models Do Not Understand Negation Shortcut learning in deep neural networks
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6ad3e1d3-2a2b-4509-be2f-5b1e12a3db5e · outbound
Vision-Language Models Do Not Understand Negation SynthCLIP: Are We Ready for a Fully Synthetic CLIP Training?
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 863bde2d-1a1e-433d-bfcd-09a3e99ecd6e · outbound
Vision-Language Models Do Not Understand Negation Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fd484441-67ad-4796-b2bf-0cc61cb951c7 · outbound
Vision-Language Models Do Not Understand Negation Quilt-1m: One million image-text pairs for histopathology
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 73531a14-883c-4bee-842c-7c0fc3a33525 · outbound
Vision-Language Models Do Not Understand Negation Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a0594ace-e525-4638-b328-70feab20beaa · outbound
Vision-Language Models Do Not Understand Negation Generative models as a data source for multiview representa- tion learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6825bbb5-161f-4b27-8c15-85c4acb02932 · outbound
Vision-Language Models Do Not Understand Negation The power of negation in english: Text, context and relevance
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b40567b7-2ebc-44d1-a989-c1c016b6515d · outbound
Vision-Language Models Do Not Understand Negation Negation in syntax–on the na- ture of functional categories and projections
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e4b8857b-2a24-492a-8b01-221b1f34db22 · outbound
Vision-Language Models Do Not Understand Negation Naturalbench: Evalu- ating vision-language models on natural adversarial samples
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 51eb013c-e984-4379-8ec6-8ce2bcc29324 · outbound
Vision-Language Models Do Not Understand Negation Compre- hending and ordering semantics for image captioning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e2d73bba-3d18-4099-bf07-0007c3bfce36 · outbound
Vision-Language Models Do Not Understand Negation Cross-modal retrieval and semantic re- finement for remote sensing image captioning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9c371f82-ffe9-4c86-b92b-0595593f0e24 · outbound
Vision-Language Models Do Not Understand Negation Microsoft coco: Common objects in context
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c376ff10-6159-490e-b8f8-24a60b6a017e · outbound
Vision-Language Models Do Not Understand Negation A visual- language foundation model for computational pathology
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 92c26dc6-4b83-4553-a6ca-f3662ab335d6 · outbound
Vision-Language Models Do Not Understand Negation CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46d0c712-e0cf-4ceb-9a8a-be0410d8928e · outbound
Vision-Language Models Do Not Understand Negation Fine-tuning llama for multi-stage text retrieval
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e298a188-b5e0-4d97-b981-26f8a39358a5 · outbound
Vision-Language Models Do Not Understand Negation Crepe: Can vision-language foundation models reason compositionally? In CVPR, 2023
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 367d25f1-9eed-41c0-bc92-6b3d86ddcd67 · outbound
Vision-Language Models Do Not Understand Negation Simple open-vocabulary object detection with vi- sion transformers
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7bb530a6-5b4e-447e-983a-dc3ff370dd56 · outbound
Vision-Language Models Do Not Understand Negation Recent advances in processing negation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 98baac9d-107f-4d94-9a8e-fc3102221a5b · outbound
Vision-Language Models Do Not Understand Negation Effect of negation in sentences on sentiment analy- sis and polarity detection
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1e9c7f92-9608-4b9a-bbb0-8f78ca2618b7 · outbound
Vision-Language Models Do Not Understand Negation Clip-it! language-guided video summarization
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8835517b-641d-4731-9c1e-a04012018817 · outbound
Vision-Language Models Do Not Understand Negation Multi-Stage Document Ranking with BERT
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5869f8b3-04a3-4e44-a339-5be9332767af · outbound
Vision-Language Models Do Not Understand Negation Synthesize diagnose and optimize: Towards fine- grained vision-language understanding
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3e17222d-e85e-4157-8b69-7f9b531211b2 · outbound
Vision-Language Models Do Not Understand Negation On guiding vi- sual attention with language specification
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 877627b6-c6ca-4b13-89d2-df8a6aa91d57 · outbound
Vision-Language Models Do Not Understand Negation Learn- ing transferable visual models from natural language super- vision
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1a058905-4956-4332-b0bb-e67c043cadab · outbound
Vision-Language Models Do Not Understand Negation Denseclip: Language-guided dense prediction with context- aware prompting
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b2cf1896-10f9-4896-a3a8-551b340ce5b3 · outbound
Vision-Language Models Do Not Understand Negation Sentence-bert: Sentence embeddings using siamese bert-networks
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 986b4a43-0592-49e9-b54d-5fdb299a5c7d · outbound
Vision-Language Models Do Not Understand Negation High-resolution image syn- thesis with latent diffusion models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 850b4137-6f56-4bb0-8303-dfb1b0199aa7 · outbound
Vision-Language Models Do Not Understand Negation Clip for all things zero-shot sketch-based image retrieval, fine- grained or not
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1a841611-2507-45f8-b381-b5259404120b · outbound
Vision-Language Models Do Not Understand Negation LAION-5b: An open large-scale dataset for train- ing next generation image-text models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b1117bc4-936e-443f-b7f3-fcd2367020e7 · outbound
Vision-Language Models Do Not Understand Negation How much can clip benefit vision-and-language tasks? In International Conference on Learning Representa- tions
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 89029596-1604-4d2e-8baa-02b6076886a1 · outbound
Vision-Language Models Do Not Understand Negation Proposalclip: Unsupervised open-category object pro- posal generation via exploiting clip cues
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 15f7bafa-b741-49ac-964c-95d392e8eba8 · outbound
Vision-Language Models Do Not Understand Negation Cliport: What and where pathways for robotic manipulation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6b9ddda6-4d36-47c2-8059-4939e8fb9485 · outbound
Vision-Language Models Do Not Understand Negation Learn "No" to Say "Yes" Better: Improving Vision-Language Models via Negations
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6d9dcd2-4121-4365-b480-4979e54b9d75 · outbound
Vision-Language Models Do Not Understand Negation Stablerep: Synthetic images from text-to- image models make strong visual representation learners
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3cd798c9-eaeb-4969-820e-39ad73b97f21 · outbound
Vision-Language Models Do Not Understand Negation Learning vision from mod- els rivals learning vision from data
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bacd2688-c153-4551-9cd1-9d3ea0705ac0 · outbound
Vision-Language Models Do Not Understand Negation Expert-level detection of pathologies from unannotated chest x-ray images via self- supervised learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 505de08f-ebd5-4b72-bb76-e06e2100f58f · outbound
Vision-Language Models Do Not Understand Negation Language models are not naysayers: an anal- ysis of language models on negation benchmarks
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 57d0e0a7-a39f-419a-8995-ab824dc0e9ca · outbound
Vision-Language Models Do Not Understand Negation Msr-vtt: A large video description dataset for bridging video and language
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 134bce19-b0fc-4a7e-8e3a-d53033da6fed · outbound
Vision-Language Models Do Not Understand Negation Real-fake: Effective training data synthesis through distribution matching
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 28592488-7af3-4134-a2ea-ce92e7c9575c · outbound
Vision-Language Models Do Not Understand Negation When and why vision- language models behave like bags-of-words, and what to do about it? In ICLR, 2023
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 56c7ca48-e364-4fab-afa8-8f56e3500ae9 · outbound
Vision-Language Models Do Not Understand Negation Lit: Zero-shot transfer with locked-image text tuning
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3b85b920-c49b-45cb-b1c9-3d6d2ee36ec1 · outbound
Vision-Language Models Do Not Understand Negation Sigmoid loss for language image pre-training
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5bdf1839-48ec-41c2-8e4d-f6c92c722db5 · outbound
Vision-Language Models Do Not Understand Negation BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31792e35-f68e-4868-919a-52efb1ade9b3 · outbound
Vision-Language Models Do Not Understand Negation Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation eebb94a9-a692-433c-88eb-a37f3f06fbda · outbound
Vision-Language Models Do Not Understand Negation Yes.” over “No
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6ff600fa-2015-4ed7-996a-b3c7c3a277ae · outbound
Vision-Language Models Do Not Understand Negation Unresolved cited work
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1367e721-a6e2-480a-80d9-672a22dd9629 · outbound
Vision-Language Models Do Not Understand Negation Unresolved cited work
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6a51e034-03ea-4726-a7a1-89714e6234a0 · inbound
TNG-CLIP:Training-Time Negation Data Generation for Negation Awareness of CLIP Vision-Language Models Do Not Understand Negation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93ca8fa6-82d7-43c1-853c-34dbd8976c31 · inbound
NegVQA: Can Vision Language Models Understand Negation? Vision-Language Models Do Not Understand Negation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af5a841d-b63e-4378-9185-5f06f40dddbf · inbound
Decomposing Visual Classification: Assessing Tree-Based Reasoning in VLMs Vision-Language Models Do Not Understand Negation
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65823167-c3b5-467e-9557-9740e1845312 · inbound
Disparities In Negation Understanding Across Languages In Vision-Language Models Vision-Language Models Do Not Understand Negation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8f28e7c4-ea6d-489f-935b-6c6cc2a6296d · inbound
Trace Mutation in Human-LLM Dialogue: The Transcript as Forensic and Mitigation Surface Vision-Language Models Do Not Understand Negation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2471455d-3f34-45b8-8239-b49ecff50ba2 · inbound
CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs Vision-Language Models Do Not Understand Negation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 10e8c895-0807-4b6d-a917-130295408816 · inbound
Uneven Evolution of Cognition Across Generations of Generative AI Models Vision-Language Models Do Not Understand Negation
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 64c36674-3a95-4912-9f7c-3014d8fc0d3a · inbound
Learning to Detect Cross-Modal Negation: An Analysis of Latent Representations and an Attention-Based Solution Vision-Language Models Do Not Understand Negation
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.