Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:01:07.332466Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 81 of 81 outbound references and 4 inbound Pith citation observations for arXiv:2506.17267.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:01:07.332466Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T05:51:29.406876Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T10:09:44.501635Z
81 of 81 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation db04e06f-f11b-45e7-a857-2b6f14e0b7e2 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfbd2932-fb48-49b2-b13d-92c274c9d958 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning A Survey of Vision-Language Pre-Trained Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37109915-1310-47d6-9543-ff27ed9b2992 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Flamingo: a visual language model for few-shot learning,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcc8c1a0-c985-4ed2-b869-d631b3815a97 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Causal Inference with Large Language Model: A Survey
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60fea556-4575-45ff-9b62-2d850b2a616b · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning CELLO: Causal Evaluation of Large Vision-Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99555b6c-afd1-4d6d-8953-d6324ebaf0e1 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Learning Transferable Visual Models From Natural Language Supervision
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ff0d1b4-07be-48bd-ae98-24a0b4f88ee3 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Measuring progress in fine-grained vision-and-language understanding,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8300da25-7d91-4df9-abed-eaece9b8617e · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Synthesize, diagnose, and optimize: Towards fine-grained vision-language understanding,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e3947761-44ac-4821-adb9-1a20a4b3d0b9 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Finer: Investigating and enhancing fine-grained visual concept recognition in large vision language models,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7d2999e9-fa44-4ff1-b40a-137410a2ded7 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Benchmarking zero-shot recognition with vision-language models: Challenges on granularity and specificity,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ab451ce1-26b1-4cb7-81d3-32137ef45aaa · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Vilta: Enhancing vision-language pre-training through textual augmentation,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b637aff9-66da-4a03-bcb8-bc90540662f9 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Towards vision-language mechanistic interpretability: A causal tracing tool for blip,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a1d558dd-dfec-4602-8c5c-0d506ba8ad8f · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning What matters when building vision-language models?
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b9f42f80-8a62-4c16-b903-fd65508ef49a · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Fine-grained alignment for cross-modal recipe retrieval,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0c8a2ede-f477-4e9f-bc3e-ac43428afc04 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Localized triplet loss for fine-grained fashion image retrieval,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 110a4311-512f-4704-9fe0-b9aaca3441f3 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning TripletCLIP: Improving Compositional Reasoning of CLIP via Synthetic Vision-Language Negatives
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3c99b13f-56ef-4af6-9ab3-ca42fbd4ccb9 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Facenet: A unified embedding for face recognition and clustering,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bbcaeb3-aba4-445b-8ba4-599906698382 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning CPL: Counterfactual prompt learning for vision and language models,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f2cdb927-37fb-41da-80bd-bbd97f3162c2 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Attention Is All You Need
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b666d97b-07a6-4e8f-afd5-494e7f79f3db · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning A survey on evaluation of large language models,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba152358-f5c9-440e-9f5e-efe7a008b601 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning A survey of visual transformers,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 846e6f77-4295-47f2-ba90-0ee581c751d9 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning A simple framework for contrastive learning of visual representations,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4f02ba1-a538-42dd-8408-84b1e0268399 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Understanding contrastive representation learning through alignment and uniformity on the hypersphere,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 02c019f0-7f4e-4b7e-b6b5-d13924477716 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b9e9e6a-efdb-4d44-a706-1028d67ab742 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Cogs: A compositional generalization challenge based on semantic interpretation,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 43a53351-d77c-4251-b4fd-cea102c36b13 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Learning what makes a difference from counterfactual examples and gradient supervision,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 08b38393-f8f0-44a3-b187-ce2bd65a3274 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Teaching clip to count to ten,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e06c7165-c2c9-4efa-b367-550c11e41768 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning DISCO: Distilling Counterfactuals with Large Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bad5841-e5b9-4482-a9f0-64aaf70d9a71 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Counterfactually measuring and eliminating social bias in vision-language pre-training models,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e09ab765-4ce8-4e63-9431-0f3250b93549 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Counterfactual attention learning for fine-grained visual categorization and re-identification,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation af420799-f337-4ec0-b5f4-19b03960bb1c · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Counterfactual samples synthesizing and training for robust visual question answering,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5733a09a-3318-489a-98d1-376fb7498080 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Qwen2.5 Technical Report
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4593c00-6d8a-4f6d-aa1d-a6a6c4f83ef0 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Conceptual 12M: Pushing web-scale image-text pre-training to recognize long-tail visual concepts,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fda6020a-f7a7-4d80-8212-4c0c609522ea · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image captioning,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 060bdcea-5d8c-4cec-9572-47b5edeed4be · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Microsoft COCO Captions: Data Collection and Evaluation Server
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efa34ba6-5906-48d1-9a18-de51b7693da2 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning ConMe: Rethinking Evaluation of Compositional Reasoning for Modern VLMs
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 866c9125-f995-4dfe-bd87-feb0efc052a2 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning When and why vision-language models behave like bags-of-words, and what to do about it?
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e0bd20ed-6479-47b0-bb65-0026b8d5f9d2 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning VL-CheckList: Evaluating Pre-trained Vision-Language Models with Objects, Attributes and Relations
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f12b6b44-3b06-4d04-a7a4-ca0e8df4fa74 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning ImageNet Large Scale Visual Recognition Challenge,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 799bc5f2-8288-4427-b732-aea50db657ef · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ed03be52-36d3-4416-990c-4cafb667f01d · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Contrasting Intra-Modal and Ranking Cross-Modal Hard Negatives to Enhance Visio-Linguistic Compositional Understanding
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c6d5e468-171b-47c6-a514-c88c7e26d27c · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Structure-CLIP: Towards Scene Graph Knowledge to Enhance Multi-modal Structured Representations
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 036de79c-a746-4d6b-8d94-91e7d50d6c8f · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Improved Baselines with Visual Instruction Tuning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 556d253d-ee8b-4367-bd2b-bd573994ec36 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3d19d3a-9461-4759-90de-978b5b273bc3 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3838e51a-9024-44e9-8b6c-4840d6b2b301 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d6eedb6-ee1e-4ec7-a636-f31252676b0a · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Evaluating Object Hallucination in Large Vision-Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8903ba8e-e6da-43b9-b714-2c24b5ee9f14 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce3d1c0b-f348-4bdc-a973-23c90907feb9 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Vision-and-Language Pretrained Models: A Survey
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f46b262-c841-40b7-afca-bda312bc53a6 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Exploring the frontier of vision-language models: A survey of current methodologies and future directions,
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04362e3a-ca1b-4aa9-b3c2-5162f4df9c23 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Vision-language models for vision tasks: A survey,
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1064166d-50a6-44ac-a19e-7dead8da82db · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Counterfactual vision and language learning,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5256965e-acca-49f2-83fb-b036e5e749f4 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Counterfactual reasoning for multi-label image classification via patching-based training,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 75ad03d2-c6ca-4845-a852-39b0a661e028 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Causal Graphical Models for Vision-Language Compositional Understanding
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f4e29da0-3367-43e2-b371-50cbf99390a1 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning CPL: Counterfactual Prompt Learning for Vision and Language Models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 82b5f82b-87b4-4f61-8a8e-38e472eac53b · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Counterfactual Visual Explanations
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e4108df-7b85-443d-af3d-3ff07f794393 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Counterfactual Attention Learning for Fine-Grained Visual Categorization and Re-identification
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation af0ade69-265d-4dfe-87f4-5bac436fca09 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 38178384-683d-4c49-b0d8-2d9434e28870 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Unresolved cited work
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b4b2c7b5-42f4-4330-97b6-da6084a79ff5 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Rules: • Modifyonly one thing(either one attribute or one causal link)
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a125a548-442b-48f1-8e97-5a7935069a43 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Example Input: A young woman holding a racket hit the ball, and the ball flew outward
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 98563f15-079f-4b01-877a-75ae0aa6db83 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Unresolved cited work
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 87f61e64-b74d-41b0-90d6-958147e1b9ef · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Output (causal):
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b3fa6ee5-d2a5-4657-a4b6-db4745d88fe1 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning dirt” with “paved roads
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 37f1a45c-cfb8-476f-877b-b90457564489 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Unresolved cited work
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b49d51f6-6c25-4ba7-827c-78703bf04ab2 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Unresolved cited work
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation efcd3bf2-24ca-48b3-95c4-1d7b15b217ad · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning causal decision points
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1a1cde11-e036-45e2-b861-38daf8f12ede · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Unresolved cited work
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5504e3cd-2e3b-4cfa-854b-3e4a660c462e · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Unresolved cited work
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 72662f1d-9f16-4588-bd8f-3ce329cae31a · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Unresolved cited work
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b4b55706-9edf-4850-8635-25e55d99d1eb · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Two people on motorcycles riding them
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ef86dff8-a5a1-4677-9571-c66da1bcdad5 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Unresolved cited work
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c2267686-b567-446f-825c-7242bcfec410 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Unresolved cited work
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dc45ff25-b0fd-42b3-9c54-1e74ba3d9c7a · outbound
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5e3b4834-cbbe-463d-ad1c-215fa2526338 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning parallel realities
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 443bbec6-8f01-47e4-a298-1b68d0e77f84 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning These are employed in Lcsd to help the model learn semantic scene boundaries
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7aa85fba-ece3-4611-85b4-63191fb81046 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Unresolved cited work
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 527bc56b-0dc5-4d77-bf56-110aaa7a9aa8 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning Unresolved cited work
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 39db48cf-7edf-4131-b359-6611f6b84ee1 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning kicking a ball
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d6f96fae-5544-4683-a0ec-c5087db8e7fc · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning COGS: A Compositional Generalization Challenge Based on Semantic Interpretation
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f17ba9ae-74f7-42d7-bef8-edc49c67d258 · outbound
CF-VLM:CounterFactual Vision-Language Fine-tuning What matters when building vision-language models?
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed3a27c0-7fbc-410e-b5e1-f1305cb5945a · inbound
OSC: Cognitive Orchestration through Dynamic Knowledge Alignment in Multi-Agent LLM Collaboration CF-VLM:CounterFactual Vision-Language Fine-tuning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f99d24fa-3e34-4deb-bb1e-d630323c6d50 · inbound
When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs CF-VLM:CounterFactual Vision-Language Fine-tuning
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e43294dc-c047-41b0-97e7-665272a05f21 · inbound
CFPO: Counterfactual Policy Optimization for Multimodal Reasoning CF-VLM:CounterFactual Vision-Language Fine-tuning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 03766e84-061f-41f4-9881-13016b986e34 · inbound
SIVA-RL: Sensitivity-Invariance Visual Alignment for Multimodal Reinforcement Learning CF-VLM:CounterFactual Vision-Language Fine-tuning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.