Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T10:24:06.917263Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 100 of 115 outbound references and 3 inbound Pith citation observations for arXiv:2412.00142.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T10:24:06.917263Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:02:42.424201Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-16T06:17:26.610695Z
100 of 115 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ac2bcad2-f61d-4522-8da8-0bb0092f96d9 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b0af4f5-1f9b-4e94-ae4e-48600d8aca7f · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Vqa: Visual question answering
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c723246f-b926-4169-b19a-6ad76d6bbf5e · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9217092-aa73-4eb3-8237-d351cacfe238 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Autoencoders, unsupervised learning, and deep architectures
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2f0c454-b8df-4a3a-b022-b71a48b53d09 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features When XGBoost outperforms GPT-4 on text classification: A case study
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2eb0f0be-de19-49d0-b26a-e5c26d488c06 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Language Models are Few-Shot Learners
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca244d32-4943-4a39-b712-4a115bbb80ac · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Fine-Tuned 'Small' LLMs (Still) Significantly Outperform Zero-Shot Generative AI Models in Text Classification
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 523fbd7a-de83-4964-8fd9-4630c910dcde · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Unified Hallucination Detection for Multimodal Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb6aebae-89ec-4ba8-ac22-de904549b4c3 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Imagenet: A large-scale hierarchical im- age database
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb04a755-f9d4-495d-8378-f998781e6d2e · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Llms to the moon? reddit market sentiment analysis with large language mod- els
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3e23a81-ccd2-4f07-9be7-d0e605ac8e6f · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Virtex: Learning vi- sual representations from textual annotations
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f2c5cdc-a32c-49f0-9f38-1a2faea766c4 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb58b3b7-75ea-4df7-8a7e-2fbf7cb6eb02 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Towards Multimodal In-Context Learning for Vision & Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df53f3e3-8d29-42fe-9603-2fec24dae47a · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Training Vision Transformers for Image Retrieval
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5194c9a0-ac5f-40ad-8991-e8ae836711df · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features KTO: Model Alignment as Prospect Theoretic Optimization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d461ff02-49ca-4f0f-b995-dd15e6850e4e · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Behr, and Nancy Kan- wisher
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e878cf0e-87c1-4f84-935a-a03c82514ac9 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features BLINK: Multimodal Large Language Models Can See but Not Perceive
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a04b953-ddb6-4016-954f-60eab479891d · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features SimCSE: Simple Contrastive Learning of Sentence Embeddings
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba3caaae-440e-4a34-952d-de863883b8e9 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d28298b-388f-478a-9ab5-b897417304ed · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features MERLIN: Multimodal Embedding Refinement via LLM-based Iterative Navigation for Text-Video Retrieval-Rerank Pipeline
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9451e03c-4b15-4071-9043-e8152a9dd315 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Eurosat: A novel dataset and deep learning benchmark for land use and land cover classification
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d606d26-7a71-4abf-93d3-d203d269c733 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features In-Context Learning Creates Task Vectors
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e74514f7-cbf2-4859-8909-dd1f81a646db · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Incorporating structured representations into pretrained vi- sion \& language models using scene graphs
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aeda20c5-fa01-45f1-808a-bce0fd2ad5f2 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Hinton G, van der Maaten
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3daabcc-1716-4cd4-a045-f9f13176563e · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Finding visual task vectors
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74179b86-648b-4da6-a344-f76e81323af5 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features SugarCrepe: Fixing Hackable Benchmarks for Vision-Language Compositionality
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b4a7ca8-970e-422b-9686-e7b598e922e9 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features LoRA: Low-Rank Adaptation of Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5be0f649-ec2a-418a-abd1-682a70217bc8 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7998ba8-cbd3-4ebe-a15e-14b1a6eadab3 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Llm2clip: Powerful language model unlock richer visual representation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d62f176-d6c9-4df6-91af-6e84b65d0a47 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Hudson and Christopher D
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bce9494-9555-495c-83a8-a5e40749a083 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Learning Object Detection from Captions via Textual Scene Attributes
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 39148b52-46c9-476e-8be6-5ee3e96d056b · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Promptbert: Improving bert sen- tence embeddings with prompts
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6802af6-95e4-4021-927d-ee99eaa91656 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Scaling Sentence Embeddings with Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 237a0d01-dcd2-4b71-b099-f69379f6f25c · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features E5-V: Universal Embeddings with Multimodal Large Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7af35b5-88f0-4098-8bdf-6118b18bba8f · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Domain specificity in face perception
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 382a63d8-e84e-4863-bc52-dfba30d2f81f · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83ef24da-f997-4211-8e96-2834f6894f9c · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Auto-Encoding Variational Bayes
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 910264f5-b8ec-440c-a391-c7395233d91d · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Autoregressive image generation us- ing residual quantization
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4161a6d-6e21-4c15-95eb-f4aebf9e4a94 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Meta-task prompting elicits embeddings from large language models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9690ea70-a02c-469a-9577-a87499e1b821 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features The power of scale for parameter-efficient prompt tuning
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c700684-3f0a-4297-863e-d216f8f1f69c · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17c967a0-1d3c-453f-b983-174d6b3af53b · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features NaturalBench: Evaluating Vision-Language Models on Natural Adversarial Samples
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac48e6c0-5c35-4fbb-8baf-fb982d627ce6 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features LLaVA-OneVision: Easy Visual Task Transfer
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eed1a4c1-c4c0-487e-837b-c911053a739d · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0f68e6c-677c-4fd6-98cc-0496d8fdf931 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features BLIP-2: bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 417cda0d-77b3-4fe0-b33d-6b42dddfdb36 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Conan-embedding: General Text Embedding with More and Better Negative Samples
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14dceaea-d9d6-4ae5-9a85-08b932e4f859 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Prefix-Tuning: Optimizing Continuous Prompts for Generation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42edebf6-d422-4266-9906-fe60478ae4e0 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Scaling language-image pre- training via masking
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 970b216e-bf6d-4868-a1c3-aa7efe0d6b62 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Your Mixture-of-Experts LLM Is Secretly an Embedding Model For Free
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6aef712-9564-4205-a949-f92b788859cd · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features GT2Vec: Large Language Models as Multi-Modal Encoders for Text and Graph-Structured Data
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fecdeb87-71e5-4db3-b77f-bc6f4f83444f · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Maire, Serge J
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32744962-ab80-4aa2-9641-9069a427bd09 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Revisiting the Role of Language Priors in Vision-Language Models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7eae4df1-b646-42d4-9a84-4703e2aef339 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Multimodality helps unimodality: Cross- modal few-shot learning with multimodal models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f529054a-834f-4495-80b8-93c892edf478 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Evaluating text-to-visual generation with image- to-text generation
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3642c95-4b37-43e5-a23d-d6fa293ad50b · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Evaluating text-to-visual generation with image- to-text generation
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ac9665c-8893-4879-a96f-923ed2e325ab · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Improved baselines with visual instruction tuning, 2023
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61d8578c-b584-4850-a45e-76d8c05952fc · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Visual instruction tuning
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81031fed-02db-49cf-93c1-094d24034ad3 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Meaning Representations from Trajectories in Autoregressive Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3808aeeb-1ae9-4b0e-bee6-da5ed897c2b9 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Decision-making with auto-encoding variational bayes
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 860f3a34-4cc4-46b0-a8e6-eba528d8b32f · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Decision-making with auto-encoding variational bayes
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10691041-26b8-4e48-b5e5-e17e4edce9ba · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Ovis: Structural Embedding Alignment for Multimodal Large Language Model
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10a89739-248d-433c-a1cc-7c3a993b0531 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Unresolved cited work
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a571fcb7-0b4d-464c-90b0-d04a4c02e882 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Distributed representations of words and phrases and their compositionality
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0e5e3f9-82ab-4d97-b0c0-b574445b0d8f · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Compositional chain of thought prompting for large multimodal models
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation defbcfcb-9ee4-4916-a941-7e112781b921 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Mteb: Massive text embedding benchmark
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc664de4-3dde-4be6-9293-770fa48bb75d · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Sentence-T5: Scalable Sentence Encoders from Pre-trained Text-to-Text Models
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3091e1f7-985f-49cf-8c5c-7448981a64db · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Large Dual Encoders Are Generalizable Retrievers
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0670cd0b-a8c1-446e-9b06-3a9ee455f178 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Automated flower classification over a large number of classes
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e46d77f-b919-4518-9962-f5637ed0830b · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features In-context Learning and Induction Heads
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c1338aa-59ca-4408-93e8-256fda08462f · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features GPT-4 Technical Report
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 010f9609-f6f2-46f5-abea-2eb796dc0a0d · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Training language models to follow instructions with human feedback
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a400b89-6cda-4ef6-ba50-1a486ee49972 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Cats and dogs
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f67ca58f-9e22-46dc-9068-2da43c2402c9 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Pytorch: An imperative style, high-performance deep learning library
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01b999a2-0cd5-4f0e-a976-53c2eaa49278 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Glove: Global vectors for word representation
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9e6dcf0b-c3ac-441d-ae24-280d741e57fb · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Learn- ing transferable visual models from natural language super- vision
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8c24296-ec00-47bb-a410-22800db5ee67 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Direct 11 preference optimization: Your language model is secretly a reward model
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8cbff334-af6f-467b-8790-ad45513f619c · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d3e79ea-988c-4ebf-82db-4d101074fea0 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Sentence-bert: Sen- tence embeddings using siamese bert-networks
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d5dc1b9c-c76a-4329-a7d7-cba5c40cfc19 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Parallel distributed processing, volume 1: Ex- plorations in the microstructure of cognition: Foundations
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a95d2ca8-ef8d-456d-b84a-7e01e8fe4a4e · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Facenet: A unified embedding for face recognition and clustering
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 48d7cab2-082e-4b9d-95a3-2acea89ca86c · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Understanding machine learning: From theory to algorithms
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a23325ca-107e-444b-bcef-bb2ae5f3046d · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Towards vqa models that can read
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b7abfb13-0a01-44e0-a800-cbe89f29c988 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Singh, Stephanie C.Y
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0e01f113-c27c-4648-8f9c-c3b1c9469ffb · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Manning, Andrew Ng, and Christopher Potts
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bac12034-a2ca-4b81-bad1-062acff02f6b · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Text classification via large language models
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d04e8732-fb87-4c32-a3a7-1fe897211233 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72b9cea0-707c-416a-9982-4fed9c1b8160 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Gemini: A Family of Highly Capable Multimodal Models
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a2f1d75-2de5-4269-b006-4b69ab2ba907 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Winoground: Probing vision and language models for visio-linguistic compositionality
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d86023fd-50ae-405b-9975-d5d6ab24ea65 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Function Vectors in Large Language Models
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6f4c5bb-3d3e-4510-b508-11423e45f1ac · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Turk and Alex Pentland
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 442f34b2-5397-4270-beb1-0e3f5c570d8c · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Neural Discrete Representation Learning
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ddc2d99-ec51-42ac-87b4-7135453d20db · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Belongie
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 61fa933b-adb0-4852-9c4e-78fca8c13fe9 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Improving Text Embeddings with Large Language Models
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc3887f1-26e9-417d-b89f-df56945ae432 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72088dc8-cff8-4351-bf27-2a52e4338ed0 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features InferAligner: Inference-Time Alignment for Harmlessness through Cross-Model Guidance
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3493a6be-8d70-4e44-b9ac-0abe3377e6cf · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features InternVideo2: Scaling Foundation Models for Multimodal Video Understanding
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66ed4700-903a-41f0-b7e8-53b3b5f642c3 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Large Language Models Are Zero-Shot Text Classifiers
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f5c77e1-fc00-44cd-9ae4-3bf37f86797e · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Finetuned Language Models Are Zero-Shot Learners
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0cd3042-7589-4866-b072-5179448ce79b · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a9b37a6-4528-40d3-ab9f-fb380a97fed8 · outbound
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Scaling Laws for Discriminative Classification in Large Language Models
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e87e8d74-2e2e-4461-b28c-1e519c258d34 · inbound
Activation Reward Models for Few-Shot Model Alignment Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f39de5d2-e178-4293-996e-9311c4bec4e9 · inbound
Filter-And-Refine: A MLLM Based Cascade System for Industrial-Scale Video Content Moderation Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10332bb5-406a-41c9-a679-076ba29eadbb · inbound
Fine-R1: Make Multi-modal LLMs Excel in Fine-Grained Visual Recognition by Chain-of-Thought Reasoning Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.