Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:11:35.513418Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 1 inbound Pith citation observation for arXiv:2505.19812.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:11:35.513418Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-28T07:23:19.428348Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T06:36:43.892355Z
48 of 48 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6c929e86-9daa-4a0a-ae4e-f983f0075e83 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Many-Shot In-Context Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7593ca6e-b579-4dcc-81f8-f9092f82acbf · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Flamingo : A visual language model for few-shot learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ef4c7895-a837-44e8-866d-182a5d776de0 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74436313-c209-4d5f-80ea-8c01ac2c68fc · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation In-Context Learning with Long-Context Models: An In-Depth Exploration
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34c7919d-a2ca-4cf7-af04-7fcd4afbe359 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 67e192e5-0c0f-4ed9-8f34-2a263735c92e · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2167d069-857a-41eb-9072-65fe975ff2e5 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Efficient Large Multi-modal Models via Visual Context Compression
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bc0f75e-6125-493f-8c57-10d178b0b7b9 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 860b85bd-4838-41f7-99ef-2752919474db · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0b8b3a9a-4426-43dc-8af1-703943d46d37 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 643dfc05-6154-42ff-b97c-a625472b495f · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15528579-1929-409d-acd8-fd534cf5f0cf · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Imagenet: A large-scale hierarchical image database
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1975f8b-7e2c-40f8-b5cf-d91c1e6c1c2e · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HD
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67dc8539-a0cb-4d7a-90c9-6d742bf8a407 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1299413-da39-438a-bf0e-0c18d0dc8edd · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation E2vpt: An effective and efficient approach for visual prompt tuning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0ec7e0b0-a40f-4510-b9a3-2b647d61eced · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Facing the Elephant in the Room: Visual Prompt Tuning or Full Finetuning?
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c152c7e-5fd1-4c63-ba77-0bcba38ae54c · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation How Well Does GPT-4V(ision) Adapt to Distribution Shifts? A Preliminary Investigation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2d172a6-b150-4e43-a188-e8811223a222 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Training Compute-Optimal Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a97cbb4a-0e38-412c-8503-cfe33ccabf46 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation J., yelong shen, Wallis, P., Allen-Zhu, Z., Li, Y., Wang, S., Wang, L., and Chen, W
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 519e0c92-c527-4f64-b436-a095b9845eb7 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c479d0e-d8d5-402a-9232-e27a2d20978e · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Multimodal task vectors enable many-shot multimodal in-context learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2418698b-fb25-4137-a8ad-8379cf3a1a2a · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation K., Patra, B., et al
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation babc7053-0aa1-404c-9f40-dea4a049286d · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Visual prompt tuning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb0a9261-7215-4c54-943f-05f6d289123c · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation A., Wang, J
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 23bc6067-5560-47dc-8812-daeea5d49daa · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation In-Context Learning with Many Demonstration Examples
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a00b5135-1371-4ce3-bbb3-dd0acbe987e5 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation SnapKV: LLM Knows What You are Looking for Before Generation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9bb4932-1b06-4bd8-b864-9d21e6b712bf · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1d887552-a737-4a45-baac-c887b14432db · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation DeepSeek-VL: Towards Real-World Vision-Language Understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8e3e0eb-4a9f-4b9a-84f1-c5ae5072c6a1 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation K., and Buehler, M
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0318de08-c0f8-4412-9937-4342799fd6f7 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Ok-vqa: A visual question answering benchmark requiring external knowledge
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a77dcc2-de59-46ea-b7dd-68f6fba2315f · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Kosmos-2: Grounding Multimodal Large Language Models to the World
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97260de0-4913-4b9e-9525-16a8bd43c526 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation S., Sayeed, K
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e63ac90b-c753-42c1-9c9e-ef41c2101dea · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Locllm: Exploiting generalizable human keypoint localization via large language model
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e3704243-3f69-43bc-9b09-978e6aca2fee · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11826797-261a-40f2-b98f-66a8f21e0a04 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation M2pt: Multimodal prompt tuning for zero-shot instruction learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6e0f3a88-6f01-436d-83d3-2e7ed9ae359b · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation CogVLM: Visual Expert for Pretrained Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6820257a-da9a-4614-8419-f4f9f994887d · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Efficient streaming language models with attention sinks
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebd177b5-26e7-4f03-8e4f-73a59f334cf0 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Pink: Unveiling the power of referential comprehension for multi-modal llms
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 362d4ae8-9baa-4683-894e-d66714b5ba9a · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Pyramidinfer: Pyramid kv cache compression for high-throughput llm inference
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 79afe392-d574-4bc9-af5a-75d37f59269f · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Yi: Open Foundation Models by 01.AI
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f7a8c52-13f7-49be-926a-c10ac801f6d5 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e26cdee-dfc5-4d36-bf20-3a7f973e4bd1 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation On the Out-Of-Distribution Generalization of Multimodal Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f215e4e-184c-4263-8b20-3192cfbb75cc · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b06defb-91f6-4e98-b706-31b9ce580a2d · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation H2o: Heavy-hitter oracle for efficient generative inference of large language models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ff75235-870d-43e0-aa6b-fe336c89a79f · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Towards automatic learning of procedures from web instructional videos
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f215767b-ef12-4d92-b20f-2bd45afa47d4 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 023b0629-f762-439d-8158-f004e14b5aad · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation Mini GPT -4: Enhancing vision-language understanding with advanced large language models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 82575eb0-9963-44bb-be2d-ef1d287bce29 · outbound
Efficient Multi-modal Long Context Learning for Training-free Adaptation write newline
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd89840e-8c0f-42d1-ae9c-b9b74dc55f22 · inbound
Hyper-ICL: Attention Calibration with Hyperbolic Anchor Distillation for Multimodal In-Context Learning Efficient Multi-modal Long Context Learning for Training-free Adaptation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.