Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:56:25.274118Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 3 inbound Pith citation observations for arXiv:2412.11475.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:56:25.274118Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:13:55.100419Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-16T10:13:25.893446Z
16 of 16 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ec0e5239-7dd6-4732-b339-de7ac23051de · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f0d4402-6331-47f9-94f2-7efc9df631d6 · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference Squid: Long Context as a New Modality for Energy-Efficient On-Device Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6fbdb3aa-9b2e-474b-9c87-03e0955f30d8 · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference Efficient and Effective Text Encoding for Chinese LLaMA and Alpaca
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63199eca-cfec-4be8-bdce-b47ddd994430 · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 496b8a2e-bd61-47f1-b964-64b98ae0a7ad · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference Ferret-UI 2: Mastering Universal User Interface Understanding Across Platforms
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2352a00-6adc-4510-ae19-c616f3862b2f · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference MobileLLM: Optimizing Sub-billion Parameter Language Models for On-Device Use Cases
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a1886ed-1fba-44c1-b721-05aada00d860 · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference Introducing orion, our first true augmented reality glasses, 2024b
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ebb69450-b345-4078-a341-33f7d557fb7a · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference MLC-LLM, 2023-2024
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3d6cf64b-e886-48b7-b655-d3e64e53dee7 · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference Ollama, 2023-2024
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4b3f9571-f3a6-4321-87be-95806625a7a4 · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference Gemma: Open Models Based on Gemini Research and Technology
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0069e66d-5914-41fd-9b25-78e40bd0ce71 · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab87ea45-a88f-423b-b55c-b993cb534dd1 · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference Sigmoid Loss for Language Image Pre-Training
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8a55622-171e-420c-9ed2-34191e0a8516 · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference The Llama 3 Herd of Models
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c22894f-f28d-41a2-8d7c-f5986ea75aac · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference Unresolved cited work
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c259c23c-8566-4ad1-bea6-68f281cee6dd · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference PaliGemma: A versatile 3B VLM for transfer
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4fd3c01-5ede-4f6a-9ed0-b8e86d0f0e5b · outbound
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference Octopus v2: On-device language model for super agent
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55649044-730d-41c5-bbe2-e7f7f061fbba · inbound
Towards a Multi-Agent Vision-Language System for Zero-Shot Novel Hazardous Object Detection for Autonomous Driving Safety OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8febbffd-7988-46e7-b6ce-d4a41c792e9f · inbound
A Review of 3D Object Detection with Vision-Language Models OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 77ac84e2-5936-4016-a7af-df25475e5e91 · inbound
AutoNeural: Co-Designing Vision-Language Models for NPU Inference OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.