Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2412.08802.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T23:20:01.029650Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T19:30:07.771573Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation a37951d5-4c65-4996-8dbf-85523f262b56 · inbound
MRAMG-Bench: A Comprehensive Benchmark for Advancing Multimodal Retrieval-Augmented Multimodal Generation jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c233abe-685d-4479-aed1-9faae35b6db7 · inbound
RGB-Pointmap Pretraining for Unified 3D Scene Understanding jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f5186e79-b65c-4286-91ea-410d98f875a5 · inbound
HIVE: Query, Hypothesize, Verify An LLM Framework for Multimodal Reasoning-Intensive Retrieval jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e6c5de1f-0bdd-4e19-893a-9e758e94f51b · inbound
jina-embeddings-v5-omni: Geometry-preserving Embeddings via Locked Aligned Towers jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8b86ce07-2594-4b42-96d6-f4e48c6c4bd3 · inbound
jina-embeddings-v5-omni: Geometry-preserving Embeddings via Locked Aligned Towers jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fb41aea0-70b2-483a-a861-eaea954a8b99 · inbound
jina-embeddings-v5-omni: Geometry-preserving Embeddings via Locked Aligned Towers jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f1cf3dec-3d74-43a2-8536-ce16658b0faf · inbound
MONET: A Massive, Open, Non-redundant and Enriched Text-to-image dataset jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9f548490-1d6f-4406-8dfd-aa5baf953dfb · inbound
One Stone, Three Birds: Self-adaptive Optimal Transport for Multi-VLM Selection, Adaptation, and Ensembling jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 20b3544d-2409-4544-8762-e46faac2b04a · inbound
Invoice Haystack: Benchmarking Document Retrieval and Visual Question Answering Under Strong Visual Homogeneity jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 67de105b-4d74-4a96-86cb-b0b461385c3a · inbound
Invoice Haystack: Benchmarking Document Retrieval and Visual Question Answering Under Strong Visual Homogeneity jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 994aec71-e745-4473-8b9b-2de9c6dc08df · inbound
ReasonCLIP-58M: Visually Grounded Commonsense Reasoning Supervision for CLIP jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5530c45c-12d9-46d5-bb05-c6171f617c16 · inbound
KoVRE: Training an Efficient Embedding Model for Korean Visual Document Retrieval jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73ae2735-f4c0-4927-ba64-3ea619c019e7 · inbound
Illuminating Visual Identity in Universal Multimodal Embeddings jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.