Pith. sign in

Paper Citation Record · LEDGER

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark

As of 20 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 4 inbound Pith citation observations for arXiv:2505.11651.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.11651 v2

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:54:30.130726Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:30:17.130322Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T10:46:52.605526Z

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 23125cfa-1128-4a07-a583-f0145315418d · outbound

This paper cites an unresolved cited work.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:54:30.452178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T20:54:30.031475Z digest=sha256:abd73fea90ea1a6f20f853d2f5694a29fddcee6e292fb407f5dc935a474e70cc

Observation e9f5e17c-d438-4a92-bccb-5021f171b295 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.037211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.037211Z digest=sha256:f4378425d7aea1458762a662b717cfefaad10296d4ef4c782099ab3b49a4b56d

Observation ffc8cf44-6d1b-48ae-97f1-420eeb5fef32 · outbound

This paper cites M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.042709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.042709Z digest=sha256:5f13375d764da2f95e5e7f5d61cc8cbcc3d545044dca441a271a1d6c301b1946

Observation d4857ef1-cae7-45b3-ad2b-4bc5d70b03f7 · outbound

This paper cites ColPali: Efficient Document Retrieval with Vision Language Models.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark ColPali: Efficient Document Retrieval with Vision Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.047966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.047966Z digest=sha256:e563900a038a147397d0a0b2a09efe62945b75a8038a041b9344d3444648831b

Observation 518cf49c-8be4-4ee2-be8b-dde6b1a47dd4 · outbound

This paper cites an unresolved cited work.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.053239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.053239Z digest=sha256:3293351df48c7b61562f3a63e6e45f7a859654b2bfe32aad28b3b9b794b6b68d

Observation 1b86ca15-0a9e-4867-a7d1-98698125ff28 · outbound

This paper cites Dense Passage Retrieval for Open-Domain Question Answering.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Dense Passage Retrieval for Open-Domain Question Answering

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.058852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.058852Z digest=sha256:770e6ca0bed6b4d8634200610b2cd9e3d2ebc7599e5159a5e1eb98575716e2ba

Observation 7aabd4e9-7949-4b83-ad2d-5e5638da8c28 · outbound

This paper cites Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.064456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.064456Z digest=sha256:6f9082ebb8e19542a581db2a6a876cf640728f4623245a47873d49d8b65a783c

Observation 8897b0ea-6568-478a-96f6-d461715f58d9 · outbound

This paper cites an unresolved cited work.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:54:30.426350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T20:54:30.069344Z digest=sha256:11455fa912faed5e128144bf55db582a0ec091cc349124cd9850fa0dc61568e4

Observation bc868860-0d08-4c82-99f5-516d985fc37a · outbound

This paper cites an unresolved cited work.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.074088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.074088Z digest=sha256:7cc0140a1269e27d10970e6f6507bb592c4092d1fe9856dda04d2145d956cf5f

Observation 4dd1a46c-f2f5-4607-a073-271d9a708217 · outbound

This paper cites PaliGemma 2: A Family of Versatile VLMs for Transfer.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark PaliGemma 2: A Family of Versatile VLMs for Transfer

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.084848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.084848Z digest=sha256:dcea076f0895e432f72a9f6ee2e3bcb2924c77c674cfb9a6d7d8a616565dd900

Observation 625069dd-f7f0-493f-9d1b-238fe66944a6 · outbound

This paper cites BEIR: A Heterogenous Benchmark for Zero-shot Evaluation of Information Retrieval Models.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark BEIR: A Heterogenous Benchmark for Zero-shot Evaluation of Information Retrieval Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.090222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.090222Z digest=sha256:445f43e8429eedb7c090c4ec249dd632882025a241d45b34ca77df9b197db51c

Observation 60f07c6c-0929-4527-afe7-07eaa2a4faa3 · outbound

This paper cites Text Embeddings by Weakly-Supervised Contrastive Pre-training.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Text Embeddings by Weakly-Supervised Contrastive Pre-training

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.096090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.096090Z digest=sha256:a5b5fd715520764e9c9dce9d1525eebd01057bbb60bf8ca8179356e6830106af

Observation 5f351bf5-dd1b-47c0-bda5-9483ddef83ff · outbound

This paper cites Improving Text Embeddings with Large Language Models.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Improving Text Embeddings with Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.100935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.100935Z digest=sha256:7b7a04f6db46449fd819fcd27ed0d50a41fb4c1ab1493b3419e58b2f3108f6c1

Observation cc45d07b-a567-4f40-959a-a6911663d3b9 · outbound

This paper cites Multilingual E5 Text Embeddings: A Technical Report.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Multilingual E5 Text Embeddings: A Technical Report

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.105825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.105825Z digest=sha256:674276c1ced1c52c3ee0526d88ff602c7a72ba773724e035cf7b1c0b14244a8b

Observation bba64d62-3b7e-415e-a69a-8251be2a3428 · outbound

This paper cites Arctic-Embed 2.0: Multilingual Retrieval Without Compromise.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Arctic-Embed 2.0: Multilingual Retrieval Without Compromise

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.110726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.110726Z digest=sha256:30da4ce3e33df994cf3507c0d2c478be4987600848db697074125d4138dc60e5

Observation be9f79f2-da11-480a-b558-893f88b1b87a · outbound

This paper cites an unresolved cited work.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.115669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.115669Z digest=sha256:3abd4148fcfd16fa82dcc8939a7bd319c66ea6b49d7675a235b02b68885a1f0e

Observation 3bf8af02-506b-4e59-bb5e-69aa7b52bdec · outbound

This paper cites mGTE: Generalized Long-Context Text Representation and Reranking Models for Multilingual Text Retrieval.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark mGTE: Generalized Long-Context Text Representation and Reranking Models for Multilingual Text Retrieval

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.125063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.125063Z digest=sha256:1179e03921e5dc242bd5a05b9095f78460b4540ceb4d865a0aa72c65424ecad9

Observation 430257e9-9e47-4d67-b5d2-866a2b4e9c87 · outbound

This paper cites GME: Improving Universal Multimodal Retrieval by Multimodal LLMs.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark GME: Improving Universal Multimodal Retrieval by Multimodal LLMs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.130726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.130726Z digest=sha256:1a5d8686cf0709bc34b6df5c7c7bb6b1ca167a47c51cec94bb07278c7f96170b

Observation 90d08dc9-7feb-483b-a852-0587107bca47 · outbound

This paper cites Making a MIRACL: Multilingual Information Retrieval Across a Continuum of Languages.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Making a MIRACL: Multilingual Information Retrieval Across a Continuum of Languages

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.120309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.120309Z digest=sha256:df827de55b6b28a144dac1c38eb820c0a85434fda410d9866de06f713e222ab9

Observation ddbb6ba8-2596-4272-ad8a-bc4bc0c14887 · outbound

This paper cites Unifying Multimodal Retrieval via Document Screenshot Embedding.

MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark Unifying Multimodal Retrieval via Document Screenshot Embedding

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:30.078700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:30.078700Z digest=sha256:f3b47e3b992b0286321d4c4a340e790b54993f94ac7a2b1acbac95320aa59040

Pith citing papers

Observation 5bdb5fe4-7cb7-43ea-9522-c6de41db5230 · inbound

Llama Nemoretriever Colembed: Top-Performing Text-Image Retrieval Model cites this paper.

Llama Nemoretriever Colembed: Top-Performing Text-Image Retrieval Model MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T19:30:17.130322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:30:17.130322Z digest=sha256:67060972d4d93c966c557066cb4ed30f7e2efc4d009c51342ee265315b5caf25

Observation d495d4e8-291d-494c-a2f6-1653f3988990 · inbound

Argus-Retriever: Vision-LLM Late-Interaction Retrieval with Region-Aware Query-Conditioned MoE for Visual Document Retrieval cites this paper.

Argus-Retriever: Vision-LLM Late-Interaction Retrieval with Region-Aware Query-Conditioned MoE for Visual Document Retrieval MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T10:46:52.606860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T04:55:04.299049Z digest=sha256:daca302fd8d7594bc523beb4368e11c48d55bbe3ffda194a4bd506964a0f1f11

Observation 4f3d91eb-fa8e-4539-8888-eb9b505080bd · inbound

MM-Matryoshka: Towards Budget-Elastic Visual Document Retrieval via a 2D Multimodal Matryoshka Training Framework cites this paper.

MM-Matryoshka: Towards Budget-Elastic Visual Document Retrieval via a 2D Multimodal Matryoshka Training Framework MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:26:45.625202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T06:59:06.801340Z digest=sha256:fbbd04bb69286efb70de8f0fa599702c14cc7b43ec5d1dc5292551c8146a8f96

Observation eae26c1d-24dd-44b8-953a-f6c4a44bdadc · inbound

KoVRE: Training an Efficient Embedding Model for Korean Visual Document Retrieval cites this paper.

KoVRE: Training an Efficient Embedding Model for Korean Visual Document Retrieval MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:08.520214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:08.520214Z digest=sha256:ae5862e72c7efc5b781a3abca0cc5329f5dc4c166d7f1d3489785ba795741010