Pith. sign in

Paper Citation Record · LEDGER

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark

As of 18 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2507.15882.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.15882 v2

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:12:03.811310Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact1
  • verified fuzzy18
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bdec5a75-e740-4419-9381-c9d9a27ecc92 · outbound

This paper cites Improving language understanding by generative pre-training.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Improving language understanding by generative pre-training

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:09.403705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:11:59.365713Z digest=sha256:c043509057447e4c8f000c30fa29b53ff5058edff19d812a448fd132a5734693

Observation 15ff5d91-7beb-4982-a532-37f0e1536a26 · outbound

This paper cites Palm: Scaling language modeling with pathways.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Palm: Scaling language modeling with pathways

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:59.401910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:59.401910Z digest=sha256:52972af3be1e2e7935ff55fea1467a58f998f0757c041f0250dac9d590cd9d77

Observation d7bf8b15-b176-45bb-9de9-0b3e549a0122 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark LLaMA: Open and Efficient Foundation Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:59.430437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:59.430437Z digest=sha256:92a5b30294849900a0b8931e8a7554de50921775c7424e3e1272a96436d581cc

Observation 5ffc3dc8-c6b8-4fe4-822d-8208fa21d627 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:59.455039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:59.455039Z digest=sha256:2ee7fdb1a8eaf71aec30145403bb628cf3c43f517dcbd3a7ae5989a86d5bb590

Observation 055f1b8a-079e-4295-a6e0-373421acef82 · outbound

This paper cites GPT-4 Technical Report.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark GPT-4 Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:59.493405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:59.493405Z digest=sha256:e4487a36b0d0210205b4674151296ded11d74b509900726587a58e861a1232ee

Observation afb4f16d-a539-4702-bdc4-b88b33ff2f43 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Gemini: A Family of Highly Capable Multimodal Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:59.549630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:59.549630Z digest=sha256:158d5c0a8ca3f7335e1bcbf9bcefe7fc618bbf644c994982df04bbe8c06fd50a

Observation 746a3555-eb67-43d1-bb57-19cc4bcab9f6 · outbound

This paper cites The Llama 3 Herd of Models.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark The Llama 3 Herd of Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:59.677645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:59.677645Z digest=sha256:4227b2294ef08e4199ad281522de8c02631c0634ac52148ad3dcec381ab20161

Observation dd1fa778-bdde-4df4-b501-cc0a34073135 · outbound

This paper cites Introducing the next generation of claude: Claude 3 model family.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Introducing the next generation of claude: Claude 3 model family

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:09.150309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:11:59.812322Z digest=sha256:1f9e8bd214d5a5c87dc57851940386ef1292b82763cb56bd2e3f483cc8228bd1

Observation 6029eb52-ba7e-4834-b8a9-72a7cc99e65f · outbound

This paper cites The amazon nova family of models: Technical report and model card.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark The amazon nova family of models: Technical report and model card

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:08.923684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:11:59.945149Z digest=sha256:f126c7b55065dbcbb9b37b67d1fc5e04ab6942a612c8e3efab158a99ccb082b6

Observation f74e43bc-4295-4db0-a940-1a8ed08f3396 · outbound

This paper cites Visionllm: Large language model is also an open-ended decoder for vision-centric tasks.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Visionllm: Large language model is also an open-ended decoder for vision-centric tasks

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:08.718134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:00.043620Z digest=sha256:67da721c1e414a678ef96265e56e6923d4a3271602b2093344ce365d4a209955

Observation 48e270f1-ef79-48ee-9c86-fb54bdc34c40 · outbound

This paper cites Adaptive Video Understanding Agent: Enhancing efficiency with dynamic frame sampling and feedback-driven reasoning.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Adaptive Video Understanding Agent: Enhancing efficiency with dynamic frame sampling and feedback-driven reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:00.161768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:00.161768Z digest=sha256:3ca0251a897a59745d8b7cf75575eaea204bc5ed995d423d5ec2a0d3cccf3066

Observation f70a710e-72fa-49f5-9473-03ad86e397d1 · outbound

This paper cites Zero-resource speech translation and recognition with llms.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Zero-resource speech translation and recognition with llms

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:08.478594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:00.359797Z digest=sha256:a939ab46865d153962a13c4176c926dce19c085d3f9878ee01f1f2cadd76594a

Observation 53e7e97b-432f-487e-9f15-55cc23f44d33 · outbound

This paper cites Legal- bench: A collaboratively built benchmark for measuring legal reasoning in large language models.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Legal- bench: A collaboratively built benchmark for measuring legal reasoning in large language models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:08.206903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:00.538822Z digest=sha256:2098f8c4230e4c72f33bb5681c85ee4fc458b6b7c5df258675131572f0fe7c6d

Observation 691f8977-ccee-4b33-91cf-ed45f729094a · outbound

This paper cites Chatlaw: Open-source legal large language model with integrated external knowledge bases.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Chatlaw: Open-source legal large language model with integrated external knowledge bases

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:07.973577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:00.658961Z digest=sha256:4d48ce31fa068c281341af764e8e88226792db83d2fcba315d8d4dad8ea7bf8a

Observation cfd8f6f6-3ff4-4515-a935-f83a85ea8b94 · outbound

This paper cites A study of generative large language model for medical research and healthcare.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark A study of generative large language model for medical research and healthcare

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:07.652417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:00.724539Z digest=sha256:fe0731d3bfdc51e0d2eadad417809285a2c87b6fd7dd712e5af7ae7740d0320c

Observation 5709f843-c7c6-4453-b980-c1b77cd516df · outbound

This paper cites Large language models in medicine.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Large language models in medicine

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:07.486410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:00.817655Z digest=sha256:238f128d991b0b4fe5b87bf2f0888f0b34636d69dd38533660760b7d9e54fbd4

Observation 2d67e545-e78e-4c93-93af-b62101901c99 · outbound

This paper cites Large language models in finance: A survey.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Large language models in finance: A survey

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:07.265020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:00.931170Z digest=sha256:f6f0ee31ac86f322b818c87ca87738c95706fe89a4061326c2d8c9ca2549612c

Observation 01eccb1b-fd52-43fb-a708-227884a5c295 · outbound

This paper cites A Survey on Large Language Models for Critical Societal Domains: Finance, Healthcare, and Law.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark A Survey on Large Language Models for Critical Societal Domains: Finance, Healthcare, and Law

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:01.088100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:01.088100Z digest=sha256:9952b354fb70325443c3c713ab034e386a768f53f7e27c9174dddf475c5afa26

Observation be36be46-14c7-442b-9148-bdf8a426487c · outbound

This paper cites GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:01.261807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:01.261807Z digest=sha256:ef34f514b2b980657d217dcd2ced2e0c4b10f2d54da146f3b9a51dd52851266c

Observation dadc202f-5e84-45c4-a3ff-3e8eceeda7ff · outbound

This paper cites Superglue: A stickier benchmark for general- purpose language understanding systems.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Superglue: A stickier benchmark for general- purpose language understanding systems

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:06.917425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:01.394095Z digest=sha256:3abe500aca58dfafe70b43508c6057ec4d55c9f1e90aecebfd9f9b5e2322fd0b

Observation 05a1a7e5-78ad-45cf-b5b1-2b7ce80cbff0 · outbound

This paper cites Needle in a haystack-pressure testing llms.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Needle in a haystack-pressure testing llms

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:06.668443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:01.566528Z digest=sha256:739b517ecc200dbe7f560f50c8e1fffb0207104ad7c384ab666324253679e697

Observation b703fee8-e780-40c9-a86a-53f5c6c9b6a9 · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:01.712774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:01.712774Z digest=sha256:7b7469c8009647e54af573ea6162b27d0b0b390e0b85ca85277c5750dd1014a9

Observation e8e3a648-420a-4b26-b5f5-3e32071c7ed3 · outbound

This paper cites Vqa: Visual question answering.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Vqa: Visual question answering

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:06.382209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:01.876492Z digest=sha256:9e3b67a00f9a7e55ff319a16c0ec75a6b9c65a1a748be169cba69e45660431a0

Observation f4e19d24-0934-444b-8fab-2888b4389a5b · outbound

This paper cites A Corpus for Reasoning About Natural Language Grounded in Photographs.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark A Corpus for Reasoning About Natural Language Grounded in Photographs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:02.036747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:02.036747Z digest=sha256:1113be82994fe348f6f18ec8bcdb2ed71c47ac86f309f189f6442586e0a918fc

Observation 2c6eae63-2f9c-449f-b856-6080ae010be5 · outbound

This paper cites MileBench: Benchmarking MLLMs in Long Context.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark MileBench: Benchmarking MLLMs in Long Context

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:02.171862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:02.171862Z digest=sha256:a2887156e85e60730094cb7fb771362ea3d06c8ee1001e6c4195826afe16a206

Observation 1e246951-2bed-4f69-9193-0ad1ef4dd56f · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:02.326329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:02.326329Z digest=sha256:e5df55f58e2518e7d8dab4d5b8cc3167ae4ad40491c70b2212d3138526a0f5ae

Observation 76c5ce3d-5e6d-464e-a5d8-2572687aa250 · outbound

This paper cites Docvqa: A dataset for vqa on document images.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Docvqa: A dataset for vqa on document images

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:06.071480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:02.479204Z digest=sha256:dace31a3b92a1be264a1887f0f7e8591e62c130cf5e1962756c891cbb62d298a

Observation 91332362-0274-45a4-a58f-852ee01988dd · outbound

This paper cites Document understanding dataset and evalua- tion (dude).

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Document understanding dataset and evalua- tion (dude)

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:05.884146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:02.641880Z digest=sha256:11c47e2f800f7254cecd70d832347277ac5fe94204c1061ef03623c419bb7310

Observation eaa94a9d-bd12-43eb-bc08-6fa51c8b11c3 · outbound

This paper cites Leave no document behind: Benchmarking long-context llms with extended multi-doc qa.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Leave no document behind: Benchmarking long-context llms with extended multi-doc qa

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:05.594366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:02.766229Z digest=sha256:71a345333c82be8d9999c0ed0fba443398aa95bb1c76392569d90308ce53bce6

Observation e8b89465-2f2f-4727-9996-0e454d557092 · outbound

This paper cites Slidevqa: A dataset for document visual question answering on multiple images.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Slidevqa: A dataset for document visual question answering on multiple images

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:05.270470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:02.916848Z digest=sha256:2a182bd4f9a10b45cc5dc51fc0007248b7e6e9c818096d2c3dcebfbd701258e6

Observation d4acb8a5-5d1b-43b1-94a9-f204c153c49a · outbound

This paper cites MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:02.996244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:02.996244Z digest=sha256:74c573f69eff29eb11095a90a0b24d6ae7b5d68dd3c8bc8765a18e66acc45ecb

Observation 527e72e5-0a7d-4e03-9788-624ac5fab665 · outbound

This paper cites Needle In A Multimodal Haystack.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Needle In A Multimodal Haystack

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:03.116125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:03.116125Z digest=sha256:a0b8a9b34eb7f8cfec718af3cc448b0cc1118c40a0b52f9691f77c0d0dc84ccd

Observation 40d0d1e0-ba4b-4b0a-abe3-874870c8707a · outbound

This paper cites M-Longdoc: A Benchmark For Multimodal Super-Long Document Understanding And A Retrieval-Aware Tuning Framework.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark M-Longdoc: A Benchmark For Multimodal Super-Long Document Understanding And A Retrieval-Aware Tuning Framework

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:03.272897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:03.272897Z digest=sha256:0e88eb7ae4a94e8dc7bfdd28ae8400998b4abd49cc472b9fbb82f0fad763b92f

Observation c64c4e5b-afd7-4f9b-a92e-1e04eed10b3e · outbound

This paper cites Pixtral 12B.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Pixtral 12B

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:03.418422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:03.418422Z digest=sha256:a1be050e6ea48dceea1da9658e3652b06e07f71c48ae948a02909855b5cdbb1c

Observation 07604ab5-94a6-416c-8a5b-08e471b4db48 · outbound

This paper cites Cross-Modal Retrieval: A Systematic Review of Methods and Future Directions.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Cross-Modal Retrieval: A Systematic Review of Methods and Future Directions

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:12:04.176645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:03.638856Z digest=sha256:a6723f3804b358c9d95f6d28a921fe5f946bf91dd0802e6b4f443b34c8293d6c

Observation 8d68d363-5b69-4a39-86db-2c1b14cf346d · outbound

This paper cites Lost in the middle: How language models use long contexts.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Lost in the middle: How language models use long contexts

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:04.935596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T16:12:03.811310Z digest=sha256:73c4b553eaf55fc4cf187f14c36058a1f47b14f0fc74685d34afbb8e797010f5

Pith citing papers

No inbound Pith citation observations are available.