Pith. sign in

Paper Citation Record · LEDGER

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents

As of 7 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2607.24748.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.24748 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T13:26:31.987399Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 13ca8334-0769-46b6-bae4-6a11d6ec6ba6 · outbound

This paper cites - Start with meta search using document titles and metadata for fast candidate selection.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Start with meta search using document titles and metadata for fast candidate selection

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.449554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.449554Z digest=sha256:5b27ff5a607ab21958a7714276f60aab5fa2020eb16cc7dcca4d4fa75748e8bc

Observation cd5af9bd-2028-41e2-a89c-c26eeef662ca · outbound

This paper cites - Perform semantic search using HyDE embeddings for conceptual similarity.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Perform semantic search using HyDE embeddings for conceptual similarity

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.606217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.606217Z digest=sha256:d6d8ede1639a9cc2da88e2f17ab2f7c3fe0515d0c2977ceb857067adc2312bf4

Observation 29cb9329-9009-4bca-847a-5ee70afa8bad · outbound

This paper cites Hu, A., Xu, H., Zhang, L., Ye, J., Yan, M., Zhang, J., Jin, Q., Huang, F., and Zhou, J.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents Hu, A., Xu, H., Zhang, L., Ye, J., Yan, M., Zhang, J., Jin, Q., Huang, F., and Zhou, J

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:29.676884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:29.676884Z digest=sha256:b2a524a1bfb23f1b1f64c3abfe9bca0ec19910eedbfb81e4f0e169ae31f28f94

Observation 8ea05033-810f-4109-8777-0960d41aebdd · outbound

This paper cites MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:29.810681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:29.810681Z digest=sha256:a882f553371c193732071968dcf2444cd554a412a9e1391f29ab10cf5e691abd

Observation d4a5edc3-07b1-4e7d-8d4c-eefd0038f00a · outbound

This paper cites - Include document metadata and chunk positions for traceability.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Include document metadata and chunk positions for traceability

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.028796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.028796Z digest=sha256:24fa9402c15ab024ad4ee4faf9f3c210a2538880fd8db43a6b100eb74c1a236c

Observation fe02b393-dc29-4800-9b77-4e56c2c67881 · outbound

This paper cites Shi, Y ., Wang, J., Shan, Z., Peng, D., Lin, Z., and Jin, L.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents Shi, Y ., Wang, J., Shan, Z., Peng, D., Lin, Z., and Jin, L

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.159099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.159099Z digest=sha256:588e6d4c54fde62e0c88b861cd907624deb5cebc2cef91e6db3dfe301fc97493

Observation f9d19562-1b51-4950-8752-e07749f432ce · outbound

This paper cites - Generate evidence from the final chunks with proper source attribution.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Generate evidence from the final chunks with proper source attribution

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.789058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.789058Z digest=sha256:59d83c512adbf399ca23c06d4102525bd9488f262ec9b40a1821e924ea8fd980

Observation d7cd0f3d-5b61-4276-a5eb-2163ad1a64d3 · outbound

This paper cites - Do not repeat the same failed search parameters.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Do not repeat the same failed search parameters

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.937442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.937442Z digest=sha256:cf3326bd2bf7fb3da38a3322775c5af09ebc226c25af2e13149b26f3cffe4565

Observation e4c3dd45-edc8-4a4a-b65a-b2bdc5af6f9e · outbound

This paper cites - Identify restrictive qualifiers in the original query.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Identify restrictive qualifiers in the original query

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.192196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.192196Z digest=sha256:6bff72d6b84d5e436ce9398f312d2b3fca52e2bfe2d659ad92d2a4d3b407d6db

Observation f9d1d3b8-7012-4545-8b20-2b108d8301c5 · outbound

This paper cites - Remove numerical constraints if not essential.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Remove numerical constraints if not essential

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.350618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.350618Z digest=sha256:d72c5e35a6af1eec6482d82519f97568f42be85e826bebe0efcd2b75ab81e930

Observation ded97a5b-c8b6-48eb-a6b7-62b5b4c2b584 · outbound

This paper cites Generate a HyDE (Hypothetical Document Embedding) query for semantic search.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents Generate a HyDE (Hypothetical Document Embedding) query for semantic search

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.446244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.446244Z digest=sha256:ca464fa3f93a3571ff9575081eec8ee0cc1e4f892efee75ac5b5fce3ba4514de

Observation 9e490cef-901d-487c-99a9-44839ced6d46 · outbound

This paper cites - Generate 1-3 sentences, approximately 100-250 tokens.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Generate 1-3 sentences, approximately 100-250 tokens

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.602117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.602117Z digest=sha256:5b610290cd846cb8546b5494081a1d81d337e63a90781e4b8ebea1eee41c5969

Observation 74003f01-ffff-4efc-8131-fa9a564f10f8 · outbound

This paper cites - Describe the type of document that would contain the answer.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Describe the type of document that would contain the answer

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.743963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.743963Z digest=sha256:cfc18f617033e7eb35c43a45e4a4b1eb4be96fe506df821abfdbd4914658d054

Observation 8119981a-6a10-4934-819a-217060ac9883 · outbound

This paper cites - Include related concepts and context that might appear in relevant documents.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Include related concepts and context that might appear in relevant documents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.868542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.868542Z digest=sha256:44651064bd53a552b88bc11eff73457a2b3e286c4d4e5d6a63cae96d0217c114

Observation 3741cb58-0fa7-4d35-ac0b-9d27198684c0 · outbound

This paper cites - Do NOT use question format.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Do NOT use question format

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.987399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.987399Z digest=sha256:d515fb4f3d56077f79522821e9fcec39d04affe364e9edaf8b65f6fa591ebc17

Observation 5ad0e590-b36d-4834-b898-3d7f258b82e3 · outbound

This paper cites emnlp-main.311/.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents emnlp-main.311/

Reference 311

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.000582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.000582Z digest=sha256:a5e76e8f62c481d0d3e652cf3c9dac375056f531c0c6efd547a4f7183e685811

Observation 90e48d1b-acc6-4ed1-8d1d-13fa881ad2bb · outbound

This paper cites Retrieval-Augmented Generation with Graphs (GraphRAG).

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents Retrieval-Augmented Generation with Graphs (GraphRAG)

Reference 735

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:29.578873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:29.578873Z digest=sha256:bacf19c0db3133bf0aafed0e3c1231e5e07b12905efd2439a5572f9754b2d9e0

Observation c5b73bbf-e22c-4165-9cb5-f1312655973b · outbound

This paper cites VisRAG2.0: Mitigating Visual Hallucinations via Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents VisRAG2.0: Mitigating Visual Hallucinations via Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation

Reference 1166

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.298761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.298761Z digest=sha256:7745448fa72e0f0eb19ab1ced6fa46ee953586a5a7c6fa4c27a03523e758ec14

Observation 10edcf7f-3bcc-42de-9138-811916a638e0 · outbound

This paper cites ISBN 979-8-89176-251-0.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents ISBN 979-8-89176-251-0

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:29.528679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:29.528679Z digest=sha256:da104ce96cda016998c00bf6be10050dbe84d51031ea15fd64c695f1019c986a

Pith citing papers

No inbound Pith citation observations are available.