Pith. sign in

Paper Citation Record · LEDGER

DocVQA: A Dataset for VQA on Document Images

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 38 inbound Pith citation observations for arXiv:2007.00398.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2007.00398 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 38 of 38 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:53:24.790502Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T21:00:09.476068Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9fc75294-2741-4978-a25d-8c0ab778eea9 · inbound

Long Context Transfer from Language to Vision cites this paper.

Long Context Transfer from Language to Vision DocVQA: A Dataset for VQA on Document Images

Reference 54

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:08:36.126572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T07:08:35.946669Z digest=sha256:aa0975ec9cbb0e7e088eded5e13bf00d35ec85c33dc765770720a9284dc54368

Observation 50d9352d-b38e-467e-b124-af0163694f8b · inbound

PaliGemma: A versatile 3B VLM for transfer cites this paper.

PaliGemma: A versatile 3B VLM for transfer DocVQA: A Dataset for VQA on Document Images

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:10:20.770367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T13:10:19.972353Z digest=sha256:672d8481bef54f2a5c8a4f090411887561f91c58ee03b57302e03c55d4230f88

Observation db46a353-cc1a-4cdc-8314-d9c82893565f · inbound

PaliGemma 2: A Family of Versatile VLMs for Transfer cites this paper.

PaliGemma 2: A Family of Versatile VLMs for Transfer DocVQA: A Dataset for VQA on Document Images

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:15:07.712195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T09:15:07.523565Z digest=sha256:f62a343c4bf3dcc67634befd9897a7232ddb848b972c6553a5818948c541c4cb

Observation 369fbbf5-48c1-4fab-94fc-9a7d3f52ddf4 · inbound

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding cites this paper.

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding DocVQA: A Dataset for VQA on Document Images

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:52:01.796335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T19:49:00.961388Z digest=sha256:739ccbfc83cd88de64eeb6f674d7cf1006e3ae41d6d4106fa38b6907d1016b12

Observation 99fdb9fa-25a6-4dd7-960c-7f37b6558ffc · inbound

Understand, Think, and Answer: Advancing Visual Reasoning with Large Multimodal Models cites this paper.

Understand, Think, and Answer: Advancing Visual Reasoning with Large Multimodal Models DocVQA: A Dataset for VQA on Document Images

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:24.790502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:24.790502Z digest=sha256:37a2edb0451ff40d20369a9924b50573b1436d7645268d12f79d006f0799d06b

Observation eab244cd-062d-451b-91ab-aaaca92a01da · inbound

Spoken question answering for visual queries cites this paper.

Spoken question answering for visual queries DocVQA: A Dataset for VQA on Document Images

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.942212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.942212Z digest=sha256:4bac1316ef89ea04c9f9ee46028104949f894022257889d72d61437bda3a5aeb

Observation 7e1f81cd-a1b2-4d68-b093-2d83107fdba0 · inbound

EffiVLM-BENCH: A Comprehensive Benchmark for Evaluating Training-Free Acceleration in Large Vision-Language Models cites this paper.

EffiVLM-BENCH: A Comprehensive Benchmark for Evaluating Training-Free Acceleration in Large Vision-Language Models DocVQA: A Dataset for VQA on Document Images

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:06.849342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:06.849342Z digest=sha256:a7db3dc7dac3601368ad4b541a48215235dfda2d61f5556c69e338f8be866f39

Observation 202d097c-9a67-4c76-85eb-19e2422e6c74 · inbound

On the Comprehensibility of Multi-structured Financial Documents using LLMs and Pre-processing Tools cites this paper.

On the Comprehensibility of Multi-structured Financial Documents using LLMs and Pre-processing Tools DocVQA: A Dataset for VQA on Document Images

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:36.789692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:27:36.789692Z digest=sha256:f4c7fe43c1482a4d565032c8e3ae279b667eb644e93d0eecfe7b929e1fd4b95e

Observation 40e38e9f-cae0-47d4-9c06-2aae302fb3f7 · inbound

VDInstruct: Zero-Shot Key Information Extraction via Content-Aware Vision Tokenization cites this paper.

VDInstruct: Zero-Shot Key Information Extraction via Content-Aware Vision Tokenization DocVQA: A Dataset for VQA on Document Images

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T17:59:05.622333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:59:05.622333Z digest=sha256:ad0d3c5a431412e2a8343c6ac9db6f4ba1365035bbb59428c68b7b4333dfdc56

Observation d6733d6f-d499-4144-9636-5486db821583 · inbound

Differential Multimodal Transformers cites this paper.

Differential Multimodal Transformers DocVQA: A Dataset for VQA on Document Images

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:20.935176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:39:20.935176Z digest=sha256:de7fe0eb3b56dd75654be7ae5b0e4280dc2acd77efeb358a9f10910c6ebf8677

Observation b8386a33-20c2-4723-803b-f2ef52d0b13a · inbound

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning cites this paper.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning DocVQA: A Dataset for VQA on Document Images

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T10:14:26.472521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:14:26.472521Z digest=sha256:28ff6ffd8c1a3da094af15408c3c8b8f189bbe373aa83d83adc068844072157e

Observation 2b5709b0-a04a-4bd6-87d9-71b46c381cc6 · inbound

VLMs-in-the-Wild: Bridging the Gap Between Academic Benchmarks and Enterprise Reality cites this paper.

VLMs-in-the-Wild: Bridging the Gap Between Academic Benchmarks and Enterprise Reality DocVQA: A Dataset for VQA on Document Images

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T11:11:20.832097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:11:20.832097Z digest=sha256:2549f475cc72556509554504b22cc0130cc07a0eea8c620b87ef17aab7d4e132

Observation 054fa4f0-9302-40b0-aa7c-d26a758aa5bf · inbound

Visual-TableQA: Open-Domain Benchmark for Reasoning over Table Images cites this paper.

Visual-TableQA: Open-Domain Benchmark for Reasoning over Table Images DocVQA: A Dataset for VQA on Document Images

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T17:42:47.586020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-18T17:37:31.837022Z digest=sha256:fe5bd06daea3ff77a7ecbf5f0a372e616f72a6601377f0dd4dc72d3be4126732

Observation ab61b841-c5db-45f1-afeb-b8d1dd3d97dd · inbound

Routing-Based Continual Learning for Multimodal Large Language Models cites this paper.

Routing-Based Continual Learning for Multimodal Large Language Models DocVQA: A Dataset for VQA on Document Images

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:55:35.336139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-18T00:52:36.700027Z digest=sha256:dcec008802f66f2c7320e6e2305029e4bc5d02cad9e9042299af4258fbff59dd

Observation 2fdfb66c-d33a-4093-b387-627443e15855 · inbound

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR cites this paper.

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR DocVQA: A Dataset for VQA on Document Images

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T20:20:11.510920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T20:19:52.701262Z digest=sha256:9c93af3075e18fa419d68b3cc99bf91b845757ec8d7d23b116fe36c0fb910ef9

Observation 2d897ada-8f28-4f91-9800-5f8d6012ed55 · inbound

Reconstructing Content with Collaborative Attention for Universal Multimodal Representation Learning cites this paper.

Reconstructing Content with Collaborative Attention for Universal Multimodal Representation Learning DocVQA: A Dataset for VQA on Document Images

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T19:41:33.738459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:41:33.738459Z digest=sha256:01c11df0f4a173b2f8674e78481f408950050772ec61f0a149da7ef7b6ea43b8

Observation e8010136-aa82-4ae3-8ded-81593f01f200 · inbound

FileGram: Grounding Agent Personalization in File-System Behavioral Traces cites this paper.

FileGram: Grounding Agent Personalization in File-System Behavioral Traces DocVQA: A Dataset for VQA on Document Images

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:45:51.276505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:33:57.643549Z digest=sha256:34130c51e9e24b93a7dbc202e85114fe2b1e60120cd5757daa437737f34051f0

Observation 756994a1-db77-498a-90b5-6873aee623f7 · inbound

Q-Zoom: Query-Aware Adaptive Perception for Efficient Multimodal Large Language Models cites this paper.

Q-Zoom: Query-Aware Adaptive Perception for Efficient Multimodal Large Language Models DocVQA: A Dataset for VQA on Document Images

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:55:52.856348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:46:26.869644Z digest=sha256:79619c8e83f0a15dc9b9173b2fa8c15e5895bcfea56be1aaf53ab3aeca6deea1

Observation 99a7eafb-0ef2-47b6-bbc3-23c10c5fed5c · inbound

Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment cites this paper.

Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment DocVQA: A Dataset for VQA on Document Images

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:41:22.377691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:30:39.410040Z digest=sha256:cef136b5c79b985aaafd88db5cf92595083502129f7d0bc41ee5cbd4f904f553

Observation f54d59ad-4dd6-4343-9bec-42260ac96d19 · inbound

Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models cites this paper.

Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models DocVQA: A Dataset for VQA on Document Images

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:16:11.146347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:14:11.941977Z digest=sha256:6e074e385e3042ef0dc18534b6ee53b02b23007f07fa23b2dc66cf33e8597445

Observation ed301277-68ed-493f-b006-79df3acf6942 · inbound

Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems cites this paper.

Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems DocVQA: A Dataset for VQA on Document Images

Reference 1

Resolution
malformed identifier
arxiv_id, observed 2026-05-10T11:50:20.766938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T11:42:13.100463Z digest=sha256:25b36dc4d563aa0797f15966c5b41848a8db681e3107a3a2110f2921e829032e

Observation 1bb123da-ab2f-465d-98e3-8289bbacef79 · inbound

ReaLB: Real-Time Load Balancing for Multimodal MoE Inference cites this paper.

ReaLB: Real-Time Load Balancing for Multimodal MoE Inference DocVQA: A Dataset for VQA on Document Images

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:36:10.208047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T01:18:27.597377Z digest=sha256:37b55b5b022dae21a1d09309b4d9e5c3d19f21423e4aa482f0ef172bcd99c956

Observation fd9d38bd-1825-406a-be64-8868f184dab1 · inbound

Visual Reasoning through Tool-supervised Reinforcement Learning cites this paper.

Visual Reasoning through Tool-supervised Reinforcement Learning DocVQA: A Dataset for VQA on Document Images

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:46:03.231775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T03:05:21.688216Z digest=sha256:e0e4564782169b1e62e8dd5465a060d51684c064f04d8230b5fe6818164980a8

Observation 38b54dfb-70d2-4d82-9413-4a23a8d5eea4 · inbound

The category of Whittaker modules over the Cartan Type Lie algebra $\bar{S}_2$ cites this paper.

The category of Whittaker modules over the Cartan Type Lie algebra $\bar{S}_2$ DocVQA: A Dataset for VQA on Document Images

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:25:40.475504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T09:08:06.592577Z digest=sha256:d3e67549c6efa062bab1d6a49fd0c2ec543b5ccde6d4d85b3b15c3d3dc901069

Observation 3b1e86e4-3282-4455-a195-5d9290b8e94a · inbound

FCMBench-Video: Benchmarking Document Video Intelligence cites this paper.

FCMBench-Video: Benchmarking Document Video Intelligence DocVQA: A Dataset for VQA on Document Images

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:21:41.346550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T17:14:19.186123Z digest=sha256:b97a252ccd8e6687b9fdb22dd7edd9188de6c220368dd3511c499d51327339f6

Observation fec6371b-9edf-4e6c-b000-4668c7c106e8 · inbound

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction cites this paper.

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction DocVQA: A Dataset for VQA on Document Images

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:26:02.150150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T15:29:17.567557Z digest=sha256:a51f2c786a855d36eb0d060c70cb499efdca13e72204a2bac4ce43d50fcda83c

Observation 430876f6-6afd-4c69-92ed-2e3c253887b0 · inbound

Focus-then-Context: Subject-Centric Progressive Visual Token Reduction for Vision-Language Models cites this paper.

Focus-then-Context: Subject-Centric Progressive Visual Token Reduction for Vision-Language Models DocVQA: A Dataset for VQA on Document Images

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:23:58.581700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-21T05:20:55.448430Z digest=sha256:937660fcdc4623b07c3af01d6dbf48a45464028c1b8485a1635c06e1238b23ba

Observation fdc13621-692c-4c1a-9d69-acb3c29209ca · inbound

Reinforcement Learning with Robust Rubric Rewards cites this paper.

Reinforcement Learning with Robust Rubric Rewards DocVQA: A Dataset for VQA on Document Images

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:43:13.742120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T07:39:21.677389Z digest=sha256:8b32d166ef2a865c99331f94ff89d6af73a98a9373af9ed239864684a5d2e041

Observation cbe0c011-0592-4df5-9f91-8eeaf596ab3f · inbound

Representation Forcing for Bottleneck-Free Unified Multimodal Models cites this paper.

Representation Forcing for Bottleneck-Free Unified Multimodal Models DocVQA: A Dataset for VQA on Document Images

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:00.977140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T22:54:10.460872Z digest=sha256:da79aa77198f4c1f045e3431d8ca7cef87eae1a6548d1c8a209b2b8c72d4adb1

Observation 4c9a5c00-a365-4c3e-89fc-85f4d9e504a0 · inbound

Representation Forcing for Bottleneck-Free Unified Multimodal Models cites this paper.

Representation Forcing for Bottleneck-Free Unified Multimodal Models DocVQA: A Dataset for VQA on Document Images

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T15:31:57.426559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T15:31:57.426559Z digest=sha256:d9c648f0b698f45e0a1cb14736a23ec0714d3cc128ab0b954496d0feb741b4df

Observation 010cabf7-d743-4021-b8fc-fcd648a84a74 · inbound

Loss Landscape Poisoning: Targeted Extraction of Unseen Training Data from LLMs cites this paper.

Loss Landscape Poisoning: Targeted Extraction of Unseen Training Data from LLMs DocVQA: A Dataset for VQA on Document Images

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:38:43.814213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T03:59:30.468854Z digest=sha256:65407cb087de566a2e02733ac8cb90f8a551c09279d1b7ad00221009113f124d

Observation 1bf935f7-a473-45a5-873b-0012da53ba03 · inbound

How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations cites this paper.

How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations DocVQA: A Dataset for VQA on Document Images

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T21:00:09.477417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-25T19:13:27.527971Z digest=sha256:cec4a2d39d1b2440d0eb31a1b78b2abfd8e94cf9b054bd6519422a46e8f5be3c

Observation d3a070c7-8575-4c1c-96f1-715e82e35265 · inbound

Combating Textual Noise and Redundancy: Entropy-Aware Dense Visual Token Pruning cites this paper.

Combating Textual Noise and Redundancy: Entropy-Aware Dense Visual Token Pruning DocVQA: A Dataset for VQA on Document Images

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T14:48:32.412815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-03T14:47:35.377391Z digest=sha256:dc3fdacd69ef68b9bb26ec8a86ab5890468b10cacd9cf70e07791d933aa96c13

Observation d0a1ea24-90f5-47ea-a7ef-7e51dccc1a17 · inbound

RADIO1D: Elastic Representations for Condensed Vision Modeling cites this paper.

RADIO1D: Elastic Representations for Condensed Vision Modeling DocVQA: A Dataset for VQA on Document Images

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T01:07:20.766474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:07:20.766474Z digest=sha256:3a106badbd490609ef891faf22af91562cedf3be5db59f4870287287e020c98a

Observation 7e001a6e-0255-4ee9-93d8-2d2459c10395 · inbound

Data Pyramid for Embodied Manipulation cites this paper.

Data Pyramid for Embodied Manipulation DocVQA: A Dataset for VQA on Document Images

Reference 253

Resolution
unresolved
no resolver link, observed 2026-07-31T06:18:55.901623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:18:55.901623Z digest=sha256:088e42c59fe24a7c507b44a63265ddee7cefd4907f67066d9f12124d7f1fe373

Observation 2774c8d8-9c4f-4ac0-b7c5-6e6d4d973c7e · inbound

DocAnnot -- Accelerating the Creation of Key Information Extraction Datasets with GenAI-Powered Auto-annotation cites this paper.

DocAnnot -- Accelerating the Creation of Key Information Extraction Datasets with GenAI-Powered Auto-annotation DocVQA: A Dataset for VQA on Document Images

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T14:38:35.492534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:38:35.492534Z digest=sha256:c230c82694d61cc6cf6c6498a7190b4a7b4a4cb6841456c7fdbc61e15825c5fa

Observation 86b507e3-44ac-49ba-a61e-a3f591c852f8 · inbound

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition cites this paper.

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition DocVQA: A Dataset for VQA on Document Images

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T02:55:23.403485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:55:23.403485Z digest=sha256:48c3415413b2a9c8977c6775e40458999be473d451cd0a9726ff3208897aea84

Observation d129ac94-b538-4ea4-b655-27c8cecd55c1 · inbound

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds cites this paper.

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds DocVQA: A Dataset for VQA on Document Images

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T04:26:29.325114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:26:29.325114Z digest=sha256:d9388947245f4e17f0a59a077f725c4e58e6fafd70814ef74bbcc82f733895ea