Pith. sign in

Paper Citation Record · LEDGER

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation

As of 13 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 2 inbound Pith citation observations for arXiv:2507.02859.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02859 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:24:57.154109Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T10:06:52.822171Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T11:46:55.443273Z

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved20
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dcf504e3-25e2-4fb6-b3f4-b1afc63c03a0 · outbound

This paper cites Lion: Empowering multimodal large language model with dual-level visual knowledge.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Lion: Empowering multimodal large language model with dual-level visual knowledge

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:56.982611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:56.982611Z digest=sha256:c0d5d4759f3b1da5c876926af6fd59a1009dda5356d4ab14b2ccdfba9d7bf8dc

Observation b9b49cf4-48b3-4927-bedb-62f79bc4c293 · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:56.987451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:56.987451Z digest=sha256:c4ebbef50cc3637085526b5cc76399c6fa66c9422b24a926fba9326a011a7bea

Observation fddf9d81-e69c-41bc-8704-d7293b380278 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Training Verifiers to Solve Math Word Problems

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:56.992240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:56.992240Z digest=sha256:ce61f58cd08e1832bcdc585d9b963e967b0f2ad1f87bafc6cc10bb0c87c46f24

Observation fde16b29-bd6f-49a1-9fc3-c35633c2d7c0 · outbound

This paper cites The Llama 3 Herd of Models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:56.997274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:56.997274Z digest=sha256:21064af23a269c197bbd05293737b79e1745aea96a8b2e790f71c6d5142b96d1

Observation 337842db-ae19-48bd-b2cc-4701c88c8015 · outbound

This paper cites ChartLlama: A Multimodal LLM for Chart Understanding and Generation.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation ChartLlama: A Multimodal LLM for Chart Understanding and Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.001879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.001879Z digest=sha256:d915c1244ba19aa9757efa5038c137c3d2450ee0bbf0f4437df63059b0dde8f6

Observation c84e23e9-1b37-4075-a4e2-ba347837b1f8 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Lora: Low-rank adaptation of large language models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.717781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.006527Z digest=sha256:8cd3659fdf2d1250597ba38b1757ae70a4d617d60f165e8f6a20707a6facca24

Observation 89640aa1-16c9-4a21-981c-4d7c069ee456 · outbound

This paper cites Icdar2019 compe- tition on scanned receipt ocr and information extraction.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Icdar2019 compe- tition on scanned receipt ocr and information extraction

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.010697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.010697Z digest=sha256:e7c43f3cf107f5849934e8080115b0773229a922d0ac93031ea8e155532d82f8

Observation 41ccbb43-9ec3-4e91-bcb1-ccd0d4e2fdf1 · outbound

This paper cites Dvqa: Understanding data visualizations via ques- tion answering.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Dvqa: Understanding data visualizations via ques- tion answering

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.014859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.014859Z digest=sha256:5f8339195155589c02912db0a9be8c106b460c8c2bc6b10e507f60d5c35596f2

Observation 75e04989-60eb-42bd-990c-34ef9714ebcb · outbound

This paper cites Large language models are zero-shot reasoners.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Large language models are zero-shot reasoners

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.686967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.019111Z digest=sha256:0572076bb6f18c9b93ddc92f1d99cd9486408c93004895c7a2380dce06667c66

Observation 9b952664-fbd5-4b9f-b041-c23bd4d6d716 · outbound

This paper cites Visual genome: Connecting language and vision using crowdsourced dense image annotations.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Visual genome: Connecting language and vision using crowdsourced dense image annotations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.024113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.024113Z digest=sha256:26d39d067cb87b06398c8e048e0eb572cdfefff8f564bb971e2d1be7baf5f4d9

Observation 70ba3b26-17d3-41d2-8670-94555f69bacc · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.028725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.028725Z digest=sha256:65651d66abae710f54217c82fd42cdf31b65e8dee3995ed885acc82dc6f88267

Observation e2f59f45-9306-4a10-ae4f-c6dca2908af5 · outbound

This paper cites Deductive verification of chain-of-thought reasoning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Deductive verification of chain-of-thought reasoning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.656097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.033101Z digest=sha256:993b212a244d26342241628008ac1a7cb5385a645cf55633fdcc4dc51cf92546

Observation eeefdf7a-eee3-4929-9f3c-0eb9cc76fd04 · outbound

This paper cites Improved baselines with visual instruction tuning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Improved baselines with visual instruction tuning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.638991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.037955Z digest=sha256:1017353187662153683c983affcd225ff14c26a26ba8cfb556b048bf276457bf

Observation f691e5eb-9093-4fec-abe8-181e7349ec9e · outbound

This paper cites Visual instruction tuning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Visual instruction tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.042273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.042273Z digest=sha256:56cf2e0bf4978a4ec844658ef606d9a25fa3c199afed5a1991cd38bd44468cae

Observation 8d73a0a5-26ab-4176-a6f0-1edefb03280d · outbound

This paper cites Nltk: The natural language toolkit.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Nltk: The natural language toolkit

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.617121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.046456Z digest=sha256:066e7eb26e5db8f7c4fb1cca13a99f0eed246c6295f9976f6a01a6d1737cbce0

Observation 730f94d4-a500-453d-b600-77ff873e777c · outbound

This paper cites Decoupled weight decay regularization, 2019.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Decoupled weight decay regularization, 2019

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.051035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.051035Z digest=sha256:946cd2b8e40d13110dfeae41f9afb54b892c880a0c255e4431a63679965b75e5

Observation f67e04b2-c603-40db-b79a-40384cc05c33 · outbound

This paper cites Dynamic prompt learning via policy gradient for semi-structured mathematical reasoning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Dynamic prompt learning via policy gradient for semi-structured mathematical reasoning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.594983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.055458Z digest=sha256:3a5e3eb93ac7392fb1496468e12aa9de1d90123f4a383a2f4d43749c6da3544f

Observation 295cb5a7-f4bc-4a59-97ab-171cb26625f2 · outbound

This paper cites Chartqa: A benchmark for question an- swering about charts with visual and logical reasoning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Chartqa: A benchmark for question an- swering about charts with visual and logical reasoning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.580807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.060651Z digest=sha256:c2ba35ae92508f27a3232b2902f45e8ab883873cc53551de87e7bb9fb56a4f06

Observation f3f7c609-a6df-4643-a018-ee3e45b957f7 · outbound

This paper cites ChartGemma: Visual Instruction-tuning for Chart Reasoning in the Wild.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation ChartGemma: Visual Instruction-tuning for Chart Reasoning in the Wild

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.064925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.064925Z digest=sha256:4d8229d2ff88df07b36aeadaf7beec69545738ce960aeffeba8aa31f51e2c31a

Observation 196330d8-89f1-494f-bb63-b6a766836d71 · outbound

This paper cites ChartAssisstant: A Universal Chart Multimodal Language Model via Chart-to-Table Pre-training and Multitask Instruction Tuning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation ChartAssisstant: A Universal Chart Multimodal Language Model via Chart-to-Table Pre-training and Multitask Instruction Tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.069320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.069320Z digest=sha256:bf3137704cb20412452a98afffd3d44f52c8ae34200c35f98a78e855bafe9957

Observation 5684f8a7-b8f4-4511-a2fa-e21875e76203 · outbound

This paper cites Selfcheck: Using llms to zero-shot check their own step-by-step reason- ing.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Selfcheck: Using llms to zero-shot check their own step-by-step reason- ing

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.567212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.073735Z digest=sha256:fb4fd013e92ec5b335a1b62d042053a5305fd0c4f9dc1694776ee353dece515b

Observation 7c4cc44b-90d6-4fef-946c-a6aa0c2d07da · outbound

This paper cites Im2text: Describing images using 1 million captioned pho- tographs.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Im2text: Describing images using 1 million captioned pho- tographs

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.552516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.078227Z digest=sha256:f0a270459a713948bc1c4b19d609c211726569e1eec653647fe4bdf21ca79081

Observation 865f1f13-391c-44c2-8d01-0426461d180b · outbound

This paper cites Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.082981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.082981Z digest=sha256:286c9b8dffc67321f6569d8cbcae81b655a8be3ad00f0979c96c94b1bd56b59c

Observation a3bd51e3-7e11-445c-b490-77522bcb375e · outbound

This paper cites Aligning large and small language models via chain-of-thought reasoning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Aligning large and small language models via chain-of-thought reasoning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.530060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.087259Z digest=sha256:5dda9eac854733d00b91d4ea8c7fd24ce45e275ef49b3f65940da5b7bad5e7aa

Observation de6b8527-f939-4f29-97c2-c8f351b2d215 · outbound

This paper cites Laion-400m: Open dataset of clip-filtered 400 million image-text pairs.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Laion-400m: Open dataset of clip-filtered 400 million image-text pairs

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.516556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.092576Z digest=sha256:479c671ac9519e164925c860169bd8b6ad9b67fa4da322d30d2f092120957c4e

Observation 6127810b-2b5a-498c-bfd3-905d4efbf858 · outbound

This paper cites Visual cot: Unleashing chain-of-thought reasoning in multi-modal language models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Visual cot: Unleashing chain-of-thought reasoning in multi-modal language models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.503066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.096740Z digest=sha256:7faebbf0146bcec4f6ad4bd0e624d1eb8b89ea3d88476934125573e1a5513395

Observation 3bcdec76-d803-4e29-ad25-d2a84edfcd44 · outbound

This paper cites Conceptual captions: A cleaned, hypernymed, im- age alt-text dataset for automatic image captioning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Conceptual captions: A cleaned, hypernymed, im- age alt-text dataset for automatic image captioning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.488546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.100783Z digest=sha256:12804206549876c205796f887a1aa42001a5af5bfbf95285063d379ca02223ab

Observation 181ad49c-580b-4edb-9735-fbf457ff94d6 · outbound

This paper cites Mome: Mixture of multimodal experts for generalist multimodal large language models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Mome: Mixture of multimodal experts for generalist multimodal large language models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.105148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.105148Z digest=sha256:9b84d6d66cb30fff9caac5027dccf924872840c0ffee8b4718d3e5ec9d140b97

Observation 3e38e282-e2d4-488f-aaf8-06f6ad2db1ce · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large lan- guage models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Chain-of-thought prompting elicits reasoning in large lan- guage models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.109588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.109588Z digest=sha256:95f149bb3eaeb56b6dd1d55667585350280bac0dd17a9c196a6deae7fc9b5008

Observation 33aafc2c-50ed-4f2a-8985-4d80ee31dcc9 · outbound

This paper cites Grounded Chain-of-Thought for Multimodal Large Language Models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Grounded Chain-of-Thought for Multimodal Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.113550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.113550Z digest=sha256:329f5b9bfea022980c87fe061fd0f6758be5a7513a76d65e8fba7b999dc18e55

Observation 4be37be1-5449-44a0-a5b5-3afe5f5a0f82 · outbound

This paper cites Visionary-r1: Mitigating shortcuts in vi- sual reasoning with reinforcement learning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Visionary-r1: Mitigating shortcuts in vi- sual reasoning with reinforcement learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.117687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.117687Z digest=sha256:ba6077e7f3be9023f6147c2ac2bf9412851934acf5ff49bc96b8145b2faa32d9

Observation 723aeddf-b3be-4a50-8615-2acb65734afc · outbound

This paper cites Falcon: Resolv- ing visual redundancy and fragmentation in high-resolution multimodal large language models via visual registers.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Falcon: Resolv- ing visual redundancy and fragmentation in high-resolution multimodal large language models via visual registers

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.456072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.121824Z digest=sha256:d4efe62024c67aa2c780633ebed34311d8695f8b810f3149b89558138752c3e4

Observation 8b606a23-4068-4979-ab81-d52df5b795f6 · outbound

This paper cites Tat-qa: A question answering benchmark on a hybrid of tab- ular and textual content in finance.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Tat-qa: A question answering benchmark on a hybrid of tab- ular and textual content in finance

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.442042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.126531Z digest=sha256:e8f9859ec43c24a6c99de1d4d7177c040429716c570cabb01e9b14a2f774fc55

Observation d1ea09ca-e4d5-477a-a6c8-8c56b48f928e · outbound

This paper cites Visual7w: Grounded question answering in images.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Visual7w: Grounded question answering in images

Reference 34

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T20:24:57.428003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.130597Z digest=sha256:09fbf274e7a10db1ca20555d87d160b82a4d5e0514aea7d547b6fcc740ba8d6d

Observation 1be332be-6cf8-4e31-9040-ebfaa80a41a8 · outbound

This paper cites an unresolved cited work.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:24:57.398146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.140629Z digest=sha256:02c289c090e3543d3389e05815f86502e167b87cc432c9c64973d2c77186a972

Observation 83a36de8-3632-41b7-93dc-e3e9ff4fc514 · outbound

This paper cites an unresolved cited work.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:24:57.383742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.145279Z digest=sha256:c2d35877625e6b8f0b5e61851ef551f7dd893a1067b52ec13b6915a870b53b7a

Observation dc3c4422-2004-4480-b7f5-4459846cb8a9 · outbound

This paper cites The table shows the number of fan letters for each day.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation The table shows the number of fan letters for each day

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.368364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.149476Z digest=sha256:ad0037a4e789ab80da695137ab3c791111649cf04c245783366dc0dba8b02af3

Observation c92ee23a-db34-49bb-b932-050b9e5fa0dd · outbound

This paper cites Opening Balance.\.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Opening Balance.\

Reference 39

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T20:24:57.352221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.154109Z digest=sha256:fdd5cd5ae08df189406483fab2c097dacbe41e5ee6b9ca5cda2e81f0f37c6e78

Observation bb02a9a4-8ac3-4115-bb41-4aa687633838 · outbound

This paper cites Thus, the total number of fan letters received on Thursday and Monday is: 204 + 271 = 475.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Thus, the total number of fan letters received on Thursday and Monday is: 204 + 271 = 475

Reference 204

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.413154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T20:24:57.135943Z digest=sha256:de4830697b7d3a75fa325880adc519aa1cee39462e0bac08f70fdc7316927538

Pith citing papers

Observation 71c9b1fc-2c69-43b6-8b4d-c87e0865845c · inbound

Balancing Image Compression and Generation with Bootstrapped Tokenization cites this paper.

Balancing Image Compression and Generation with Bootstrapped Tokenization Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-02T11:46:55.444845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T03:07:33.054518Z digest=sha256:654023ffa7b3dafb82a8c4003a7b32d7f95f0bea08ad97964827778441dd48ae

Observation e9852caa-6d97-453e-b997-8bb10ba5d27c · inbound

Answer-Conditioned Chain-of-Thought Distillation for Few-Shot Industrial Vision with Small VLMs cites this paper.

Answer-Conditioned Chain-of-Thought Distillation for Few-Shot Industrial Vision with Small VLMs Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-14T10:06:52.822171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:06:52.822171Z digest=sha256:5a184a0cb9f54341709e2125905f2493993c248a256ccf33be8789c638c878fd