Pith. sign in

Paper Citation Record · LEDGER

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy

As of 9 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 3 inbound Pith citation observations for arXiv:2506.09958.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09958 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:42:02.726920Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:57:51.206063Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T03:56:21.604762Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact4
  • verified fuzzy5
  • unresolved42
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1c52b270-b9ff-4969-b23b-a8c9878cbd98 · outbound

This paper cites Foundation metrics for evaluating effectiveness of healthcare conversations powered by generative AI.npj Digital Med., 7(82):1–14, March 2024.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Foundation metrics for evaluating effectiveness of healthcare conversations powered by generative AI.npj Digital Med., 7(82):1–14, March 2024

Reference 1

Resolution
malformed identifier
no resolver link, observed 2026-08-07T04:42:02.560852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.560852Z digest=sha256:2e1105db9535abb5230f626e9acff146b974dbc9952fe4acf25d7d1dc7ff0d0d

Observation a2c6ce8e-7a85-4ee8-bff5-7d04aa538d59 · outbound

This paper cites Flamingo: a Visual Language Model for Few-Shot Learning.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Flamingo: a Visual Language Model for Few-Shot Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.564894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.564894Z digest=sha256:63bcfb411d554c42e74659a9b78a50e862d95e619c0aa1c016fad99f351438fb

Observation a011d625-0122-4965-8059-336b9f3babff · outbound

This paper cites A deep learning framework for quality assessment and restoration in video endoscopy.arXiv, April 2019.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy A deep learning framework for quality assessment and restoration in video endoscopy.arXiv, April 2019

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.568482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.568482Z digest=sha256:bb87002276e78edb366dd1ec9741a30aa4fca3f15b4493f89ae241af29940a52

Observation 742143e1-a8a9-4ca3-8b5c-06eadc2925fb · outbound

This paper cites Qwen2.5-VL Technical Report.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.572023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.572023Z digest=sha256:878b6c1edbdc75a6b951c860f55267cd92b93b360116ddd5ae5a9f28cfba5d1b

Observation 9b01ca2c-d2e7-4f42-b1c4-a6b01b57a0d9 · outbound

This paper cites Vision–Language Model for Visual Question Answering in Medical Imagery.Bioengineering, 10(3):380, March 2023.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Vision–Language Model for Visual Question Answering in Medical Imagery.Bioengineering, 10(3):380, March 2023

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.575479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.575479Z digest=sha256:6257509ef3dc99b82fdaa47087535f17f2558bbe110b71f891cec620435d0018

Observation c8071b00-5997-42eb-b362-7c04b953944f · outbound

This paper cites Smedsrud, Steven Hicks, Debesh Jha, Sigrun L.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Smedsrud, Steven Hicks, Debesh Jha, Sigrun L

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.578778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.578778Z digest=sha256:b757aa0ff0e5e7ce33644fe94d4df9e809b038c4cb8402fc8ff247294769baf7

Observation c193399a-2514-4e52-8e1b-6bf57876587d · outbound

This paper cites Iglovikov, and Alexandr A.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Iglovikov, and Alexandr A

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.581769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.581769Z digest=sha256:4bae323a873e5549653cd475ad0a6b38df31f29496646f41afdcae564a74b2b3

Observation 4b9b7d8e-612e-4d26-8bcb-8f5485610ce8 · outbound

This paper cites A Simple Framework for Contrastive Learning of Visual Representations.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy A Simple Framework for Contrastive Learning of Visual Representations

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.584675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.584675Z digest=sha256:c751eb17bf3d2436036b0160e3f37c30ff7e5002b35576d8b7386a3f191e523a

Observation b1d98937-3b9b-4a47-92f4-fc1afd4bd9c3 · outbound

This paper cites R-LLaVA: Improving Med-VQA Understanding through Visual Region of Interest.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy R-LLaVA: Improving Med-VQA Understanding through Visual Region of Interest

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.588945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.588945Z digest=sha256:52c672d9353d4794374ce8c86783e4fee2645507cbb8aa1e6f4f9ce4d6a81f09

Observation e0f0c41c-5985-4413-907c-2d2629c589ea · outbound

This paper cites Generative Models in Medical Visual Question Answering: A Survey.Appl.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Generative Models in Medical Visual Question Answering: A Survey.Appl

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.591678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.591678Z digest=sha256:4da12dccd165f85df03e361d5b881a3164c5e539a8b14f94095ca9216efd0148

Observation a3b1350f-f88c-4339-96fa-4f4a2a8c620f · outbound

This paper cites LLM-based NLG Evaluation: Current Status and Challenges.Computational Linguistics, pages 1–27, 2025.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy LLM-based NLG Evaluation: Current Status and Challenges.Computational Linguistics, pages 1–27, 2025

Reference 11

Resolution
malformed identifier
no resolver link, observed 2026-08-07T04:42:02.594102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.594102Z digest=sha256:22adf997b993ca4676f3e7c92fd815e0c4e1ddd761a215a39dfbc496c44b020d

Observation 1c7bb54d-d855-4835-93c3-d9ae3e933a37 · outbound

This paper cites Hicks, Vajira Thambawita, P ˚ al Halvorsen, and Michael A.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Hicks, Vajira Thambawita, P ˚ al Halvorsen, and Michael A

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.597646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.597646Z digest=sha256:a60d5f61783e50aab52f90b833eb42689388388a0a90a1dce207c582afa9b238

Observation b8cbc6fb-8c62-40bc-bf01-60df620470cb · outbound

This paper cites Medgemma hugging face, May 2025.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Medgemma hugging face, May 2025

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.839627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:42:02.600361Z digest=sha256:1660b67b61a9901fe17b43df9e100b8c572de0aaf0b6d5c722d0c4dc306aa4dd

Observation 2d9bf78c-e0a5-4e92-8415-b4f5e3352157 · outbound

This paper cites LaPA: Latent Prompt Assist Model For Medical Visual Question Answering.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy LaPA: Latent Prompt Assist Model For Medical Visual Question Answering

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:42:03.038603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:42:02.603313Z digest=sha256:3d7d744ba21923fbb5c853c61d539148b7a770b6ff825561bfac1d4bb04665b9

Observation 059da0cf-f138-4789-8bb7-8129831f33a1 · outbound

This paper cites DiN: Diffusion Model for Robust Medical VQA with Semantic Noisy Labels.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy DiN: Diffusion Model for Robust Medical VQA with Semantic Noisy Labels

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:42:03.024840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:42:02.606380Z digest=sha256:55c84288edd84bba798e8f9704a33ed94e70fa03307c5aea3afff207158e29b5

Observation 73acb3fc-ee4d-4efb-9116-8e12b4afe5dd · outbound

This paper cites Vision-language models for medical report generation and visual question answering: a review.Front.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Vision-language models for medical report generation and visual question answering: a review.Front

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.609354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.609354Z digest=sha256:c482826be47915c3d90d7e213ce432eaa04dd27d5a48eba9896fd04156157447

Observation ce258bb0-d390-4241-ad2d-c3a69acb344b · outbound

This paper cites DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.612276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.612276Z digest=sha256:3175fce3c5da21e20a4c37024b269cc3b858107b50a732854211549a70500ef9

Observation ba445f8d-489f-472a-be4c-234269813756 · outbound

This paper cites Overview of imageclefmedical 2023-medical visual question answering for gastrointestinal tract.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Overview of imageclefmedical 2023-medical visual question answering for gastrointestinal tract

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.613594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:42:02.617533Z digest=sha256:7985e0bbdc3d54c9bce565a3d7f185044e75286e5f32000f7cc8f33bd8174331

Observation 0dc39b54-26f4-4063-9f9c-6e19054be7dc · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy LoRA: Low-Rank Adaptation of Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.620794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.620794Z digest=sha256:0cafe59776a83a1a0f2b7fae5a726b6c40590ae32465c60782e1c86ca64389ed

Observation 8aee339d-ca74-4715-94d7-19d6be5718b2 · outbound

This paper cites Summers, and Yingying Zhu.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Summers, and Yingying Zhu

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.624394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.624394Z digest=sha256:0278884cb591e610fb3d0a201f9952a1ce43f467056a19ee8cd2add777c3324a

Observation 84594a3c-9911-445d-91f4-5d41e69b0f21 · outbound

This paper cites Sadman Hafiz, Jamin Rahman Jim, Md.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Sadman Hafiz, Jamin Rahman Jim, Md

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.627513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.627513Z digest=sha256:cb7eeebdf0ddc5485f1e14052d76369931be57146977e9cd5d9f3b7cf20cdcf3

Observation eff2e767-9c63-482a-920e-2a60e8a62874 · outbound

This paper cites Hicks, Vajira Thambawita, Enrique Garcia-Ceja, Michael A.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Hicks, Vajira Thambawita, Enrique Garcia-Ceja, Michael A

Reference 23

Resolution
verified exact
doi, observed 2026-08-07T04:42:02.991850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:42:02.631069Z digest=sha256:6ba7b5419b01625e5bd8bfd9d5a5ed376c730f3e55a7bf6e0f972a4400beaa9b

Observation bb2fe99e-3373-4845-ab3f-2a81e219314a · outbound

This paper cites Lau, Soumya Gayen, Asma Ben Abacha, and Dina Demner-Fushman.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Lau, Soumya Gayen, Asma Ben Abacha, and Dina Demner-Fushman

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.633750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.633750Z digest=sha256:9cd96d95bd1f4538561fd9fff2b783ad10d4378be9195d221b8616fff793f82d

Observation d8645d43-43fa-4eed-ad94-0683aa80f03b · outbound

This paper cites Meteor: an automatic metric for MT evaluation with high levels of correlation with human judgments.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Meteor: an automatic metric for MT evaluation with high levels of correlation with human judgments

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.636692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.636692Z digest=sha256:2b51b8dff6f93691d872f813a3476c057189ac27e6148a083934933802db44a5

Observation 9b4184d9-7b52-4b7d-a755-5e76e336007f · outbound

This paper cites LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.639446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.639446Z digest=sha256:69eb4d4e722c8a24fdba1fd7a68efb6be4cbae0b768c009be6323df882b86540

Observation 6374a26f-cdb1-4954-b58a-21362fd715d5 · outbound

This paper cites Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.642311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.642311Z digest=sha256:2a006c503b9837c202d5ee911ce5bee3f158f2e326c13ba4f16e9f450333b8a3

Observation 5f2a347e-1273-423f-8ea4-455a9a976e34 · outbound

This paper cites Candidate-Heuristic In-Context Learning: A new framework for enhancing medical visual question answering with LLMs.Information Processing & Management, 61(5):103805, September 2024.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Candidate-Heuristic In-Context Learning: A new framework for enhancing medical visual question answering with LLMs.Information Processing & Management, 61(5):103805, September 2024

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.645547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.645547Z digest=sha256:a1a3bb06d7d7f9d5b01f5be0b4740aca95e522cb824c1319f5f4a8957db79082

Observation b3ee46df-b7b5-4e5a-86c2-0948147c2b90 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Rouge: A package for automatic evaluation of summaries

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.648531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.648531Z digest=sha256:969fb288347024f26be85c8b4433ee184bb65a74b69b2c54caf8d19842312b2a

Observation 2c6ec00c-b26a-4396-806b-9c645233b806 · outbound

This paper cites Medical visual question answering: A survey.Artif.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Medical visual question answering: A survey.Artif

Reference 30

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T04:42:03.491149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:42:02.651405Z digest=sha256:acadd55ec75f7f66a34396eb521753b169634260ae1d2a58a9601320ad858bf7

Observation 4927d41e-e237-47b8-aa8b-a2ea1a5640b1 · outbound

This paper cites Slake: A Semantically-Labeled Knowledge-Enhanced Dataset For Medical Visual Question Answering.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Slake: A Semantically-Labeled Knowledge-Enhanced Dataset For Medical Visual Question Answering

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.654115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.654115Z digest=sha256:54c7e23a96bbca4f008aa185c56e7c1fa78fa44b03c4a7ca4a97f14ab85bf149

Observation 454e4dff-4980-4af8-8a5a-ac47abd2aee0 · outbound

This paper cites Visual Instruction Tuning.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Visual Instruction Tuning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.657045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.657045Z digest=sha256:6c22a93721db2be66e1b1805a6a734fdeee082038a56ca672a77df0e6675f504

Observation 35c11c11-694e-4dd3-92a9-a62c5d80325b · outbound

This paper cites Peft: State-of- the-art parameter-efficient fine-tuning methods.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Peft: State-of- the-art parameter-efficient fine-tuning methods

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.463784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:42:02.659998Z digest=sha256:aa5c11622ecb7d3f0e9f06d6d467caed72106e2e1890d80f00ba8424149a7d1f

Observation 2f0cbe2a-c590-420f-9b06-cbd8b6b11149 · outbound

This paper cites Med-Flamingo: a Multimodal Medical Few-shot Learner.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Med-Flamingo: a Multimodal Medical Few-shot Learner

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.665674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.665674Z digest=sha256:6357b33a6ffc7806a56e087a87f91e0c2d411f2f50d1469ef3503ed868bcdc5e

Observation 27c44666-2662-4c50-8ac3-b0c032f09de0 · outbound

This paper cites GPT-4 Technical Report.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy GPT-4 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.668362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.668362Z digest=sha256:65142a332c83c48836e41ad949d01456a248004be241793a7543f8ea145d5356

Observation 63bbee4d-3107-4a92-94b2-4d782acf47e3 · outbound

This paper cites Chaudhari, et al.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Chaudhari, et al

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.671397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.671397Z digest=sha256:d7e2abedd897eb6bebc35fa6f12b0cd14a2b86447e36a36866f59dfba51f6662

Observation de414271-f059-4fcd-86d9-dd21f162ea55 · outbound

This paper cites BLEU: a method for automatic evaluation of ma- chine translation.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy BLEU: a method for automatic evaluation of ma- chine translation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.673870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.673870Z digest=sha256:fcbddea609558fe0cfb40438c1b90071f13441ce4705d63fbe14162cbde6586d

Observation c9f0e415-519b-4534-b083-ff7561b001d0 · outbound

This paper cites chrF: character n-gram F-score for automatic MT evaluation.ACL Anthology, pages 392–395, September.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy chrF: character n-gram F-score for automatic MT evaluation.ACL Anthology, pages 392–395, September

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.314693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:42:02.676750Z digest=sha256:57fd606c7d43691c4140e2d925219de556533f6a49316f17f175a242cfc1663c

Observation 71e43144-29fe-42b6-bee7-f243b85df2ec · outbound

This paper cites ZeRO: Memory Optimizations Toward Training Trillion Parameter Models.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy ZeRO: Memory Optimizations Toward Training Trillion Parameter Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.682342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.682342Z digest=sha256:0639660cd9be70d32cb7eb3e338559cb31c6f665a9442836b6e1ee76a49af0f4

Observation 12a40f5f-3c5e-486e-86aa-fae06d9bfec2 · outbound

This paper cites Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.685440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.685440Z digest=sha256:bde92b37898b5c27a8b2044003cbb593bcd9109bda7ebdd306f790950d3a4388

Observation 3d5da4d0-b840-4da3-a59f-969e4b5f2476 · outbound

This paper cites BLEURT: Learning Robust Metrics for Text Generation.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy BLEURT: Learning Robust Metrics for Text Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.689211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.689211Z digest=sha256:477690add4ba4112f3f9a7f24ab574434a620493eefdfe1f64e541ec28b727e1

Observation 24d49581-6a9f-44fb-848d-238a30735c5b · outbound

This paper cites Pfohl, Heather Cole-Lewis, et al.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Pfohl, Heather Cole-Lewis, et al

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.692791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.692791Z digest=sha256:fb6e7ea16596c417b085e68a5835942f4422a8bd0bc6948859333eb1e3128936

Observation c7b19ffe-ac7f-4563-9322-33b9c37baa25 · outbound

This paper cites Qwen3 Technical Report.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Qwen3 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.695513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.695513Z digest=sha256:0a2b1ae84f43641712519f8c3cd2b8fc0a4d8f6f1b8310e8c426c9d93a9a6605

Observation ea035ea9-0fd3-475e-bb1f-20b208f33a15 · outbound

This paper cites Sanders, Yuchen Liu, Kennarey Seang, Bach Xuan Tran, Atanas G.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Sanders, Yuchen Liu, Kennarey Seang, Bach Xuan Tran, Atanas G

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.698427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.698427Z digest=sha256:e248a38f3fc32bf845da3441c0c164f83cacd27e8e73730b882010d1ae478857

Observation 61c26755-4ced-440d-9ee4-61af9d046faf · outbound

This paper cites Wilkinson, Michel Dumontier, IJsbrand Jan Aalbersberg, Gabrielle Appleton, Myles Axton, Arie Baak, Niklas Blomberg, Jan-Willem Boiten, Luiz Bonino da Silva Santos, Philip E.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Wilkinson, Michel Dumontier, IJsbrand Jan Aalbersberg, Gabrielle Appleton, Myles Axton, Arie Baak, Niklas Blomberg, Jan-Willem Boiten, Luiz Bonino da Silva Santos, Philip E

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.701389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.701389Z digest=sha256:1b62fd513f7273c9307cc209e627fa4edb889491e2c3bd905c5e15608244d49e

Observation fab20e9a-a383-49f6-9eb6-f9752df995cf · outbound

This paper cites an unresolved cited work.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:42:04.164751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:42:02.704076Z digest=sha256:c5138d893192bd574f0246f94cc5c71db1b70de9ba2eb430af0176ec40a89e3b

Observation e22cee96-9e5c-420f-825a-ebdb89da6b06 · outbound

This paper cites Parameter-Efficient Fine-Tuning Methods for Pretrained Language Models: A Critical Review and Assessment.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Parameter-Efficient Fine-Tuning Methods for Pretrained Language Models: A Critical Review and Assessment

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.706457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.706457Z digest=sha256:53fdbf8d990b73c66b21c7dca9734d0cec89fd69ee395e9e3c730f3a2187cb11

Observation ede2d0b1-2c9f-4fa3-b54b-9e6713528d75 · outbound

This paper cites MedFrameQA: A Multi-Image Medical VQA Benchmark for Clinical Reasoning.arXiv, May 2025.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy MedFrameQA: A Multi-Image Medical VQA Benchmark for Clinical Reasoning.arXiv, May 2025

Reference 49

Resolution
verified exact
doi, observed 2026-08-07T04:42:02.854127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:42:02.709606Z digest=sha256:d82558140794311adf09f507ebaded6753e1ea6b941c81392565a65eb811f0fe

Observation f5190ec1-fb2e-44d6-8a72-8eb1b5c4be42 · outbound

This paper cites Fine-grained Adaptive Visual Prompt for Generative Medical Visual Question Answering.AAAI, 39(9):9662–9670, April 2025.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Fine-grained Adaptive Visual Prompt for Generative Medical Visual Question Answering.AAAI, 39(9):9662–9670, April 2025

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.712182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.712182Z digest=sha256:501b9f7b1cfeb848e65b59496f2c68b7144d172ea5bdb5c644cb0e40f74b9580

Observation 5f569dee-710c-4b5a-8fbc-8f61c8bd0d7e · outbound

This paper cites Medical Visual Question Answering via Conditional Reasoning.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Medical Visual Question Answering via Conditional Reasoning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.013804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:42:02.715622Z digest=sha256:6a0f57a71cdfdc76733dee51bb7dd1d593220b2c98a7ece2ed0d495ceebad1cd

Observation 9685044a-1a2a-4a01-b705-e61c3f63213f · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy BERTScore: Evaluating Text Generation with BERT

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.720800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.720800Z digest=sha256:60460b3c77afebe25460a2e6c8a45643a366be445ca1cab491e0f51574758ff8

Observation fb0332a3-b5c3-4810-82d8-acbe4ada430d · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.723789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.723789Z digest=sha256:e51776ea4a99c50fd9377fbd7bb4660ad2e5d168ec05c5acc1f16a8d3e4edc97

Observation 24983bd3-10c5-4138-8d99-7ff0ad5ad655 · outbound

This paper cites SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.726920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.726920Z digest=sha256:3cc4740dd62921ad491d91496bac909b27569d0fcc55c1807c0f97926835ffcf

Observation 9973781f-b476-41fa-82b6-60c86ed56032 · outbound

This paper cites an unresolved cited work.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Unresolved cited work

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.679485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.679485Z digest=sha256:27462c5c0ff754a87455bdf7337bc2101f50dec0e5fc7ab4255a025c9030146d

Observation 7fd676bb-74cc-4ad6-8035-32ac396b560a · outbound

This paper cites an unresolved cited work.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Unresolved cited work

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.718367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.718367Z digest=sha256:cb3e7d3c719fb6a785831bdff57609d9b7c23480fd2dcae9d7f07a2c2ba8f407

Pith citing papers

Observation f0beebc1-9999-4b7e-97b3-738c0a256f0c · inbound

Multimodal AI for Gastrointestinal Diagnostics: Tackling VQA in MEDVQA-GI 2025 cites this paper.

Multimodal AI for Gastrointestinal Diagnostics: Tackling VQA in MEDVQA-GI 2025 Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:57:51.206063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:57:51.206063Z digest=sha256:0427df8cca6a760bdd986feefc8ef65cbecab9d4d3fa5d87d70e0eaca5c3d2e3

Observation 16a6d077-dde9-4722-9d55-5f74cf186cc2 · inbound

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology cites this paper.

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:56:21.606430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:55:55.359488Z digest=sha256:df815c959f694dfd5f6d59308e3c62d91fef4037914c7a5f0db66b2ba36b3d3c

Observation e4faedd8-0162-413c-94cc-6b6cde36c72b · inbound

Measuring and Improving Complex-Atomic Answer Consistency in Endoscopic VQA cites this paper.

Measuring and Improving Complex-Atomic Answer Consistency in Endoscopic VQA Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T16:57:30.636879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:57:30.636879Z digest=sha256:188351454a997518589627f350499ffd16a6b668d85483fd132972254506d96a