Pith. sign in

Paper Citation Record · LEDGER

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy

As of 22 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 3 inbound Pith citation observations for arXiv:2506.09958.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09958 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:42:02.726920Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:57:51.206063Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T03:56:21.604762Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact4
  • verified fuzzy5
  • unresolved42
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1c52b270-b9ff-4969-b23b-a8c9878cbd98 · outbound

This paper cites Foundation metrics for evaluating effectiveness of healthcare conversations powered by generative AI.npj Digital Med., 7(82):1–14, March 2024.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Foundation metrics for evaluating effectiveness of healthcare conversations powered by generative AI.npj Digital Med., 7(82):1–14, March 2024

Reference 1

Resolution
malformed identifier
no resolver link, observed 2026-08-07T04:42:02.560852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.560852Z digest=sha256:2e1105db9535abb5230f626e9acff146b974dbc9952fe4acf25d7d1dc7ff0d0d

Observation a2c6ce8e-7a85-4ee8-bff5-7d04aa538d59 · outbound

This paper cites Flamingo: a Visual Language Model for Few-Shot Learning.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Flamingo: a Visual Language Model for Few-Shot Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.564894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.564894Z digest=sha256:d54cdf4de6832dca1c7bc3758ae337bcb122d6a2b7084999151698bcfb7012a7

Observation a011d625-0122-4965-8059-336b9f3babff · outbound

This paper cites A deep learning framework for quality assessment and restoration in video endoscopy.arXiv, April 2019.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy A deep learning framework for quality assessment and restoration in video endoscopy.arXiv, April 2019

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.568482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.568482Z digest=sha256:bb87002276e78edb366dd1ec9741a30aa4fca3f15b4493f89ae241af29940a52

Observation 742143e1-a8a9-4ca3-8b5c-06eadc2925fb · outbound

This paper cites Qwen2.5-VL Technical Report.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.572023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.572023Z digest=sha256:206cabbd95a7ff7a7d1eafa13819f12f2c128a6f59ba7f704c153fb82a4d1d22

Observation 9b01ca2c-d2e7-4f42-b1c4-a6b01b57a0d9 · outbound

This paper cites Vision–Language Model for Visual Question Answering in Medical Imagery.Bioengineering, 10(3):380, March 2023.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Vision–Language Model for Visual Question Answering in Medical Imagery.Bioengineering, 10(3):380, March 2023

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.575479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.575479Z digest=sha256:6257509ef3dc99b82fdaa47087535f17f2558bbe110b71f891cec620435d0018

Observation c8071b00-5997-42eb-b362-7c04b953944f · outbound

This paper cites Smedsrud, Steven Hicks, Debesh Jha, Sigrun L.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Smedsrud, Steven Hicks, Debesh Jha, Sigrun L

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.578778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.578778Z digest=sha256:b757aa0ff0e5e7ce33644fe94d4df9e809b038c4cb8402fc8ff247294769baf7

Observation c193399a-2514-4e52-8e1b-6bf57876587d · outbound

This paper cites Iglovikov, and Alexandr A.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Iglovikov, and Alexandr A

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.581769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.581769Z digest=sha256:4bae323a873e5549653cd475ad0a6b38df31f29496646f41afdcae564a74b2b3

Observation 4b9b7d8e-612e-4d26-8bcb-8f5485610ce8 · outbound

This paper cites A Simple Framework for Contrastive Learning of Visual Representations.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy A Simple Framework for Contrastive Learning of Visual Representations

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.584675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.584675Z digest=sha256:c323e3eb58e9f31e61adf18ba0ec7e6b64fae95915efa6356fe07d0f77556260

Observation b1d98937-3b9b-4a47-92f4-fc1afd4bd9c3 · outbound

This paper cites R-LLaVA: Improving Med-VQA Understanding through Visual Region of Interest.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy R-LLaVA: Improving Med-VQA Understanding through Visual Region of Interest

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.588945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.588945Z digest=sha256:c875d16dfe444eee9324d19ce327992c8759f0a74bf647633ed9f12e9838ae17

Observation e0f0c41c-5985-4413-907c-2d2629c589ea · outbound

This paper cites Generative Models in Medical Visual Question Answering: A Survey.Appl.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Generative Models in Medical Visual Question Answering: A Survey.Appl

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.591678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.591678Z digest=sha256:4da12dccd165f85df03e361d5b881a3164c5e539a8b14f94095ca9216efd0148

Observation a3b1350f-f88c-4339-96fa-4f4a2a8c620f · outbound

This paper cites LLM-based NLG Evaluation: Current Status and Challenges.Computational Linguistics, pages 1–27, 2025.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy LLM-based NLG Evaluation: Current Status and Challenges.Computational Linguistics, pages 1–27, 2025

Reference 11

Resolution
malformed identifier
no resolver link, observed 2026-08-07T04:42:02.594102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.594102Z digest=sha256:22adf997b993ca4676f3e7c92fd815e0c4e1ddd761a215a39dfbc496c44b020d

Observation 1c7bb54d-d855-4835-93c3-d9ae3e933a37 · outbound

This paper cites Hicks, Vajira Thambawita, P ˚ al Halvorsen, and Michael A.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Hicks, Vajira Thambawita, P ˚ al Halvorsen, and Michael A

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.597646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.597646Z digest=sha256:a60d5f61783e50aab52f90b833eb42689388388a0a90a1dce207c582afa9b238

Observation b8cbc6fb-8c62-40bc-bf01-60df620470cb · outbound

This paper cites Medgemma hugging face, May 2025.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Medgemma hugging face, May 2025

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.839627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T04:42:02.600361Z digest=sha256:cae7073e8467ee950636cb0e2ebfde05d7e8084ddd6737d8132be0478f903cd7

Observation 2d9bf78c-e0a5-4e92-8415-b4f5e3352157 · outbound

This paper cites LaPA: Latent Prompt Assist Model For Medical Visual Question Answering.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy LaPA: Latent Prompt Assist Model For Medical Visual Question Answering

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:42:03.038603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T04:42:02.603313Z digest=sha256:cd8473349bbb62163a3d60ec79401e1e04cbc4d5b782c10636334f667fbb7042

Observation 059da0cf-f138-4789-8bb7-8129831f33a1 · outbound

This paper cites DiN: Diffusion Model for Robust Medical VQA with Semantic Noisy Labels.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy DiN: Diffusion Model for Robust Medical VQA with Semantic Noisy Labels

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:42:03.024840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T04:42:02.606380Z digest=sha256:8b615292bade391546c2637e711a3ccd6fbc96f0ebb34f4b8cc546ab4052424c

Observation 73acb3fc-ee4d-4efb-9116-8e12b4afe5dd · outbound

This paper cites Vision-language models for medical report generation and visual question answering: a review.Front.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Vision-language models for medical report generation and visual question answering: a review.Front

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.609354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.609354Z digest=sha256:c482826be47915c3d90d7e213ce432eaa04dd27d5a48eba9896fd04156157447

Observation ce258bb0-d390-4241-ad2d-c3a69acb344b · outbound

This paper cites DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.612276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.612276Z digest=sha256:9589112ab7c8f0f26f42095f5e5310502863f256badb5eb66c9e549152006b14

Observation ba445f8d-489f-472a-be4c-234269813756 · outbound

This paper cites Overview of imageclefmedical 2023-medical visual question answering for gastrointestinal tract.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Overview of imageclefmedical 2023-medical visual question answering for gastrointestinal tract

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.613594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T04:42:02.617533Z digest=sha256:072bbb6dbc644012c8569d73f25f23b6eb998caebb2f22b27017aac361ea4fac

Observation 0dc39b54-26f4-4063-9f9c-6e19054be7dc · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy LoRA: Low-Rank Adaptation of Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.620794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.620794Z digest=sha256:a9e6ca9ded7eed4b2a3854c00fb68b5c8c1718c329d9ce425e9b384841f8db81

Observation 8aee339d-ca74-4715-94d7-19d6be5718b2 · outbound

This paper cites Summers, and Yingying Zhu.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Summers, and Yingying Zhu

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.624394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.624394Z digest=sha256:0278884cb591e610fb3d0a201f9952a1ce43f467056a19ee8cd2add777c3324a

Observation 84594a3c-9911-445d-91f4-5d41e69b0f21 · outbound

This paper cites Sadman Hafiz, Jamin Rahman Jim, Md.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Sadman Hafiz, Jamin Rahman Jim, Md

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.627513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.627513Z digest=sha256:cb7eeebdf0ddc5485f1e14052d76369931be57146977e9cd5d9f3b7cf20cdcf3

Observation eff2e767-9c63-482a-920e-2a60e8a62874 · outbound

This paper cites Hicks, Vajira Thambawita, Enrique Garcia-Ceja, Michael A.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Hicks, Vajira Thambawita, Enrique Garcia-Ceja, Michael A

Reference 23

Resolution
verified exact
doi, observed 2026-08-07T04:42:02.991850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T04:42:02.631069Z digest=sha256:da4f0d7de2905140418eea9bbe06a57c7b60833fe46d8ee1a048dee02a45e295

Observation bb2fe99e-3373-4845-ab3f-2a81e219314a · outbound

This paper cites Lau, Soumya Gayen, Asma Ben Abacha, and Dina Demner-Fushman.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Lau, Soumya Gayen, Asma Ben Abacha, and Dina Demner-Fushman

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.633750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.633750Z digest=sha256:9cd96d95bd1f4538561fd9fff2b783ad10d4378be9195d221b8616fff793f82d

Observation d8645d43-43fa-4eed-ad94-0683aa80f03b · outbound

This paper cites Meteor: an automatic metric for MT evaluation with high levels of correlation with human judgments.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Meteor: an automatic metric for MT evaluation with high levels of correlation with human judgments

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.636692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.636692Z digest=sha256:2b51b8dff6f93691d872f813a3476c057189ac27e6148a083934933802db44a5

Observation 9b4184d9-7b52-4b7d-a755-5e76e336007f · outbound

This paper cites LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.639446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.639446Z digest=sha256:69eb4d4e722c8a24fdba1fd7a68efb6be4cbae0b768c009be6323df882b86540

Observation 6374a26f-cdb1-4954-b58a-21362fd715d5 · outbound

This paper cites Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.642311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.642311Z digest=sha256:57dc48b5c104d385321daff93c1c9ba0044a9c3445a32891f8fed1ed861b7a20

Observation 5f2a347e-1273-423f-8ea4-455a9a976e34 · outbound

This paper cites Candidate-Heuristic In-Context Learning: A new framework for enhancing medical visual question answering with LLMs.Information Processing & Management, 61(5):103805, September 2024.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Candidate-Heuristic In-Context Learning: A new framework for enhancing medical visual question answering with LLMs.Information Processing & Management, 61(5):103805, September 2024

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.645547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.645547Z digest=sha256:a1a3bb06d7d7f9d5b01f5be0b4740aca95e522cb824c1319f5f4a8957db79082

Observation b3ee46df-b7b5-4e5a-86c2-0948147c2b90 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Rouge: A package for automatic evaluation of summaries

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.648531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.648531Z digest=sha256:969fb288347024f26be85c8b4433ee184bb65a74b69b2c54caf8d19842312b2a

Observation 2c6ec00c-b26a-4396-806b-9c645233b806 · outbound

This paper cites Medical visual question answering: A survey.Artif.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Medical visual question answering: A survey.Artif

Reference 30

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T04:42:03.491149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T04:42:02.651405Z digest=sha256:a85fb815b591f73d9514be806c7aba7a855d6fd801805a4ccf8746d0d83c6037

Observation 4927d41e-e237-47b8-aa8b-a2ea1a5640b1 · outbound

This paper cites Slake: A Semantically-Labeled Knowledge-Enhanced Dataset For Medical Visual Question Answering.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Slake: A Semantically-Labeled Knowledge-Enhanced Dataset For Medical Visual Question Answering

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.654115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.654115Z digest=sha256:54c7e23a96bbca4f008aa185c56e7c1fa78fa44b03c4a7ca4a97f14ab85bf149

Observation 454e4dff-4980-4af8-8a5a-ac47abd2aee0 · outbound

This paper cites Visual Instruction Tuning.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Visual Instruction Tuning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.657045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.657045Z digest=sha256:6c22a93721db2be66e1b1805a6a734fdeee082038a56ca672a77df0e6675f504

Observation 35c11c11-694e-4dd3-92a9-a62c5d80325b · outbound

This paper cites Peft: State-of- the-art parameter-efficient fine-tuning methods.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Peft: State-of- the-art parameter-efficient fine-tuning methods

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.463784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T04:42:02.659998Z digest=sha256:7b11b112d419d82f6dd76e6b9c91ac867b67acc23b00237cf90ae7b3c9cfed0c

Observation 2f0cbe2a-c590-420f-9b06-cbd8b6b11149 · outbound

This paper cites Med-Flamingo: a Multimodal Medical Few-shot Learner.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Med-Flamingo: a Multimodal Medical Few-shot Learner

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.665674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.665674Z digest=sha256:27a59fe412f924711fc0abf75fd92b6f38219a39978167bb811a0418e34f6183

Observation 27c44666-2662-4c50-8ac3-b0c032f09de0 · outbound

This paper cites GPT-4 Technical Report.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy GPT-4 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.668362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.668362Z digest=sha256:f101c2eca1f5ee187deaf9ca9f7c6a8392967a30b796c3593f59833e8fc88c7a

Observation 63bbee4d-3107-4a92-94b2-4d782acf47e3 · outbound

This paper cites Chaudhari, et al.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Chaudhari, et al

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.671397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.671397Z digest=sha256:d7e2abedd897eb6bebc35fa6f12b0cd14a2b86447e36a36866f59dfba51f6662

Observation de414271-f059-4fcd-86d9-dd21f162ea55 · outbound

This paper cites BLEU: a method for automatic evaluation of ma- chine translation.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy BLEU: a method for automatic evaluation of ma- chine translation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.673870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.673870Z digest=sha256:fcbddea609558fe0cfb40438c1b90071f13441ce4705d63fbe14162cbde6586d

Observation c9f0e415-519b-4534-b083-ff7561b001d0 · outbound

This paper cites chrF: character n-gram F-score for automatic MT evaluation.ACL Anthology, pages 392–395, September.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy chrF: character n-gram F-score for automatic MT evaluation.ACL Anthology, pages 392–395, September

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.314693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T04:42:02.676750Z digest=sha256:f469786b9d0c0d765f2c13708f6595706b27ac4642e31682e1bcd90441d3e746

Observation 71e43144-29fe-42b6-bee7-f243b85df2ec · outbound

This paper cites ZeRO: Memory Optimizations Toward Training Trillion Parameter Models.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy ZeRO: Memory Optimizations Toward Training Trillion Parameter Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.682342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.682342Z digest=sha256:0639660cd9be70d32cb7eb3e338559cb31c6f665a9442836b6e1ee76a49af0f4

Observation 12a40f5f-3c5e-486e-86aa-fae06d9bfec2 · outbound

This paper cites Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.685440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.685440Z digest=sha256:f312e240e61dcd8517fc3520c06d79fcef2352d3378cb4cd98be9a452f9e3a5e

Observation 3d5da4d0-b840-4da3-a59f-969e4b5f2476 · outbound

This paper cites BLEURT: Learning Robust Metrics for Text Generation.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy BLEURT: Learning Robust Metrics for Text Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.689211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.689211Z digest=sha256:f88309f300c0e325a4b86c6166bbcdc2482eee325f0b3fc07ed19125ab2d2528

Observation 24d49581-6a9f-44fb-848d-238a30735c5b · outbound

This paper cites Pfohl, Heather Cole-Lewis, et al.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Pfohl, Heather Cole-Lewis, et al

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.692791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.692791Z digest=sha256:fb6e7ea16596c417b085e68a5835942f4422a8bd0bc6948859333eb1e3128936

Observation c7b19ffe-ac7f-4563-9322-33b9c37baa25 · outbound

This paper cites Qwen3 Technical Report.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Qwen3 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.695513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.695513Z digest=sha256:0a2b1ae84f43641712519f8c3cd2b8fc0a4d8f6f1b8310e8c426c9d93a9a6605

Observation ea035ea9-0fd3-475e-bb1f-20b208f33a15 · outbound

This paper cites Sanders, Yuchen Liu, Kennarey Seang, Bach Xuan Tran, Atanas G.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Sanders, Yuchen Liu, Kennarey Seang, Bach Xuan Tran, Atanas G

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.698427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.698427Z digest=sha256:e248a38f3fc32bf845da3441c0c164f83cacd27e8e73730b882010d1ae478857

Observation 61c26755-4ced-440d-9ee4-61af9d046faf · outbound

This paper cites Wilkinson, Michel Dumontier, IJsbrand Jan Aalbersberg, Gabrielle Appleton, Myles Axton, Arie Baak, Niklas Blomberg, Jan-Willem Boiten, Luiz Bonino da Silva Santos, Philip E.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Wilkinson, Michel Dumontier, IJsbrand Jan Aalbersberg, Gabrielle Appleton, Myles Axton, Arie Baak, Niklas Blomberg, Jan-Willem Boiten, Luiz Bonino da Silva Santos, Philip E

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.701389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.701389Z digest=sha256:1b62fd513f7273c9307cc209e627fa4edb889491e2c3bd905c5e15608244d49e

Observation fab20e9a-a383-49f6-9eb6-f9752df995cf · outbound

This paper cites an unresolved cited work.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:42:04.164751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T04:42:02.704076Z digest=sha256:b566a882f07a91999e975f72f9b424e5a2fb0cafb81d59ed5bc128a917d39633

Observation e22cee96-9e5c-420f-825a-ebdb89da6b06 · outbound

This paper cites Parameter-Efficient Fine-Tuning Methods for Pretrained Language Models: A Critical Review and Assessment.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Parameter-Efficient Fine-Tuning Methods for Pretrained Language Models: A Critical Review and Assessment

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.706457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.706457Z digest=sha256:c3267791dbe5698b910f04234236c0cee2ba4c649452fb271babfc8f80db6714

Observation ede2d0b1-2c9f-4fa3-b54b-9e6713528d75 · outbound

This paper cites MedFrameQA: A Multi-Image Medical VQA Benchmark for Clinical Reasoning.arXiv, May 2025.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy MedFrameQA: A Multi-Image Medical VQA Benchmark for Clinical Reasoning.arXiv, May 2025

Reference 49

Resolution
verified exact
doi, observed 2026-08-07T04:42:02.854127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T04:42:02.709606Z digest=sha256:62684fd2fece6fc7082dd5006a3010b7542f5b6352f1d4146187bd598b215924

Observation f5190ec1-fb2e-44d6-8a72-8eb1b5c4be42 · outbound

This paper cites Fine-grained Adaptive Visual Prompt for Generative Medical Visual Question Answering.AAAI, 39(9):9662–9670, April 2025.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Fine-grained Adaptive Visual Prompt for Generative Medical Visual Question Answering.AAAI, 39(9):9662–9670, April 2025

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.712182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.712182Z digest=sha256:501b9f7b1cfeb848e65b59496f2c68b7144d172ea5bdb5c644cb0e40f74b9580

Observation 5f569dee-710c-4b5a-8fbc-8f61c8bd0d7e · outbound

This paper cites Medical Visual Question Answering via Conditional Reasoning.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Medical Visual Question Answering via Conditional Reasoning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.013804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T04:42:02.715622Z digest=sha256:26411c3bcd8f0ea4c46020de450829969add79f6b13b44a99e1a518badcc17e7

Observation 9685044a-1a2a-4a01-b705-e61c3f63213f · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy BERTScore: Evaluating Text Generation with BERT

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.720800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.720800Z digest=sha256:60460b3c77afebe25460a2e6c8a45643a366be445ca1cab491e0f51574758ff8

Observation fb0332a3-b5c3-4810-82d8-acbe4ada430d · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.723789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.723789Z digest=sha256:a41bee2a5d044db1f127a2ba05b043b6b4f862ad5c257755fa44c0fce86b4449

Observation 24983bd3-10c5-4138-8d99-7ff0ad5ad655 · outbound

This paper cites SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.726920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.726920Z digest=sha256:2a0db8fd5f176b8425d5c3986bd9fb4f90beecad960a13c1c6c2889fb08dacbd

Observation 9973781f-b476-41fa-82b6-60c86ed56032 · outbound

This paper cites an unresolved cited work.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Unresolved cited work

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.679485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.679485Z digest=sha256:27462c5c0ff754a87455bdf7337bc2101f50dec0e5fc7ab4255a025c9030146d

Observation 7fd676bb-74cc-4ad6-8035-32ac396b560a · outbound

This paper cites an unresolved cited work.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Unresolved cited work

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.718367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.718367Z digest=sha256:cb3e7d3c719fb6a785831bdff57609d9b7c23480fd2dcae9d7f07a2c2ba8f407

Pith citing papers

Observation f0beebc1-9999-4b7e-97b3-738c0a256f0c · inbound

Multimodal AI for Gastrointestinal Diagnostics: Tackling VQA in MEDVQA-GI 2025 cites this paper.

Multimodal AI for Gastrointestinal Diagnostics: Tackling VQA in MEDVQA-GI 2025 Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:57:51.206063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:57:51.206063Z digest=sha256:92bd278231d41f76a2eea2f7de6838d7773d5ab0be082387de773c79b4426ac0

Observation 16a6d077-dde9-4722-9d55-5f74cf186cc2 · inbound

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology cites this paper.

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:56:21.606430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-12T03:55:55.359488Z digest=sha256:f354191b48f542e8f9d3169e59e4490301372556b48a049edfb83388f3d79619

Observation e4faedd8-0162-413c-94cc-6b6cde36c72b · inbound

Measuring and Improving Complex-Atomic Answer Consistency in Endoscopic VQA cites this paper.

Measuring and Improving Complex-Atomic Answer Consistency in Endoscopic VQA Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T16:57:30.636879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:57:30.636879Z digest=sha256:dc7a134c76f15fae8fe94846782eff2e00ed300588fd1a31e129fa4697dcd34b