Pith. sign in

Paper Citation Record · LEDGER

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG

As of 12 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 1 inbound Pith citation observation for arXiv:2508.06496.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.06496 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:50:54.486096Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T06:30:16.612345Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact1
  • verified fuzzy38
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 14e827d4-3a7f-4295-8a7b-82a5f0ff9681 · outbound

This paper cites The inception team at vqa-med 2020: Pre- trained vgg with data augmentation for medical vqa and vqg.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG The inception team at vqa-med 2020: Pre- trained vgg with data augmentation for medical vqa and vqg

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.068303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.306983Z digest=sha256:1dee314949c8cb90cb85c41db80e9175ae7c1ffe25205618d5f9ccb12feb418c

Observation eb557038-88f4-4a43-b0ee-c850e2ccf7de · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Flamingo: a visual language model for few-shot learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.312689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.312689Z digest=sha256:0e8506ec6674c6bccd7263cf00a2eea2d3f6c9a990f36abf286af71fe7f2f985

Observation 8602554d-77e0-436c-94d1-8e251e02e49a · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Bottom-up and top-down attention for image captioning and visual question answering

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.051579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.317920Z digest=sha256:8d8c57577f1ae3aa969f85fb1d8b2f1ee668149f732ececf53e25552afde5098

Observation 80d00b1a-e995-4d4e-b5be-d985abfe1037 · outbound

This paper cites Vqa: Visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Vqa: Visual question answering

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.040451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.321588Z digest=sha256:df2f2efa6e2621b0f5b3e0820d64a6f005fbc855b4b7b91ffc075886bbb69f3c

Observation 6a4c2e4a-1e8d-441f-a620-affcf1635601 · outbound

This paper cites A-okvqa: A dataset for adap- tive open knowledge visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG A-okvqa: A dataset for adap- tive open knowledge visual question answering

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.029805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.325609Z digest=sha256:eeed88c85c382b819a9fe270bd46e1df6f203b613d4f14bd5f386569e4e4aedd

Observation e792d4f3-7a7e-4d15-96f2-df33914590e8 · outbound

This paper cites Ocr-vqa: A dataset for visual question answering with text in the wild.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Ocr-vqa: A dataset for visual question answering with text in the wild

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.019110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.329110Z digest=sha256:a1aa2023f07739bd7e3484e3972710ed81ee721c8a8c0599900be97fb0462572

Observation cb724ada-2609-4830-b5cc-349b420eec7c · outbound

This paper cites Okvqa: A dataset for open knowledge visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Okvqa: A dataset for open knowledge visual question answering

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.008024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.334117Z digest=sha256:e7e4d90a0f8e3babc397ff65a0f2f63c25e1c8ee1a03bc00943a8d4d4a3e58d5

Observation d71c85ab-0fa7-44fa-a174-c2d049d6783a · outbound

This paper cites Veagle: Advancements in multimodal rep- resentation learning.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Veagle: Advancements in multimodal rep- resentation learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.997357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.337541Z digest=sha256:1dc04830820832e95d0cba10853c4490a2a46f9cee319e06d4e5598ef2301af4

Observation 4f6b2003-8a87-4d07-904a-7beaaa28f0c6 · outbound

This paper cites Visualgpt: Data-efficient adaptation of pretrained language models for image captioning.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Visualgpt: Data-efficient adaptation of pretrained language models for image captioning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.986231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.341065Z digest=sha256:86c0d5bbb46da32c80e4c012be523924f6a28f27260c48de3279a96fe81ab2e3

Observation e7ccec12-a2b3-4400-957f-a558af30c738 · outbound

This paper cites Unsupervised Learning of KB Queries in Task-Oriented Dialogs.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Unsupervised Learning of KB Queries in Task-Oriented Dialogs

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T15:50:54.599287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.344783Z digest=sha256:5d5404df1492bbb6167a234ff32ca021f5b6498e078a554095849ee5646aa6b2

Observation b4f33e8a-ebaa-41b2-958b-5da2bf0b5877 · outbound

This paper cites Multi-level Chaotic Maps for 3D Textured Model Encryption.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Multi-level Chaotic Maps for 3D Textured Model Encryption

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:50:54.578734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.348419Z digest=sha256:7e213b4d8e0a9524f22a11dadb1f480e22068f3e202bf92e0b15439c6ea30c77

Observation 8cae71b0-d663-4c17-b572-2f014ef36f99 · outbound

This paper cites From local to global: A graph rag approach to query-focused sum- marization, 2025.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG From local to global: A graph rag approach to query-focused sum- marization, 2025

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.975927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.352614Z digest=sha256:ca5b36404bc960dabe7f1e4366531aa53ca00d766af307c9589838d6121542c4

Observation 39130238-8533-4adb-8d0d-1a2a8dd27a6d · outbound

This paper cites an unresolved cited work.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:50:54.966356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.356485Z digest=sha256:8cc5e0dc111a99aa3887e92434a46a42f226cf4be147210b0c0300a7f0e996e6

Observation dbbee4d6-7915-4b50-a854-cd9903bea7e6 · outbound

This paper cites Prompt learning with knowledge graphs for zero- shot relation extraction.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Prompt learning with knowledge graphs for zero- shot relation extraction

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.955411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.360598Z digest=sha256:aad508276113b2d54f44800347dc0439b6b7ed149890eb8ba4b7d24e2458de1f

Observation da219817-ea59-44c6-9930-56e48ceb66f2 · outbound

This paper cites Imagebind: One embedding space to bind them all.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Imagebind: One embedding space to bind them all

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.945171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.363866Z digest=sha256:6ea1deccf192925553c60923d6acd5a161f82eb9d6c443e24dffffb0adbbfb07

Observation 3cb271ab-e6c5-424f-881e-0af75fd87ede · outbound

This paper cites Girshick.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Girshick

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.935163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.366895Z digest=sha256:c2473c419f39f2e30af1c18cd2d0555984ecbf177aea23ff9682b080a3d849f2

Observation 06590fca-c945-4c04-95dc-75762879fd07 · outbound

This paper cites MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.370611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.370611Z digest=sha256:e45de7d19e4be7b95c9ef8af8a7b2270fb49c18415b78e66d82aca54168146e2

Observation 2612983a-d807-440f-95a2-f987135eebee · outbound

This paper cites Bliva: A simple multimodal llm for better handling of text-rich visual questions.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Bliva: A simple multimodal llm for better handling of text-rich visual questions

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.923922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.374634Z digest=sha256:7796bef15c48e4ccc83791ded4886813882bcc384a846d95607114242e8ec776

Observation b568daf0-229e-4f1d-9114-669fb58e0758 · outbound

This paper cites Interpretable medical image visual question answering via multi-modal relationship graph learning.Med- ical Image Analysis, 97:103279, 2024.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Interpretable medical image visual question answering via multi-modal relationship graph learning.Med- ical Image Analysis, 97:103279, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.914096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.378530Z digest=sha256:ea4e1b273eeea4d9a9bd6048f9a88a65ca1a17249488e1676fff38d58bb9da3c

Observation cc70cdb8-06c0-4033-b745-1084bcd88cc1 · outbound

This paper cites Retrieval-augmented gener- ation for knowledge-intensive nlp tasks.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Retrieval-augmented gener- ation for knowledge-intensive nlp tasks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.903043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.382745Z digest=sha256:e789ebba4f9f844766b2ba3544f437a0cc8523579e53f4e45696518fba88b0f8

Observation 8f399d90-b07c-4872-86ac-fd52636dae0e · outbound

This paper cites Llava-med: Training a large language- and-vision assistant for biomedicine in one day.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Llava-med: Training a large language- and-vision assistant for biomedicine in one day

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.891926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.386519Z digest=sha256:43be795f22533182dcea17979c64417aed3ce2ce41e3512caec8e402236440b8

Observation ff4757b2-8f03-42d8-80f1-8cd705676a68 · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.881829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.390443Z digest=sha256:784e610d33f7e2e86c8d7e156c8ddf15a08cbce4d6d24d4203065a8ea6e61dd8

Observation 11c4b357-ce72-453d-8fb9-f29579305448 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.870066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.394082Z digest=sha256:e046b93cba0ec04e21a4a38439d08cb40651b2d098f8bb0db1bef30ec5093010

Observation 3f73eeb8-3a04-4743-ac04-452e461aecbd · outbound

This paper cites Masked vision and language pre-training with uni- modal and multimodal contrastive losses for medical visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Masked vision and language pre-training with uni- modal and multimodal contrastive losses for medical visual question answering

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.858316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.399164Z digest=sha256:57d7b456fcab33967dbb76e58e3a90d604f1a1cb7b5180686a1ed4513b697056

Observation 3cc3c98d-0afc-4647-826f-f6afb737aacb · outbound

This paper cites Lawrence Zitnick.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Lawrence Zitnick

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.846804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.402739Z digest=sha256:18b09a6bfdf0090bc914c4515198cafcf116260179776808e4900c1277a25f87

Observation 004cbde5-821d-475d-9258-b497ccce5156 · outbound

This paper cites Visual instruction tuning.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Visual instruction tuning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.834938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.406219Z digest=sha256:94388f9bc8044b5f844d5a9d150fcfa6090dfef7595cdb6d8f597efb76eeb7d2

Observation 02c580d1-817a-4479-8799-9b05251950e6 · outbound

This paper cites Can we Trust Chatbots for now? Accuracy, reproducibility, traceability; a Case Study on Leonardo da Vinci's Contribution to Astronomy.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Can we Trust Chatbots for now? Accuracy, reproducibility, traceability; a Case Study on Leonardo da Vinci's Contribution to Astronomy

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T15:50:54.548150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.409579Z digest=sha256:382871270e408acffba455cc2e84412691f4a12743dce3a6fb2f71df8d34b008

Observation 79a35448-1e64-4128-be24-e8998fcdfbd9 · outbound

This paper cites Foundation models for generalist medi- cal artificial intelligence.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Foundation models for generalist medi- cal artificial intelligence

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.822316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.413049Z digest=sha256:1b371424464e1d819d6a53a9c7fa11f43721fe99725b59cd8374803b711f8c22

Observation df2e7e2e-9c2a-46e5-bc64-df92b5195c67 · outbound

This paper cites Med-flamingo: a multimodal medical few-shot learner.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Med-flamingo: a multimodal medical few-shot learner

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.809268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.417280Z digest=sha256:fc8289600c1fe9b6d4f5f1e0df2222b139848a8513f0959debaa671430a5e0eb

Observation 5e96866b-a842-4045-a505-371176785f11 · outbound

This paper cites K-pathvqa: Knowledge-aware multimodal rep- resentation for pathology visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG K-pathvqa: Knowledge-aware multimodal rep- resentation for pathology visual question answering

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.797037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.420725Z digest=sha256:1d79cc05348341637ff41afe35fcebe89f2e46f0f2c0d9b47d8af3c84d99c24d

Observation 44b3a3d5-27fd-4a0a-bc36-d7e7d952c21d · outbound

This paper cites Overcoming data limi- tation in medical visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Overcoming data limi- tation in medical visual question answering

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.785426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.423859Z digest=sha256:4649a8e4081613e58d61c9c57b0e29dabc0b43784944c42a94787aac990afcf6

Observation c237a9e1-a7c8-4057-b873-569db0cf5b05 · outbound

This paper cites St-vqa: Visual question answering with a focus on scene texts.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG St-vqa: Visual question answering with a focus on scene texts

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.773459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.427548Z digest=sha256:9faf70984065fe844e8fb0e27c5bb835f30a82d5d3fc9e8ff4a633cc5d934500

Observation 2c23d8ba-83b4-42c1-8a05-38b2adcde634 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Learning transferable visual models from natural language supervi- sion

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.430867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.430867Z digest=sha256:f981cffb277f8ad844eced07dd5aca6457665389289b7cc5b3ffc9b53c1d1514

Observation a36ae1f0-dd3f-49b4-87ba-1cfe568394dd · outbound

This paper cites Docvqa: A dataset for document visual question answer- ing.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Docvqa: A dataset for document visual question answer- ing

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.754225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.434466Z digest=sha256:83f013274e5b81aa4235f1318d3137361e7c12964df43bca38418e1b0061b67c

Observation bf596faf-5546-48a2-b1bf-d2465cda53ee · outbound

This paper cites Maivar-t: Multimodal audio-image and video action recognizer using transformers.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Maivar-t: Multimodal audio-image and video action recognizer using transformers

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.742797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.438125Z digest=sha256:c63ff300ea2b199539df35bc516ec2e3893be4493faef6802b66a6f6069d4a0e

Observation 549747e7-e4be-4ec6-9b43-30f5f64eeb55 · outbound

This paper cites Sharma, A.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Sharma, A

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.733101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.441123Z digest=sha256:a57cd4c3a48aafe17f497d22902de1632334caf7326111ddbf82cf1cec61f292

Observation a02c14d9-bb8f-41b0-9b45-e0a1cba96ad5 · outbound

This paper cites Multimodal few-shot learning with frozen language models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Multimodal few-shot learning with frozen language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.722054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.444663Z digest=sha256:61f2f2d090ecc62363ac24526b9eb6f5954af7b0a4f8166f9185daffbf9a2338

Observation 347f422c-576d-4397-a796-06d37ee99ffa · outbound

This paper cites Attention is all you need.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Attention is all you need

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.711077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.447854Z digest=sha256:7060616be69aca9d128a59e5332fc67b59c12db27cb1b6f0d963126e96945bd5

Observation a1a75ff0-74c1-46f4-ba75-805f6df4ab5a · outbound

This paper cites Medical graph rag: Towards safe medical large language model via graph retrieval-augmented generation.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Medical graph rag: Towards safe medical large language model via graph retrieval-augmented generation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.699090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.451614Z digest=sha256:782b5cb458722563c2b6c54e0f775d8fbafdf65d63aa47c16f5b629d47a9821c

Observation 92cc7a48-8678-4f67-8c0b-8788de69aa39 · outbound

This paper cites Mmed-rag: Versatile multimodal rag system for medical vi- sion language models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Mmed-rag: Versatile multimodal rag system for medical vi- sion language models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.687549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.454992Z digest=sha256:07a9c12a661d6fad0092c3aaa3c61b1ffe5e240703fa36507caaf9b96fecfd71

Observation 6b28f0b1-c14a-4bd1-af65-fe1ac6c9db2d · outbound

This paper cites Rule: Reliable multimodal rag for factuality in medical vision language models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Rule: Reliable multimodal rag for factuality in medical vision language models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.676197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.458130Z digest=sha256:5444b65508b457defda17ef92fdb78403437bf177d4a2755717bdf1d4dcd541d

Observation 60dfb2fe-7e4f-4ab4-b54e-b52444cd22fe · outbound

This paper cites Ramm: Retrieval-augmented biomedical visual question answering with multi-modal pre-training.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Ramm: Retrieval-augmented biomedical visual question answering with multi-modal pre-training

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.664309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.461891Z digest=sha256:5ec2daf214042d76269b687c6cc3abcced9e0e79b1a8f76672e6739154d4543d

Observation 58fe69e3-9993-41a6-b5dd-69ff91bacd1d · outbound

This paper cites A generalist vision–language foundation model for diverse biomedical tasks.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG A generalist vision–language foundation model for diverse biomedical tasks

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.650767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.465268Z digest=sha256:3dd451dd9a20ba63d71c3c7a534a29dd9c123c01ede18075c12c71abd5605cd5

Observation e7c6b56e-e8e5-425d-836e-adad4a8b1ab4 · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.469168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.469168Z digest=sha256:f2f18391847bce549c3f090e8e6daef5685a65aa257c5de03c603bfb07772c53

Observation e7eee508-72dc-4667-9893-68d86374efbf · outbound

This paper cites Multimodal representation learning by alternating uni- modal adaptation.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Multimodal representation learning by alternating uni- modal adaptation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.638093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.474110Z digest=sha256:f9985e4fdcf7bb73a71076fd2cba64ec40f72c956c42483973b0d9fd2756f024

Observation 409a3c91-1642-4f8a-a0c8-d3c77c495ad1 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.477871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.477871Z digest=sha256:5711244a50c5fef14dfd48d935986d210aa0d76e892f928031569bd9f061829f

Observation f91b943d-e1e0-46c1-89c2-96937c131f6d · outbound

This paper cites an unresolved cited work.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:50:54.625452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.482486Z digest=sha256:17eade6cd4e1bedf859c1fc60e7aea6c8ed4015f3b4cf21cdfd9ac42ad4e015f

Observation 93c184ca-0c8a-4628-bf3f-73fa882a2742 · outbound

This paper cites Crossclr: Cross-modal contrastive learning for multi-modal video representations.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Crossclr: Cross-modal contrastive learning for multi-modal video representations

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.612976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T15:50:54.486096Z digest=sha256:47cc6883735eac34517bde35a30315ac967cbcc68d27a5e3f2847ec06c712373

Pith citing papers

Observation f7428869-4c5b-484c-9168-979a7864cfe6 · inbound

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy cites this paper.

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG

Reference 164

Resolution
unresolved
no resolver link, observed 2026-07-14T06:30:16.612345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:30:16.612345Z digest=sha256:451b02edffc9809fd95a93efe59539d22763e11f966f4eb9665708145d92a8e8