Pith. sign in

Paper Citation Record · LEDGER

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

As of 18 August 2026, this Paper Citation Record lists 85 of 85 outbound references and 11 inbound Pith citation observations for arXiv:2411.14522.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14522 v2

Coverage vector

measured 85 of 85 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:16:27.982410Z

measured 96 of 96 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:52:07.139701Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T01:39:23.953906Z

Reference resolution

85 of 85 outbound references displayed

  • verified exact0
  • verified fuzzy35
  • unresolved49
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f0809c6c-e530-4956-bcdd-165374031977 · outbound

This paper cites GPT-4 Technical Report.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.585908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.585908Z digest=sha256:5dbd7a5e97a8f36ee2a8d5227ce8b4308c46f94b7b8567cf9c07dfc0998d544d

Observation bb68ff5b-f2c4-4c68-8a85-b779a8caedd9 · outbound

This paper cites The claude 3 model family: Opus, sonnet, haiku.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI The claude 3 model family: Opus, sonnet, haiku

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.591221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.591221Z digest=sha256:c944e1cdbf8a349efe4b2561554c222866831bc25f7a6096393611e0b66e48f8

Observation ebd4e356-e6fd-4315-94c0-9df366c4fc05 · outbound

This paper cites Home - Retina Image Bank, 2024.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Home - Retina Image Bank, 2024

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.595870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.595870Z digest=sha256:8ed049d5b5ff6a62b4d3aa9db364e739ab94836e795b1bb3354c854eb7b3506e

Observation 939980e2-769d-4fab-bca9-f9733750eeb7 · outbound

This paper cites OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.600653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.600653Z digest=sha256:f9ee387d3db8d1baa380f88ce4b75132caee906bf30a9c63d5807d3b2d8fd737

Observation 737d2341-da84-4bb5-a81c-6ec16d908136 · outbound

This paper cites Blossom orca v3.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Blossom orca v3

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.605710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.605710Z digest=sha256:1f47615e9fad5122b86056567162d328d3657675437881ee129db16510ea183a

Observation 3c358ec7-ce60-4f89-a2ff-f805515b47bd · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.610430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.610430Z digest=sha256:34cb8c342330652f0fea30b824ec1981d6cee231b26c236cf23411202f1b057f

Observation 5ff89aa8-1982-4eac-a181-7e472d9800e5 · outbound

This paper cites COIG-CQIA: Quality is All You Need for Chinese Instruction Fine-tuning.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI COIG-CQIA: Quality is All You Need for Chinese Instruction Fine-tuning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.615383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.615383Z digest=sha256:45f9ebc18d181a5bead40cc814158a47b5f182279ff4a180d3d9aed5b5352b20

Observation 7f1b43fe-ab6d-45a7-a574-e3609967799f · outbound

This paper cites Vqa-med: Overview of the medical visual question answering task at imageclef 2019.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Vqa-med: Overview of the medical visual question answering task at imageclef 2019

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.621116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.621116Z digest=sha256:8ef6cc052da2fb46ef3b1dfa0be10716f7c36027eea63c6bef253c0fc64d1600

Observation 8dc39d1f-77dc-4245-8dae-722d79aca42e · outbound

This paper cites Cosmopedia, 2024.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Cosmopedia, 2024

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.626151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.626151Z digest=sha256:50021c49320038ef9e337596dd15c0b9c2b86e9c088fb990e87a7b802320824f

Observation 7abd5be5-6e47-4733-8029-1b3ca22517b0 · outbound

This paper cites Leetcode dataset, 2023.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Leetcode dataset, 2023

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.280400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.631296Z digest=sha256:d48b69958869326454c076515588c4a674b34cb6b58b0057cbfcc4f199083bfa

Observation 0b54dee0-d328-40fe-8383-6bd19b4df47f · outbound

This paper cites An augmented benchmark dataset for geometric question answering through dual parallel text en- coding.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI An augmented benchmark dataset for geometric question answering through dual parallel text en- coding

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.264936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.635487Z digest=sha256:ebadac8958e993296dbc2ba138deec63d4f147974031a90903c5fcfdf4436944

Observation 85d4dbcf-82ad-45f9-88e4-b9f37d25c8c8 · outbound

This paper cites CheXpert Plus: Augmenting a Large Chest X-ray Dataset with Text Radiology Reports, Patient Demographics and Additional Image Formats.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI CheXpert Plus: Augmenting a Large Chest X-ray Dataset with Text Radiology Reports, Patient Demographics and Additional Image Formats

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.639887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.639887Z digest=sha256:01a3f56cd91af5b884f595c91a0fddb29ea2fcc4223c355dce2cbf37fd5f120b

Observation d842d6a9-bf21-4984-a324-0eb0a2b022ac · outbound

This paper cites ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.644551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.644551Z digest=sha256:2e9ae2db028bdf8c2ba2593d130a3238e4e229acf1c9778b255f01d370c1a558

Observation 00198be9-9255-40dc-a781-d94e84bcdfb4 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.649840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.649840Z digest=sha256:509d425e5a506d6c70637bdcb2c48dccc7a3fec17abf1f6d6faf51681280aa48

Observation 6452a82c-112c-497c-8344-a58c7e70e083 · outbound

This paper cites ShareGPT4V: Improving Large Multi-Modal Models with Better Captions.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.654549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.654549Z digest=sha256:facfaf93e7c28892f99005b95b2556acfcd19c9038efa6123fd9d5db1800467c

Observation 9992be07-b41d-4250-8c0d-7bb53dc2a890 · outbound

This paper cites GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.659001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.659001Z digest=sha256:77291e11f445017e364330c126a9c8550167f57dfb043909affa1f54fa23a31f

Observation 67103c64-7a6c-4c77-a765-e4fa706aa943 · outbound

This paper cites InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.663595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.663595Z digest=sha256:fe0c3c3e2dbc71c73255528a275899af8a39f532896ee51b4fe872943ad9b221

Observation d99355f5-d6e6-4409-a946-df250c1b60a2 · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.668566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.668566Z digest=sha256:ce439b189da64f45467c2f6998db30631a24cb26275542a591df01e07b30163e

Observation 1dcfadc6-fc6b-452b-bc74-a8276bdc61db · outbound

This paper cites Xtuner: A toolkit for efficiently fine-tuning llm.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Xtuner: A toolkit for efficiently fine-tuning llm

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.249357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.673094Z digest=sha256:e3f50a5fa033e24de73ecc8c17f27d3d1d3076dee2a0ed480a4979625f2a54a8

Observation e77aa06e-8cda-4729-9e83-b4e5ea7a6843 · outbound

This paper cites Instructblip: Towards general- purpose vision-language models with instruction tuning.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Instructblip: Towards general- purpose vision-language models with instruction tuning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.234179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.677471Z digest=sha256:5c9745009fa4199444d218f9cb10d57e32cb67561c9f4a5d0ce7f7520019b58f

Observation ad70fd0b-a3f1-4d5b-9d30-30cc3d69bbca · outbound

This paper cites Preparing a collection of radiology examinations for distribution and re- trieval.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Preparing a collection of radiology examinations for distribution and re- trieval

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.217940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.682661Z digest=sha256:8d525a7ad57ae79bb5668d79eda2837a272649ad17b5bf222b460da9c7d00922

Observation bc1e1e32-9ddb-49b9-9fdc-a15944844fbd · outbound

This paper cites Cogview: Mastering text-to-image generation via transformers.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Cogview: Mastering text-to-image generation via transformers

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.200954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.687092Z digest=sha256:dcd140aa047c9d589cf33931b5eb416a662c9211c183ed51bfd134bbea6fae5a

Observation 6acb6d61-6d11-4d9a-a197-33b0bf516e35 · outbound

This paper cites Enhancing chat language models by scaling high- quality instructional conversations, 2023.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Enhancing chat language models by scaling high- quality instructional conversations, 2023

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.184316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.691663Z digest=sha256:6d9129e901c4defbf715fc8c21cd0534eec5c7d2dcf6573d5fadf6272c6f0d63

Observation 3d36b248-f3b7-431d-9b55-156ba62e6399 · outbound

This paper cites InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.695949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.695949Z digest=sha256:d829db0ef6258b901b9b2519787c3fd113bbf882a6a80b0cca458637f7ca056b

Observation 9535ac0b-a992-421a-9ea8-04cfdbcbfd07 · outbound

This paper cites PaLM-E: An Embodied Multimodal Language Model.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI PaLM-E: An Embodied Multimodal Language Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.700784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.700784Z digest=sha256:07344b08720db9d52a4541d27ba531e533a4432b9084d5a4bba3882182ab4d7b

Observation 948e59e8-2f08-439e-8009-5b0feccbab3e · outbound

This paper cites Vlmevalkit: An open- source toolkit for evaluating large multi-modality models.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Vlmevalkit: An open- source toolkit for evaluating large multi-modality models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.168528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.705216Z digest=sha256:3991e6cca15262d4cb53f49a273fb66e020edcd2f3ca53db54b2cccb01187dd6

Observation 8a090375-375b-410d-93ad-3a419b62b743 · outbound

This paper cites Lima: Less is more for alignment, 2023.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Lima: Less is more for alignment, 2023

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.153816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.710153Z digest=sha256:6942f71160e3af7c2bf5bd99596ca0bf9180f8d1bfcd32de4ca59f129a992a42

Observation f82a13f6-f47a-4bd2-b682-ac2ec7b82f6f · outbound

This paper cites LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.714258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.714258Z digest=sha256:48d699aafb7578a0afbeb70e230c7158bbb8e4cbe27fc5dd96b8eb13f3c7e36f

Observation 2eb5b23a-ea07-4bb6-9655-5760089ae459 · outbound

This paper cites GSCo: Towards Generalizable AI in Medicine via Generalist-Specialist Collaboration.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI GSCo: Towards Generalizable AI in Medicine via Generalist-Specialist Collaboration

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.719300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.719300Z digest=sha256:ca9f10ebb7c3eec72bc19be3847c82ff84c787663c05f1c2cb1a7bbe1ff61e53

Observation d8bc2560-4eca-4777-828a-8138c7af8401 · outbound

This paper cites PathVQA: 30000+ Questions for Medical Visual Question Answering.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI PathVQA: 30000+ Questions for Medical Visual Question Answering

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.724943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.724943Z digest=sha256:7ac61f1d570142714d5a0dc27deef0603a4ddec75227db2d5d330eabf2e733a0

Observation d05aa105-d6ff-4b0d-b562-f2a77d7db6d7 · outbound

This paper cites Medical-diff-vqa: A large- scale medical dataset for difference visual question answer- ing on chest x-ray images, 2023.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Medical-diff-vqa: A large- scale medical dataset for difference visual question answer- ing on chest x-ray images, 2023

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.139048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.729955Z digest=sha256:adb92f8a6e76bf00ec4477bb2a5ea3622159f29fc6b4057d8f4d3d5ddcc13c5e

Observation ba74ef87-2fdb-4792-879b-d01e3d2f952d · outbound

This paper cites Omnimedvqa: A new large-scale comprehensive evaluation benchmark for medical lvlm.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Omnimedvqa: A new large-scale comprehensive evaluation benchmark for medical lvlm

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.123637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.735065Z digest=sha256:5bd54b2eadf011338db14142749dddb41dfd6fcf0d5ae67d3d64ac9284f5e948

Observation ff3c5f8b-8fd7-4e65-9ae8-75f7947ad433 · outbound

This paper cites Quilt-1m: One million image-text pairs for histopathology.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Quilt-1m: One million image-text pairs for histopathology

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.739731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.739731Z digest=sha256:b3ac1a290e4306dfa26b07f40eb60b392546825cd339fae4fd830dc66c674e9e

Observation f30d8a76-7125-4b8e-ada2-812d6923897e · outbound

This paper cites Quilt-1m: One million image-text pairs for histopathology.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Quilt-1m: One million image-text pairs for histopathology

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.098556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.744344Z digest=sha256:df4bcdf7cb12907799d7cc599a428fd0b683ef264619b488e6ebe2b7e296b678

Observation ace6742a-394a-42b2-bfdb-b3ea4158bf6e · outbound

This paper cites Mimic-cxr, a de- identified publicly available database of chest radiographs with free-text reports.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Mimic-cxr, a de- identified publicly available database of chest radiographs with free-text reports

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.083428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.748724Z digest=sha256:b106cca23b57b31dc715d35afb115b6af5616d33cd531ae1db9811f5084baaab

Observation 846fe4a7-c08e-457f-b314-e69ff2ca614c · outbound

This paper cites Dvqa: Understanding data visualizations via ques- tion answering.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Dvqa: Understanding data visualizations via ques- tion answering

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.753361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.753361Z digest=sha256:ba80fa473fd6479e297e0f3069e101b1d2e8b4265aa28d9ecfea30b9830cc396

Observation a9550fcf-9ead-481f-afd1-ec4c9bd02839 · outbound

This paper cites A diagram is worth a dozen images.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI A diagram is worth a dozen images

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.758298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.758298Z digest=sha256:a2e30b54bd6afdacad7d21f4b2a740ee3c17f6f734f535733c74c55061a4a927

Observation 422491cf-ac34-4dec-a00f-76c566fb2159 · outbound

This paper cites Ocr-free document understanding transformer.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Ocr-free document understanding transformer

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.050300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.762995Z digest=sha256:faf831cc379259db54b1be19e430c64b1e07b2a5080c37fe3d20aba2611bb7d2

Observation 084ee0bd-36b7-4457-ac56-28d29ee3c24d · outbound

This paper cites A dataset of clinically generated visual questions and answers about radiology images.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI A dataset of clinically generated visual questions and answers about radiology images

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.034671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.767451Z digest=sha256:74f9a084675fe5a5e58364ee5c0579e9ae7607b80ebfbf66d73e5109c06a17fe

Observation 2ae0acb5-eee4-4e2c-a796-97a2c323fb55 · outbound

This paper cites Llava-med: Training a large language- and-vision assistant for biomedicine in one day.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Llava-med: Training a large language- and-vision assistant for biomedicine in one day

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:29.020269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.771951Z digest=sha256:7badad4e6a78a516236dc76d0558678c9b9bca3488ba8cd359c6934405882b60

Observation 66cb7b1c-534e-4578-ae99-22a6b0fec26b · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.775959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.775959Z digest=sha256:aa927cc87d8524ec639384e6844ff5c60b4a8e05ebcd76b62c472931dc389e2f

Observation f0ead3e3-c8a0-4351-92a9-c8ec737d7e05 · outbound

This paper cites Seeing and understanding: Bridging vision with chemical knowledge via chemvlm.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Seeing and understanding: Bridging vision with chemical knowledge via chemvlm

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.780679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.780679Z digest=sha256:f8e55a07dd19cdaae195200836bcbf1d18f347f6759759843027cc16d16d41aa

Observation 7dd5f11b-00a0-4169-8a00-84f9a18958a0 · outbound

This paper cites Pmc-clip: Con- trastive language-image pre-training using biomedical docu- ments.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Pmc-clip: Con- trastive language-image pre-training using biomedical docu- ments

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.995216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.786209Z digest=sha256:61c666793605c44c9cc521426ac98a5a9feef85b4be17a46b5765650412de160

Observation 614c8785-f960-42cf-a6f7-d6223ac5951b · outbound

This paper cites Slake: A semantically-labeled knowledge- enhanced dataset for medical visual question answering.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Slake: A semantically-labeled knowledge- enhanced dataset for medical visual question answering

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.980089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.790508Z digest=sha256:ffe3a618f7692f2133aa56ddf7315a7eb6c1614ed99eb96bd97cb27bd39e623f

Observation c8bc8183-cf94-451e-9376-1e6c221afa9f · outbound

This paper cites Visual instruction tuning, 2023.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Visual instruction tuning, 2023

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.965005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.795480Z digest=sha256:09e7bc217ecdffa295fc0d388ae54473cc76fda9c5e24570cb79b4289bb61a41

Observation c280c82f-eb05-476b-951e-368da4b01900 · outbound

This paper cites Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.800062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.800062Z digest=sha256:a9e7348eacd604aa4e63cf01e7366d9cb3fe6ae69af01d90b12d89ceb070eaf8

Observation bb30a723-cae1-4d6e-af03-8c5af8a9c59c · outbound

This paper cites LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.804972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.804972Z digest=sha256:a8e8aacc38e6764774c9707375651fa122376bb802760e02f27e3bb771b2a132

Observation a9067b3d-23bc-4783-85ae-ad813262062f · outbound

This paper cites Qilin-Med-VL: Towards Chinese Large Vision-Language Model for General Healthcare.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Qilin-Med-VL: Towards Chinese Large Vision-Language Model for General Healthcare

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.809953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.809953Z digest=sha256:ea77775c4f4d69144d66cfb8eb423248e5b5243942da1aa437cba6b1817a8911

Observation d5db7375-b621-4c75-ad9e-f5eec85839a8 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.814735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.814735Z digest=sha256:6d2fb2a1b01697eb327baae5738fd50639e7d38ab8fca3b41a40011773f35f87

Observation 01181550-8cd0-4fb9-9e27-495406497941 · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.819817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.819817Z digest=sha256:1dd17088a3ef36ad25c398a2c02253e6d7ea572f9d2ae4a132b10a829655a1f1

Observation 0494fc53-150b-42ec-85e3-4b4ae42c5b68 · outbound

This paper cites Docvqa: A dataset for vqa on document images.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Docvqa: A dataset for vqa on document images

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.940226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.824469Z digest=sha256:20e591cf3a1e82b51d8101c3293183e70adff6b9901dd5ab8d415364303bd8aa

Observation cc0f8001-9fab-47d0-af8a-0862f495cf19 · outbound

This paper cites Orca-math: Unlocking the potential of slms in grade school math, 2024.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Orca-math: Unlocking the potential of slms in grade school math, 2024

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.924769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.828628Z digest=sha256:8740b488978f6381a97b0db8947755209dc492c7eff6cdcf8aacfd91cbae2d64

Observation 00368572-52cc-40f1-8a26-fa8e34513d9f · outbound

This paper cites Med-flamingo: a multimodal medical few-shot learner.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Med-flamingo: a multimodal medical few-shot learner

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.908155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.832932Z digest=sha256:338dbf7f9b1abbb9a7c45d9a12062d78fe3dbb7f82743e30755857a9a951341c

Observation bc8763a9-3187-47bd-a685-f42c07709cc5 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Learning transferable visual models from natural language supervi- sion

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.837170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.837170Z digest=sha256:d4f8fda7896754775e14b0b89e4d9c01abb2422fc661ae10eaebed5ea2ba7992

Observation 01d7b33c-d69c-41b7-aa1f-5eef4d7074ad · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.841408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.841408Z digest=sha256:b258a7289101bcc58e208d7d505473f6e284b1240039fa804505499a5b564f0f

Observation b819e4e5-efab-4e82-b283-9c263b4af9c7 · outbound

This paper cites Seco de Herrera, et al.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Seco de Herrera, et al

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.881963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.845726Z digest=sha256:d670cc8ab9f8c8b98b7461e0781c47830e4387726b4babaece4e6a2de154e4e1

Observation bac7800f-9f96-41eb-ba65-82294ed302ee · outbound

This paper cites Capabilities of Gemini Models in Medicine.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Capabilities of Gemini Models in Medicine

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.850285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.850285Z digest=sha256:dc7f919ef6762486961de81a9c3fb7cc258502020b85794ce6ca9cc72c6a3ddb

Observation bb1443b6-b347-4b3d-be41-72aae5223b28 · outbound

This paper cites Large language models encode clinical knowledge.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Large language models encode clinical knowledge

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.865635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.855141Z digest=sha256:a681d1cc3e899158ed88363790f710ec86cdef4b49aa15d23ce09b19d25336a7

Observation 1b0907b5-9f76-4b44-b231-bc57dcbc8375 · outbound

This paper cites MedICaT: A Dataset of Medical Images, Captions, and Textual References.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI MedICaT: A Dataset of Medical Images, Captions, and Textual References

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.859906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.859906Z digest=sha256:fd71364c83f39c37f2b433c79a396d597fcbb1900a1c21ac8b9d90168ff25505

Observation adfe8103-d840-4aa2-b589-e46d6b6d813b · outbound

This paper cites Hashimoto.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Hashimoto

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.850185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.864071Z digest=sha256:8d14056bf56d54cc7ca2d9dee404aad6101ce392ddd54bef774b824656bf856d

Observation 8bc57a48-baaf-4bbc-a4f5-019338a9bb5d · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Gemini: A Family of Highly Capable Multimodal Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.869235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.869235Z digest=sha256:5adfeb6699dcb37d275c3264ba37f9ac64fedf786b442747013f8c7e7fe50a43

Observation 4c1ded91-8ed5-40c9-886a-ab80b647eba1 · outbound

This paper cites Internlm: A multilingual language model with progressively enhanced capabilities, 2023.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Internlm: A multilingual language model with progressively enhanced capabilities, 2023

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.834619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.874061Z digest=sha256:51e9c5a7aa9b4c617795f339bfa050374e50deec180005b2873d489de3098cd8

Observation 012f75d3-df3f-428e-bc0f-65b3b58fbe29 · outbound

This paper cites Openhermes 2.5: An open dataset of synthetic data for generalist llm assistants, 2023.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Openhermes 2.5: An open dataset of synthetic data for generalist llm assistants, 2023

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.819417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.878676Z digest=sha256:432f19c06b345320be727dc34c43dceeb2d407d75c255c79b743f92902fb1dc1

Observation 2d717345-e6c4-4d7c-b475-1959ea84eb6b · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI LLaMA: Open and Efficient Foundation Language Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.883868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.883868Z digest=sha256:a470aab40bf2fb2ea2e95afb37800787cbce79442e1cb2e065ca199727e80fc7

Observation 975b2ba1-b3bd-432c-902f-2bd51c9a0077 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.888474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.888474Z digest=sha256:4b92ff6cbdb18df9868bcd789af4b5546c9b53a05703818c19d21bc522d03792

Observation 91903ecb-3c1c-4234-917f-356fd40d1a59 · outbound

This paper cites Towards gen- eralist biomedical ai.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Towards gen- eralist biomedical ai

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.893658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.893658Z digest=sha256:6f9417887fc30bcdb85e5760b20e126a810e054ebe3fecefcd6c9f62cd66dca7

Observation 9223ef79-fa0b-47a8-abdf-f2919e1588e2 · outbound

This paper cites PMC-LLaMA: Towards Building Open-source Language Models for Medicine.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI PMC-LLaMA: Towards Building Open-source Language Models for Medicine

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.898170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.898170Z digest=sha256:b2ba0a5a4283a8fd9ad8b504a136d08c8272b6a56d1723d7a4a78ffa063edf54

Observation fcb4f0f9-ac6e-44f9-afbb-d8f71ccee68e · outbound

This paper cites Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.903306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.903306Z digest=sha256:2f25d0894cafa32ebbfcb65c54f4ec690ad59c76c2c28a0ff03a6b4adf6ffd1e

Observation 95d0985f-facc-4500-99cf-a5463af0032f · outbound

This paper cites MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.908322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.908322Z digest=sha256:39feb3c129e8c10213947cbff0664e3e38e7b1ebdaaeb0e39e4d727d75f15043

Observation e2fd5e40-6ffa-47b7-a945-2e14b621c775 · outbound

This paper cites LLaVA-UHD: an LMM Perceiving Any Aspect Ratio and High-Resolution Images.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI LLaVA-UHD: an LMM Perceiving Any Aspect Ratio and High-Resolution Images

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.913904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.913904Z digest=sha256:e27e98b18a842e6f4eeabe83f62d57920dbeb68f39f4f5fa2dc864f0a36fe13b

Observation 9bea8762-ce0a-467e-8158-f1c2cf4504c6 · outbound

This paper cites Firefly( 流 萤): 中 文 对 话 式 大 语 言 模 型.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Firefly( 流 萤): 中 文 对 话 式 大 语 言 模 型

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.793733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.918585Z digest=sha256:c6078eeb342c263faadef03d13f3dd047ba491e22a08d1d25473991dabc4e0c2

Observation 53169a36-9c93-4476-a02a-e31960efaaa1 · outbound

This paper cites SA-Med2D-20M Dataset: Segment Anything in 2D Medical Imaging with 20 Million masks.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI SA-Med2D-20M Dataset: Segment Anything in 2D Medical Imaging with 20 Million masks

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.923422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.923422Z digest=sha256:a59de37ee1a79150379e1ccac77b50f6ba5b0b2071fe05231457211894efc2d2

Observation 589337fe-1275-4edb-a5c6-ff4f47647968 · outbound

This paper cites mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.928118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.928118Z digest=sha256:99fb29d8aeb6fd617ebdcbea20cc2f9305839fde78a49f53a62d1662bd986636

Observation 0b35edca-080d-4e12-890b-22ac3a091234 · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Yi: Open Foundation Models by 01.AI

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.932740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.932740Z digest=sha256:a5a0c05e0575d962bdc49ba91406f747c6eeb8472378a59985110b2172fb84ab

Observation 61836b9d-123c-4ef2-931e-d527a78ede2d · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.777508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.937065Z digest=sha256:6f785f863371217b2d36042826705817a8cef0e25eb2aa162fb7ddbf66859e1d

Observation e910100c-211f-4468-b854-0a80a34ee52c · outbound

This paper cites A generalist vision–language foundation model for diverse biomedical tasks.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI A generalist vision–language foundation model for diverse biomedical tasks

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.941299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.941299Z digest=sha256:0e00cf68490e8680a315b1910654951ba9d1bcd847f34bbd7049a8d60fd0efa1

Observation f476cd61-7b25-46f1-b91d-13dc0b7abaee · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.945529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.945529Z digest=sha256:13fe58094627bc443b862563fcada9879d88ac18ec79eb52c8ba574eb48445cf

Observation f34ccd04-4706-454a-9b63-2557f7b4aa6e · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 78

Resolution
malformed identifier
no resolver link, observed 2026-08-12T15:16:27.949842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.949842Z digest=sha256:260ef2eb664d12729e73496a3657665ddaf138f76a78f4191aaa5f4f3532347a

Observation 10508207-032f-4067-b8ca-c9d628aa9eaa · outbound

This paper cites The first step involves an initial quality screening using a multimodal large language model (such as GPT-4o) to automatically assess the quality of the pairs.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI The first step involves an initial quality screening using a multimodal large language model (such as GPT-4o) to automatically assess the quality of the pairs

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.752077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.954491Z digest=sha256:6f0041671b17465c9733b96aa136706f9761e4b12b7de46bfe0531125728627e

Observation b2f6f970-88dc-4bb1-bcf7-ace042973598 · outbound

This paper cites an unresolved cited work.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Unresolved cited work

Reference 80

Resolution
unresolved
raw_fallback, observed 2026-08-12T15:16:28.736140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.959150Z digest=sha256:579926f9964483ff41ebce2b881ce81be8fa1c9f3adbb787775f45fc1f9b7922

Observation 2e368cbb-8672-4b61-b8c3-b3217dec7ca5 · outbound

This paper cites If the average score of the 30 image-text pairs from a dataset is less than or equal to 3 (on a scale of 1 to 5), the dataset is classified as having satisfactory quality.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI If the average score of the 30 image-text pairs from a dataset is less than or equal to 3 (on a scale of 1 to 5), the dataset is classified as having satisfactory quality

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.719427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.964286Z digest=sha256:d511a052e4b3c406d4ebde78bbe6e777ea059771b887fc79738d5e3f8d2c3e70

Observation 03b51d9b-25f2-4485-aa2b-6889a99144c2 · outbound

This paper cites - Assess whether there are any medical errors or misleading information that could impact diagnostic or treatment deci- sions.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI - Assess whether there are any medical errors or misleading information that could impact diagnostic or treatment deci- sions

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.702577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.969048Z digest=sha256:a3fd4dfd34512c65cec47ab93cc5594f5f0ecf77c2658785e9d3f12b58030c40

Observation 09d4e0a6-fd43-4e39-b16f-f4adbfef6216 · outbound

This paper cites - Assess whether the language is fluent and natural, and if there are any overly complex sentence structures that might hinder information transmission.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI - Assess whether the language is fluent and natural, and if there are any overly complex sentence structures that might hinder information transmission

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.685519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.973479Z digest=sha256:c1b28df36ca4dd43ad85a65d11edbd1ea2e6b6fccf77281d11e1478602b49a90

Observation 0de715af-47ef-4c43-9858-5ec3f5452e93 · outbound

This paper cites - Assess whether any essential information is missing, omitted, or incomplete.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI - Assess whether any essential information is missing, omitted, or incomplete

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.669622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.977961Z digest=sha256:c7707c52de3facb8378590a69b6a8723479feb5728b2d6662b2a798b297940dc

Observation 3f55f2e7-f557-4a4f-a3b8-076e01b899ed · outbound

This paper cites - Ensure that the imaging data is appropriately interpreted, the descriptions are clear and accurate, and they align with the medical diagnosis.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI - Ensure that the imaging data is appropriately interpreted, the descriptions are clear and accurate, and they align with the medical diagnosis

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:16:28.654564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T15:16:27.982410Z digest=sha256:7f11dd05ab46731b3ba0319b43219fdad9a9e6f2ac92de7324f6e079b06dc222

Pith citing papers

Observation ad7a026c-852b-4e47-8001-2594a43bac8c · inbound

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling cites this paper.

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Reference 134

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:23:58.182172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T13:23:57.588851Z digest=sha256:3b28d9521790e8b35df606e445e7f13dd492e50fd513cc240ad70348664c486f

Observation f0942baf-09d8-4fac-9bdc-b7bcfcf6d01c · inbound

Efficient Few-Shot Medical Image Analysis via Hierarchical Contrastive Vision-Language Learning cites this paper.

Efficient Few-Shot Medical Image Analysis via Hierarchical Contrastive Vision-Language Learning GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T20:10:05.454351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:10:05.454351Z digest=sha256:a8cff31e131e688aa129c2fc31d170da7c89da7228016a27121d5c30f5f39bdc

Observation 6c2296e6-84bd-4304-8071-13b82696856b · inbound

EndoChat: Grounded Multimodal Large Language Model for Endoscopic Surgery cites this paper.

EndoChat: Grounded Multimodal Large Language Model for Endoscopic Surgery GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T18:24:42.030232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:24:42.030232Z digest=sha256:861603c565011a2718a315714a4986aa889290f16812957b254fd61749fa049e

Observation f3e95069-62e4-42b4-a2a1-85888ad466fa · inbound

MM-Skin: Enhancing Dermatology Vision-Language Model with an Image-Text Dataset Derived from Textbooks cites this paper.

MM-Skin: Enhancing Dermatology Vision-Language Model with an Image-Text Dataset Derived from Textbooks GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T22:52:07.139701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:52:07.139701Z digest=sha256:4e7d21cdf6fa985d3a77da15e37861864aeec5180f00dff3ae71e3903ee4ffd3

Observation 8bacb87c-e481-4cec-ab32-258917010cad · inbound

Constructing Ophthalmic MLLM for Positioning-diagnosis Collaboration Through Clinical Cognitive Chain Reasoning cites this paper.

Constructing Ophthalmic MLLM for Positioning-diagnosis Collaboration Through Clinical Cognitive Chain Reasoning GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T14:51:09.732041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:51:09.732041Z digest=sha256:5931d84a13bf0e3b6b5225de9b2a9d20f5f73d3e42afe1c7753898e4f6540208

Observation ea982593-0a13-4fd5-bd2a-332eb7dacf9b · inbound

Towards Better Dental AI: A Multimodal Benchmark and Instruction Dataset for Panoramic X-ray Analysis cites this paper.

Towards Better Dental AI: A Multimodal Benchmark and Instruction Dataset for Panoramic X-ray Analysis GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T19:27:40.703899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:27:40.703899Z digest=sha256:8d0ef9011026361cb96a996aff07dd912482351fa00eefeb3841d6635ebe0253

Observation caf1a38d-2a9a-425f-b46e-274e53eb0425 · inbound

MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution cites this paper.

MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:41:25.823690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-07T14:02:54.395566Z digest=sha256:4f8b686e6065c2d1bfd0a72b1cfe08d727cae596ec4fd0cff95bb1da610d7eeb

Observation e9db9ebb-da10-459d-b82c-70780d3ad8e9 · inbound

MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution cites this paper.

MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-21T00:53:53.315397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-21T00:50:15.010488Z digest=sha256:3a4f4464212cbe7cadb43dc7dcf65d61b57ae90f7d99456e583a2b506769a63e

Observation 0c5f7290-e351-41db-ab89-abcf432d32c3 · inbound

MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution cites this paper.

MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:15:44.022138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-01T09:10:28.244690Z digest=sha256:332e3fecc328095406bc46f2c80d5ccd6ad66ce291d65b30fc06523cb3bba6f8

Observation b74a0c42-9500-4b84-9145-94c246601ae6 · inbound

MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution cites this paper.

MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-07-04T01:39:23.956373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-04T01:35:11.966506Z digest=sha256:1092b47ffb6277bad3078539b7f7462b45a647467afc471a35b38d517031a74e

Observation 09c3a83a-c961-4552-979d-203866364e18 · inbound

Aloe-Vision: Robust Vision-Language Models for Healthcare cites this paper.

Aloe-Vision: Robust Vision-Language Models for Healthcare GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T18:25:57.950271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-29T02:02:47.472868Z digest=sha256:2d9d6c7565fa6ae9bebf3178309eb88976e95b60258e0c6b7e9c9ce1fe9dc42f