Pith. sign in

Paper Citation Record · LEDGER

M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 89 inbound Pith citation observations for arXiv:2404.00578.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.00578 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 89 of 89 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 89 of 89 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:46:51.397088Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

16
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f184404e-2b64-4587-83fb-ee63fc83351c · inbound

MAIRA-Seg: Enhancing Radiology Report Generation with Segmentation-Aware Multimodal Large Language Models cites this paper.

MAIRA-Seg: Enhancing Radiology Report Generation with Segmentation-Aware Multimodal Large Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:17.663360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:42:17.663360Z digest=sha256:72380687809c89cc2524e4a491c2a47fd7f2041f9cdf77ed1c65a6fef7210a93

Observation 0df6d565-e58a-4be5-a0b9-3d01336fa0a3 · inbound

Large Language Model with Region-guided Referring and Grounding for CT Report Generation cites this paper.

Large Language Model with Region-guided Referring and Grounding for CT Report Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T14:15:55.972678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:15:55.972678Z digest=sha256:ac991f4b29beed332244c8a6db0e8e55b8147a00fd62818157b80d288045857d

Observation 4266b35d-f31e-4f48-b8ca-ed32fb35b455 · inbound

MG-3D: Multi-Grained Knowledge-Enhanced 3D Medical Vision-Language Pre-training cites this paper.

MG-3D: Multi-Grained Knowledge-Enhanced 3D Medical Vision-Language Pre-training M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T20:21:00.016626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:21:00.016626Z digest=sha256:35e8d0b13b0c65b148df1d36571a59bc688809e78efd6f49fc967112b1c7e3e2

Observation e14dedcb-cd1d-49fe-9acf-919d218b88b4 · inbound

From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities cites this paper.

From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T14:43:41.768579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:43:41.768579Z digest=sha256:29e8b7374f64b038ded117b5760893eca8411dbe78618654e5c934ba93f08f02

Observation bd1a39e1-ef5d-4104-a2bd-d47a6d4d4e11 · inbound

Read Like a Radiologist: Efficient Vision-Language Model for 3D Medical Imaging Interpretation cites this paper.

Read Like a Radiologist: Efficient Vision-Language Model for 3D Medical Imaging Interpretation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-11T13:03:53.081260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:03:53.081260Z digest=sha256:6cf1e10cd8fb64f7b180f0366d48651d3549feabde39521bc96e647a02823628

Observation 7cb83e7c-f1bd-4cc5-81ab-ee7c9947bc7a · inbound

ObjVariantEnsemble: Advancing Point Cloud LLM Evaluation in Challenging Scenes with Subtly Distinguished Objects cites this paper.

ObjVariantEnsemble: Advancing Point Cloud LLM Evaluation in Challenging Scenes with Subtly Distinguished Objects M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T11:54:12.097240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:54:12.097240Z digest=sha256:a79ff4d90df251db0b3a17e79fb177acf3df1fa20b5a8e99c10221e9119d8472

Observation 68ba6cf1-c3d4-4c02-9d24-eeddc0dce09d · inbound

RadGPT: Constructing 3D Image-Text Tumor Datasets cites this paper.

RadGPT: Constructing 3D Image-Text Tumor Datasets M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T21:31:28.402503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:31:28.402503Z digest=sha256:5ecc1acf264cc80e5aab2886ced5b22abb7be7d3e47cd467604caa9ae9f5fcd1

Observation 9c44e8e2-4920-4a7b-867d-393199847498 · inbound

Generalization of Medical Large Language Models through Cross-Domain Weak Supervision cites this paper.

Generalization of Medical Large Language Models through Cross-Domain Weak Supervision M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T17:36:27.848303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:36:27.848303Z digest=sha256:6507f1621e06ec0d74b7352a81e65a84a123c8bd6647e659010a142f816bb20a

Observation b1d6dc9e-ab59-4121-b2a5-e9048a3241bf · inbound

From large language models to multimodal AI: A scoping review on the potential of generative AI in medicine cites this paper.

From large language models to multimodal AI: A scoping review on the potential of generative AI in medicine M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 117

Resolution
unresolved
no resolver link, observed 2026-08-07T22:13:31.441167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T22:13:31.441167Z digest=sha256:7c0a5e275cd1c34de660e0b461740d2b843a3bb593d50541961ab1de561eb2f0

Observation 07856dd3-ad0f-4e42-8495-7a7a36f2ca8d · inbound

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding cites this paper.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.397088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.397088Z digest=sha256:69d254741f3f0c7e49cf53f8d85bcf22d32369c58693777dda868f9bbb5ceb63

Observation 0abec588-bf00-4e50-ae31-19e61f2e006e · inbound

A Review of 3D Object Detection with Vision-Language Models cites this paper.

A Review of 3D Object Detection with Vision-Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T10:13:25.100455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:13:25.100455Z digest=sha256:eb89adca9e26117edf6e64b1101b1af3ba658df2a8a4e62f5e03102ad80d233c

Observation 603f2386-1211-4df2-b3dd-0cd525f62106 · inbound

Enhancing Surgical Documentation through Multimodal Visual-Temporal Transformers and Generative AI cites this paper.

Enhancing Surgical Documentation through Multimodal Visual-Temporal Transformers and Generative AI M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T05:43:42.108010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:43:42.108010Z digest=sha256:1bf9dc7293c8b80a0019ba913f634e80b85cda8f99ea257911091e4145f47c71

Observation 80844250-8238-47c8-8c15-f6ae933f852f · inbound

Multimodal Large Language Models for Medicine: A Comprehensive Survey cites this paper.

Multimodal Large Language Models for Medicine: A Comprehensive Survey M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 276

Resolution
unresolved
no resolver link, observed 2026-08-16T05:32:55.020176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:32:55.020176Z digest=sha256:5dea5625076b94f56d61a5013709d69c96ce51a0ac843eb19778aff30061bc12

Observation ab5d0953-ae96-403a-a3b1-2556b9ae33fd · inbound

Joint Generalized Cosine Similarity: A Novel Method for N-Modal Semantic Alignment Based on Contrastive Learning cites this paper.

Joint Generalized Cosine Similarity: A Novel Method for N-Modal Semantic Alignment Based on Contrastive Learning M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T23:55:36.037745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:55:36.037745Z digest=sha256:1abce4f4baaf4f62aa72b892ed7aeb8813d6e0e6155acada988135721a212c85

Observation 81b998e8-84cd-4ebd-82ad-3ab02563bc02 · inbound

VCRBench: Exploring Long-form Causal Reasoning Capabilities of Large Video Language Models cites this paper.

VCRBench: Exploring Long-form Causal Reasoning Capabilities of Large Video Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T21:59:33.495030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:59:33.495030Z digest=sha256:247f6817155f4510b44e8885762b1bbfc41f155fc085fd3ca424ff30e8c74e92

Observation 302528a7-045a-4bef-b088-bff4db262e57 · inbound

UniCAD: Efficient and Extendable Architecture for Multi-Task Computer-Aided Diagnosis System cites this paper.

UniCAD: Efficient and Extendable Architecture for Multi-Task Computer-Aided Diagnosis System M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T21:43:30.669345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:43:30.669345Z digest=sha256:8ce3264e4fea59806bf74bccccb25658fb9924cc94fdf4e28025e137cc8e90ec

Observation 7d745b69-d3cf-4a0f-9d48-b46624b1e27f · inbound

CT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question Answering cites this paper.

CT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question Answering M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:28.332256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:28.332256Z digest=sha256:e49c1c64a9c717aa590b4f3127a42801a0aace26052e53667a544ced4bbadbe5

Observation 80578d81-f810-4025-a65d-6a6b47436b5f · inbound

Are Vision Language Models Ready for Clinical Diagnosis? A 3D Medical Benchmark for Tumor-centric Visual Question Answering cites this paper.

Are Vision Language Models Ready for Clinical Diagnosis? A 3D Medical Benchmark for Tumor-centric Visual Question Answering M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:26:28.812740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:26:28.812740Z digest=sha256:b71565771d8e5cd20aecff3a34ec520341d2a575c282fac48d014b36b15da42c

Observation 5d681346-a2ad-4700-910d-eed18fd9a76b · inbound

Medical Large Vision Language Models with Multi-Image Visual Ability cites this paper.

Medical Large Vision Language Models with Multi-Image Visual Ability M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:10.371955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:10.371955Z digest=sha256:e864b6111a59063673f09f34a398e3e6d2dfc06e8977c12e4bff0a7a5c7a817a

Observation 7a954a7c-d5a8-4442-ba12-c67648decd15 · inbound

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding cites this paper.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.628602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.628602Z digest=sha256:949bd2bf8df410035c19e3121a53a17930bb50a8be09655fba13ffd91176bb4e

Observation 8ce087ba-5932-4f0f-b9eb-e4043082d566 · inbound

CAPO: Reinforcing Consistent Reasoning in Medical Decision-Making cites this paper.

CAPO: Reinforcing Consistent Reasoning in Medical Decision-Making M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T20:10:52.036790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:10:52.036790Z digest=sha256:8105742c4d89ab6bade963b67f5cf73dc1b24946459bac28862dff08f67fc749

Observation 0b4e8ea3-63f2-4cec-96ff-fcc70fa5150d · inbound

Domain Specific Benchmarks for Evaluating Multimodal Large Language Models cites this paper.

Domain Specific Benchmarks for Evaluating Multimodal Large Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 185

Resolution
unresolved
no resolver link, observed 2026-08-07T00:39:42.343862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:39:42.343862Z digest=sha256:a328fb9174976f93cdbc7589479a8b1cf761d88882db1c44b406578df3de1af7

Observation 470dc869-5bc2-4946-ac5e-dde649d80ed8 · inbound

MedErr-CT: A Visual Question Answering Benchmark for Identifying and Correcting Errors in CT Reports cites this paper.

MedErr-CT: A Visual Question Answering Benchmark for Identifying and Correcting Errors in CT Reports M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T23:10:45.988146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:10:45.988146Z digest=sha256:209ed64e54825bbe8919415f48dead36cafcdd72e6d6fbbc40abced14a710d5f

Observation 798aebe8-d62b-4fb3-a650-7e1e5543e530 · inbound

Recurrent Visual Feature Extraction and Stereo Attentions for CT Report Generation cites this paper.

Recurrent Visual Feature Extraction and Stereo Attentions for CT Report Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:12.203692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:32:12.203692Z digest=sha256:d17a95e5e4f27777fe4613d5bac458d7092555306a6ec977e211bb5f53d94da6

Observation 3ec38a6d-29f8-4b41-ae3e-40ac27ff54f7 · inbound

Region-Aware Multimodal Large Language Model via SlowFast Tokenization and Pseudo-Mask Guidance for 3D CT Report Generation cites this paper.

Region-Aware Multimodal Large Language Model via SlowFast Tokenization and Pseudo-Mask Guidance for 3D CT Report Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:45.574493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:45.574493Z digest=sha256:98f186b657d0cdb4294b73425a15edabe8b10f0798824a8ea6d09d78431dc86f

Observation 0ae81d1b-914f-41bd-894c-5e888dec489c · inbound

CRISP-SAM2: SAM2 with Cross-Modal Interaction and Semantic Prompting for Multi-Organ Segmentation cites this paper.

CRISP-SAM2: SAM2 with Cross-Modal Interaction and Semantic Prompting for Multi-Organ Segmentation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:53:29.391871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:53:29.391871Z digest=sha256:cef75fb170c131a5ce7ba5d080313a3bf3d1e586af73d867b2b3c034928849a0

Observation 85e4cc4c-d148-4b74-ae54-7312b2313783 · inbound

Prompt Mechanisms in Medical Imaging: A Comprehensive Survey cites this paper.

Prompt Mechanisms in Medical Imaging: A Comprehensive Survey M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 198

Resolution
unresolved
no resolver link, observed 2026-08-06T22:02:25.661416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:02:25.661416Z digest=sha256:a08edcfb5c343d2c2bbceedaeb492c1e95c145c58e3eee66d22b970f348aca05

Observation 26534c7d-5826-4238-84f1-39e4d1a7b55a · inbound

MedGround-R1: Advancing Medical Image Grounding via Spatial-Semantic Rewarded Group Relative Policy Optimization cites this paper.

MedGround-R1: Advancing Medical Image Grounding via Spatial-Semantic Rewarded Group Relative Policy Optimization M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:04:54.657204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:04:54.657204Z digest=sha256:d3f27a9d6647a42b041a43e15eb9351a46a27d46d83622b2244ce1cf162aa21c

Observation e3ea3aeb-2a9b-4ddc-9717-3d168e17b137 · inbound

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing cites this paper.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:39.777120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:39.777120Z digest=sha256:da38fb664576f73f3cb15d1469dd936a6d5210230b04f3ee407149f57ea4408e

Observation 3827a0b4-4f05-43e1-a11f-de13308982b9 · inbound

Uncertainty-Driven Expert Control: Enhancing the Reliability of Medical Vision-Language Models cites this paper.

Uncertainty-Driven Expert Control: Enhancing the Reliability of Medical Vision-Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:06:23.568402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:06:23.568402Z digest=sha256:bdc5e56ebdd3820473c7eb3ce2ec92ca4e557848e8deea713aedb51fc3c41e80

Observation 4d1a4189-a266-44f6-bd16-250cd314af5f · inbound

A Comprehensive Survey of Electronic Health Record Modeling: From Deep Learning Approaches to Large Language Models cites this paper.

A Comprehensive Survey of Electronic Health Record Modeling: From Deep Learning Approaches to Large Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:18.939180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:42:18.939180Z digest=sha256:657091578ace15d49f64c73e6bebe8f0592ae39b3e20e844c68c87c02c268d7c

Observation 429a6ed5-dac3-4571-ae83-b392dadae28f · inbound

Analysis of Image-and-Text Uncertainty Propagation in Multimodal Large Language Models with Cardiac MR-Based Applications cites this paper.

Analysis of Image-and-Text Uncertainty Propagation in Multimodal Large Language Models with Cardiac MR-Based Applications M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:48.004934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:39:48.004934Z digest=sha256:005887117ab2bd8e52d37e9c2b090c832211fdb6b8a659e992e258be19e945a5

Observation ac148cf1-fe75-4dd7-9fa9-ff26107f2d1a · inbound

Cardiac-CLIP: A Vision-Language Foundation Model for 3D Cardiac CT Images cites this paper.

Cardiac-CLIP: A Vision-Language Foundation Model for 3D Cardiac CT Images M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T12:12:57.424476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:12:57.424476Z digest=sha256:eb52557d9cf275b10d838ab02bf4b8a42e28d05ef14ee29753e9ba321df814c8

Observation 5569499f-13f0-42fd-ab8f-8d331e38777e · inbound

Disorder-induced stress-flow misalignment in soft glassy materials revealed using multi-directional shear cites this paper.

Disorder-induced stress-flow misalignment in soft glassy materials revealed using multi-directional shear M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T23:25:36.790483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:25:36.790483Z digest=sha256:d9f6f5cd617fd2ef10514b1c68049097726a1b93800781cd984b2a556be3063b

Observation 1c8df07b-4282-4af2-aa22-24f49dff84f6 · inbound

VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine cites this paper.

VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:00.948324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:00.948324Z digest=sha256:a27186001f769564e6810e6d0167079c4899fb59435cac93e582d913ec8b48b5

Observation c9f01c67-46b2-4be2-8c81-5e915db4d451 · inbound

Unified Supervision For Vision-Language Modeling in 3D Computed Tomography cites this paper.

Unified Supervision For Vision-Language Modeling in 3D Computed Tomography M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T12:29:50.230776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:29:50.230776Z digest=sha256:53150ad0e025a8e7ed4fa9bdd77306160bc90be7aceb667a45eda70058253f3b

Observation 148242bc-beaf-4689-a852-a917e615e22d · inbound

Discrete Prompt Tuning via Recursive Utilization of Black-box Multimodal Large Language Model for Personalized Visual Emotion Recognition cites this paper.

Discrete Prompt Tuning via Recursive Utilization of Black-box Multimodal Large Language Model for Personalized Visual Emotion Recognition M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T13:29:52.943085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:29:52.943085Z digest=sha256:0bdd3fc3797b4c7a8e2e3bfe9f2f091d2c5d6a6f00d37edf306a9c8d8293ab75

Observation 6aba8ddf-b582-48be-b877-5da0a05f5a3d · inbound

SimCroP: Radiograph Representation Learning with Similarity-driven Cross-granularity Pre-training cites this paper.

SimCroP: Radiograph Representation Learning with Similarity-driven Cross-granularity Pre-training M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T20:51:06.573174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:51:06.573174Z digest=sha256:d8d0772964c6f2054ab98568e9fb93bc355e9704ff077266ba72b8bd46bdcfdf

Observation e3004081-6425-4a0d-ae43-e085ed86ab35 · inbound

Enhancing 3D Medical Image Understanding with Pretraining Aided by 2D Multimodal Large Language Models cites this paper.

Enhancing 3D Medical Image Understanding with Pretraining Aided by 2D Multimodal Large Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T19:48:40.847417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:48:40.847417Z digest=sha256:2fddc0dd24dab03289db8967b220dbc33370ba1fde3b1d60528e17231ef3254c

Observation d6dd44cf-da80-4248-986b-23d2752283de · inbound

MITS: A Large-Scale Multimodal Benchmark Dataset for Intelligent Traffic Surveillance cites this paper.

MITS: A Large-Scale Multimodal Benchmark Dataset for Intelligent Traffic Surveillance M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T20:33:33.091166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:33:33.091166Z digest=sha256:26cde3ad5ff2c23622d73c200ec548728e9429660e29eac5c9ed9c1db2f7b111

Observation 005877fd-9168-4760-b693-704595f14130 · inbound

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation cites this paper.

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T13:51:50.397726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:51:50.397726Z digest=sha256:35cdf1afeeb59c4aa7aad30bf47a087cae69058606f9f2964c06831fcef03bc2

Observation fe6b51a2-07cb-4b1c-b50b-ffbe728f3d2e · inbound

IBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs for Universal Biomedical Object Referring and Segmentation cites this paper.

IBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs for Universal Biomedical Object Referring and Segmentation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:33:10.136009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-16T17:31:33.903063Z digest=sha256:c3d16859b060cdec9e1c22a34f6f000183ee1e81429794bc0c7847e2202b9914

Observation bd8a8576-6537-45f3-ae60-098771088330 · inbound

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space cites this paper.

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T18:16:31.288447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:16:31.288447Z digest=sha256:1a303b5d59da654cb7d298c372dcdba083c252b4d2c47fa2f175b023866278b6

Observation 3f6ea956-afc8-4a51-acc3-c0eb28d0c475 · inbound

Medical Image Spatial Grounding with Semantic Sampling cites this paper.

Medical Image Spatial Grounding with Semantic Sampling M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-14T21:05:40.877444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T21:05:40.877444Z digest=sha256:b255f352a66fda97699fbb99b881e78f319cb9a716137283b596c70b67fead5c

Observation 48436c65-7b87-48aa-a06b-56a5ae0b7220 · inbound

Machine Learning for Network Attacks Classification and Statistical Evaluation of Adversarial Learning Methodologies for Synthetic Data Generation cites this paper.

Machine Learning for Network Attacks Classification and Statistical Evaluation of Adversarial Learning Methodologies for Synthetic Data Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-13T23:01:41.380237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T23:01:41.380237Z digest=sha256:0ee89477408b8630c278c11156ba35a6a5b9cb462d7cdb21d7d4e3c42f565b17

Observation 4a339af0-f232-449c-bfe7-6bf03b6ac85a · inbound

Visual Instruction-Finetuned Language Model for Versatile Brain MR Image Tasks cites this paper.

Visual Instruction-Finetuned Language Model for Versatile Brain MR Image Tasks M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:43:15.199309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-13T20:39:11.270334Z digest=sha256:b0a992872fbf99f663b1729e76eb66b0b7b7b5260bac2c0e4d53645385900113

Observation 138461d1-9c0c-4b30-9c66-22adfc8bbb42 · inbound

Learning Robust Visual Features in Computed Tomography Enables Efficient Transfer Learning for Clinical Tasks cites this paper.

Learning Robust Visual Features in Computed Tomography Enables Efficient Transfer Learning for Clinical Tasks M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:03:00.237995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-13T16:59:24.785428Z digest=sha256:ef4af7f5559d1e30c09f22b846d35c25a644089410b4eb0a111d5f8a74292853

Observation cb2eda88-b33c-444e-ba19-e771842c9bfd · inbound

Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis cites this paper.

Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:20:58.323661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T16:43:02.337806Z digest=sha256:f25eca5add54b5ead802f67f2e6615c992561e9ca1e11fc097172227e772adc2

Observation 86fcf4aa-8ba2-4b64-8b48-0db5b84b8ff9 · inbound

Representation geometry shapes task performance in vision-language modeling for CT enterography cites this paper.

Representation geometry shapes task performance in vision-language modeling for CT enterography M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:41:02.482725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T15:53:30.757910Z digest=sha256:c63c608ce0265f78bfe3dc20422b297148d8d21f641a6554f13aecc85f355c99

Observation a46b70fe-bfd3-4dbb-ac7c-b72bff565382 · inbound

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography cites this paper.

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:55:03.876696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T10:54:35.000783Z digest=sha256:5ab862e2ab5608e1b86e919eac454f0375185f9a2ecfc94fdce13460bdd82556

Observation 9d036104-0408-4a7d-9184-2671cf3fd640 · inbound

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography cites this paper.

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T19:46:06.596310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T19:46:06.596310Z digest=sha256:049f7765faa78c6951a3f4cac9e003f03f82d269c704d85113cf1f7307344926

Observation 7506e170-b439-4c9d-90ed-1800a6e1e08a · inbound

Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis cites this paper.

Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:27:51.878808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T08:25:57.646822Z digest=sha256:a9d6d8dd544c54dd3e232ed893253ecd45739ce2398a46a0ed54ad39d829db2c

Observation 770f8a69-294e-4fd9-82e1-1ea0c6d32b50 · inbound

MedScribe: Clinically Grounded CT Reporting through Agentic Workflows cites this paper.

MedScribe: Clinically Grounded CT Reporting through Agentic Workflows M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T05:55:31.001916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-08T19:22:09.493354Z digest=sha256:b8d36cdb3b194d9267d66d92e87079030204370c83d1fc4fde48c6cd4cb4aab7

Observation 44219f25-ef43-45c8-b1d2-95882c382539 · inbound

CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs cites this paper.

CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:41:09.638022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-08T14:49:53.357083Z digest=sha256:b438168192fc099278a7a512888460ca845cf5ba9df0f2a33c3bec6b5778e278

Observation d3d81bc0-3430-4360-8cc9-34c57dae9ada · inbound

Lost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models cites this paper.

Lost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:56:27.075904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T01:31:41.805037Z digest=sha256:aa5b7207aaa691fa799ac8e914c8a77c8b445f99319f2dfedbb3c36aefa784ef

Observation d85bdc3c-c26b-417b-b411-e0023f455b7e · inbound

Lost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models cites this paper.

Lost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T13:25:45.163252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-30T23:19:20.522732Z digest=sha256:2104e4df83fc4a06ab996dc8b2eeb01c9b473a204c9d8d3d1e46dcff4d5337c6

Observation 529e51a5-2359-4a87-b6cd-7304dd5839c8 · inbound

DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents cites this paper.

DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:41:43.481584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T04:02:36.159920Z digest=sha256:515d623381551f54ee5535573cb7d1b991854ffbefbae7be3534f4d5a2da54d0

Observation 22380ac2-6bc9-48f5-af9f-b6797fc2e88d · inbound

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology cites this paper.

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:56:21.475486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T03:55:55.359488Z digest=sha256:3f199dd1ce38bfbbfd18bd35823fcd4d28d737c89232411b23f9620506a0578a

Observation 05262b06-a4d4-4746-ad24-b0d3da38668b · inbound

M3Net: A Macro-to-Meso-to-Micro Clinical-inspired Hierarchical 3D Network for Pulmonary Nodule Classification cites this paper.

M3Net: A Macro-to-Meso-to-Micro Clinical-inspired Hierarchical 3D Network for Pulmonary Nodule Classification M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:59:27.387088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-14T20:55:57.769186Z digest=sha256:2b5dbde231d6d112a0e359a3f952728a878b3452dc32f6c790abde7299915910

Observation 7fb91275-cba4-4666-9659-b98a272329ea · inbound

CA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image Understanding cites this paper.

CA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:57:52.874218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-14T19:57:47.615619Z digest=sha256:d49855288a748d2952983d90578ec11edb8956d899054edf5170a003fce46544

Observation ca6b1399-553e-431c-ba0c-540cf4e34249 · inbound

CA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image Understanding cites this paper.

CA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T01:19:20.424112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-04T01:10:31.958276Z digest=sha256:4329875c52d9241b934cc449d2b221c3a1509de21f014ad2604dcb75b693c8c9

Observation 9d3c9cbb-ecde-4fb4-a6df-a3eaeb18f26f · inbound

Segmentation, Detection and Explanation: A Unified Framework for CT Appearance Reasoning cites this paper.

Segmentation, Detection and Explanation: A Unified Framework for CT Appearance Reasoning M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:23:37.668034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-20T18:20:38.720544Z digest=sha256:d8d9b56719a3e872012d6d618abbb2485fe844eb17084b43f276517d0d682db6

Observation 1ae3aedd-9ba1-4ffd-ab96-2a7b455fcc9c · inbound

Regulating Anatomy-Aware Rewards via Trajectory-Integral Feedback for Volumetric Computed Tomography Analysis cites this paper.

Regulating Anatomy-Aware Rewards via Trajectory-Integral Feedback for Volumetric Computed Tomography Analysis M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T08:09:51.827961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-21T08:06:02.176934Z digest=sha256:9a4b4c912102888755e85499eff77a3e4952a24c6634813fe9cde67c0069441e

Observation 2c03b03f-12fe-4206-b2d3-81676cdeb256 · inbound

NeuroQA: A Large-Scale Image-Grounded Benchmark for 3D Brain MRI Understanding cites this paper.

NeuroQA: A Large-Scale Image-Grounded Benchmark for 3D Brain MRI Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:59:45.761587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-21T06:54:55.254082Z digest=sha256:6b2cac616ef3bb06fcb94b6f263fa626f321a20e1de5628f64b438f12e620172

Observation 506a630d-b3a8-4a77-9c5b-3bbbfaea9a63 · inbound

An Open Multi-Center Whole-Body FDG PET/CT Foundation Model for Tumor Segmentation cites this paper.

An Open Multi-Center Whole-Body FDG PET/CT Foundation Model for Tumor Segmentation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:21:12.683535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-22T07:20:10.061711Z digest=sha256:40c1f987860da2299d3203cd8f2be2cd2984c57bb592fd072a3ab20360192c8c

Observation ec6342dc-bd07-4b37-824a-f33f47af1b31 · inbound

SliceWorld: A Predictive and Controllable World-State Model for CT Report Generation cites this paper.

SliceWorld: A Predictive and Controllable World-State Model for CT Report Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:04:44.781779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-30T13:57:09.053990Z digest=sha256:dc564fb827760eda0df4d023b22956d9f9f15f7e9e225bd64754d453e8d936d0

Observation b7c10c63-759f-4782-83d1-4cf296c7de96 · inbound

MedVol-R1: Reward-Driven Evidence Grounding for Volumetric Reasoning Segmentation cites this paper.

MedVol-R1: Reward-Driven Evidence Grounding for Volumetric Reasoning Segmentation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:33:50.891797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-29T18:26:10.873395Z digest=sha256:75155b25385a83d609bb53618138358976718278a3bf62c6ce1ca4a310bd7436

Observation 39b0853b-74e0-4be6-bbfc-bf2773dab466 · inbound

Astra: a generalizable report generation foundation model for 3D computed tomography cites this paper.

Astra: a generalizable report generation foundation model for 3D computed tomography M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:02:50.405512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-28T23:16:12.442334Z digest=sha256:5a9b2ff1a472d4cd74d9e66034f208a0362c9676f81e61ba186c5eba1f798481

Observation 71075a4f-7495-4349-a7e9-a8c86250dc02 · inbound

CardioLens: Revealing the Clinical Reality Gap of MLLMs via Multi-Sequence Cardiac MRI Evaluations cites this paper.

CardioLens: Revealing the Clinical Reality Gap of MLLMs via Multi-Sequence Cardiac MRI Evaluations M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:33:15.355195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-29T08:28:48.174508Z digest=sha256:3860954d7dbb8b683dc744c945bf75401bf7d013b342774e8ddd7b7a5ae886b4

Observation a0f3b2d1-8329-4eb8-8fb6-8a321e9d304f · inbound

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training cites this paper.

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:22:34.679589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-28T19:16:42.139096Z digest=sha256:f9e99fdb1c2b71e3e246bc78df083047544bbefe2383e6af612eea1f34e3c089

Observation af9e1472-583d-4ac1-a46a-fffed3d884a6 · inbound

Multi-Granularity 3D Kidney Lesion Characterization from CT Volumes cites this paper.

Multi-Granularity 3D Kidney Lesion Characterization from CT Volumes M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:06:44.625727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-28T07:06:49.770326Z digest=sha256:962e08499c7d23d0b217fd7509e8e42183d1252c14e798ce8c778d7a8df1edb5

Observation defe729f-ada0-415d-8659-a556e91a25c9 · inbound

UniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQA cites this paper.

UniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQA M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 165

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T09:47:59.536498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-27T10:21:12.782864Z digest=sha256:5e8e2da8c3296f76399b24f32024bfffba802bc43fb1d7ca48c40ac35dd46263

Observation 36072f2b-9064-4ff3-85b0-5f89be424a96 · inbound

Venice-H1: Failure-Aware Query Re-Ranking with Multi-Scale Grid Signatures for Referring Image Segmentation cites this paper.

Venice-H1: Failure-Aware Query Re-Ranking with Multi-Scale Grid Signatures for Referring Image Segmentation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-26T11:29:24.583964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-26T11:22:05.498644Z digest=sha256:85fd4e36601f7ba50ac7b458fc5072f8014055b81aa7cb2768245495f3e42d12

Observation f994ea63-e58b-437c-b779-31ae20145b28 · inbound

E-MRL: Cross-view Aligned Evidence-driven Multimodal Reinforcement Learning for Reliable 3D Tumor Analysis cites this paper.

E-MRL: Cross-view Aligned Evidence-driven Multimodal Reinforcement Learning for Reliable 3D Tumor Analysis M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T12:49:52.204124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-26T06:02:30.459845Z digest=sha256:2a73bdc06a80bdeca65b64a93558437de1b229e2f78698e77ef57aca5accf7fc

Observation 3789f371-cba5-43a9-85ca-9b5e5f876c78 · inbound

MRI2Rep: Autoregressive Structured Report Generation for 3D Liver MRI cites this paper.

MRI2Rep: Autoregressive Structured Report Generation for 3D Liver MRI M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T20:00:08.150092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-25T20:55:03.803170Z digest=sha256:0254c20aeeb6acb074bf663015de477bb2f813978b95f7b940ce9db3d2e5ad13

Observation 06ff476a-e955-41d5-84e8-9510e2c81c41 · inbound

TRACE: A Concept Bottleneck Model for Longitudinal 3D Glioblastoma Response Assessment cites this paper.

TRACE: A Concept Bottleneck Model for Longitudinal 3D Glioblastoma Response Assessment M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:54:19.947742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-30T06:52:42.092509Z digest=sha256:8870e169e1cb716ccecd2e6cedc8b9e14f16c0e7f71faa10d0eddc2c602e117d

Observation f22dbbf4-8137-445d-8294-8f6817a00ba9 · inbound

TRACE: A Concept Bottleneck Model for Longitudinal 3D Glioblastoma Response Assessment cites this paper.

TRACE: A Concept Bottleneck Model for Longitudinal 3D Glioblastoma Response Assessment M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:39:00.916214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-03T22:29:18.856647Z digest=sha256:503e5e490727bfe67f98b0a07fefeb7854590fd84e406450c281e42d0b3bd032

Observation fa7b3241-1b38-45b5-9007-bddae5600c32 · inbound

Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models cites this paper.

Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T03:31:45.407270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:31:45.407270Z digest=sha256:992079652ae8c4d689829d6da6a7724865638e6dd8f904386593dee266b51a2c

Observation 1e39b3dd-2d4e-4ff0-bd57-106651514303 · inbound

Multi-LLM Collaborative MRI Report Generation for Visual Instruction Tuning in Brain Oncology cites this paper.

Multi-LLM Collaborative MRI Report Generation for Visual Instruction Tuning in Brain Oncology M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T01:44:20.068459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:44:20.068459Z digest=sha256:43e3922470338e8cc5e23ccf24afd8d3d980f467da35a94280e8b0269f7c1e2a

Observation 9e6e25e0-7723-432d-8ada-360944d57f86 · inbound

Cheap Probes Predict Expensive Training in 3D-CT Vision--Language Models cites this paper.

Cheap Probes Predict Expensive Training in 3D-CT Vision--Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T06:15:35.089393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:15:35.089393Z digest=sha256:bd74ad8c24e0f756267d22554af8922f67200d3e09f9b4bb8f26a5cdbe8a48c1

Observation 9b759028-84a6-47a2-ba64-7c60aba90e4a · inbound

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding cites this paper.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.953458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.953458Z digest=sha256:c7ad619252931cc95f47529461edec57a3b9f355b6eea562ce44b9bd1ba2ab6e

Observation c0c13d68-673d-4e3d-8620-2bbc419e8d3d · inbound

Rad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed Tomography cites this paper.

Rad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed Tomography M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-01T00:38:25.843369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:38:25.843369Z digest=sha256:bfe0986d1635ef5c7e90f7f13bb36a451e1653b65011e35a86a1bf29398d906e

Observation dfb6c282-c217-4b04-887e-8be7488d40c8 · inbound

MedARC: Training-Free Adaptive Redundancy Compression of Visual Tokens for 3D Medical Vision-Language Models cites this paper.

MedARC: Training-Free Adaptive Redundancy Compression of Visual Tokens for 3D Medical Vision-Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T13:31:49.000325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:31:49.000325Z digest=sha256:1dbf255c95057188392c597aebe47669e25e361619a8809526783455d296bcd1

Observation da340ddf-4a0f-4494-8801-ff5e61ec899e · inbound

Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA cites this paper.

Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T05:36:44.874743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:36:44.874743Z digest=sha256:eb3b251f779099a4a926edf4234309d862d99a7927b698b8b8dba5257476cf98

Observation ddb36263-bcb2-4c54-af9f-00d976494409 · inbound

Learning How Much, Not Just What: Cross-Patient Burden Order for CT Vision-Language Pretraining cites this paper.

Learning How Much, Not Just What: Cross-Patient Burden Order for CT Vision-Language Pretraining M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T01:00:54.281987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:00:54.281987Z digest=sha256:ce19e31b3129b976cc048d0f616ef1096d6c47f45e8f487f08c6170c8c2dd22a

Observation bdc3c00f-85c1-44b6-866d-4c7cdd0290d8 · inbound

ORCA: ORgan-Centroid Aggregation for Training-Free 3D CT Visual Token Compression cites this paper.

ORCA: ORgan-Centroid Aggregation for Training-Free 3D CT Visual Token Compression M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T00:46:03.549411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T00:46:03.549411Z digest=sha256:2b4191a02afccb8f3159184d101b9922f8eb10824c699cf1e22951a1c97d8b29

Observation ce51c620-3de8-41b5-bf9e-30bebf1ba67a · inbound

CARVE: Cross-Slice Anisotropic Reallocation of Visual Evidence for Efficient 3D Medical Volume Understanding cites this paper.

CARVE: Cross-Slice Anisotropic Reallocation of Visual Evidence for Efficient 3D Medical Volume Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T14:41:24.448336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:41:24.448336Z digest=sha256:a0fdc3fd4f399293b15a330e8139ba926d3fe4d90ffbd58f131b578691d5384b

Observation ecc433c6-e8e9-4b99-a967-8151a028bf13 · inbound

Resolution Meets Reduction: Efficient Visual Context for 3D Radiology Report Generation cites this paper.

Resolution Meets Reduction: Efficient Visual Context for 3D Radiology Report Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:15.838226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:15.838226Z digest=sha256:5282549e83831951bba8974693be1395172c4ce14f8e8ef1d861cbfa19f6d285

Observation 74802f81-024e-4f16-a846-98c0c0750a94 · inbound

HounsWorld: A Multimodal World Model for Hidden Patient-State Readout, Reconstruction, and Simulation cites this paper.

HounsWorld: A Multimodal World Model for Hidden Patient-State Readout, Reconstruction, and Simulation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T21:00:35.740016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:00:35.740016Z digest=sha256:b5fe658e8df3ef2265525e95a2b22b070070083daaf60b30ec623aa858a52ce3