Pith. sign in

Paper Citation Record · LEDGER

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation

As of 10 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2607.09142.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.09142 v3

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T07:43:33.938358Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved59
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6a1ad8a4-881d-46b2-8649-94a2b439bebc · outbound

This paper cites Digital health market statistics 2026: Market size, investment and user adoption, 2026.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Digital health market statistics 2026: Market size, investment and user adoption, 2026

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:25.688712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:25.688712Z digest=sha256:88a55928b3fd07f77f04d16b5c9de6ecefaa5f30a3a19822591902e70c156e78

Observation eee4d6a4-b9de-4afc-ae16-2833153e5602 · outbound

This paper cites Telemedicine and AI in healthcare: Top platforms, trends & 2026 outlook.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Telemedicine and AI in healthcare: Top platforms, trends & 2026 outlook

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:25.786509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:25.786509Z digest=sha256:aa861f321c3c0713846c94159d99adf3f50019520c2d5319be0c6dc8043c8abf

Observation 876d928c-fc1c-4adf-9ce4-4963652761ec · outbound

This paper cites Annual results announcement for the year ended december 31, 2025.HKEXnews,.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Annual results announcement for the year ended december 31, 2025.HKEXnews,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:25.910724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:25.910724Z digest=sha256:1c7f81f70c5d25eeeeba3e9cf3f188969ff409b2d13526bd27a5f110ed0557d4

Observation 73599099-f7c4-40ec-b3a6-f37940c1d36d · outbound

This paper cites Jd health introduces groundbreaking LLM-powered suite for comprehensive online and in-hospital healthcare scenarios.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Jd health introduces groundbreaking LLM-powered suite for comprehensive online and in-hospital healthcare scenarios

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:26.122716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:26.122716Z digest=sha256:809534cbd76d54733315001f37d83175172984260a450af03cb2c44d7bcba327

Observation ff60bbc3-5ecb-4029-8655-37e3fc8bcbc0 · outbound

This paper cites Large language models in medicine.Nature Medicine, 29(8):1930–1940, 2023.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Large language models in medicine.Nature Medicine, 29(8):1930–1940, 2023

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:26.257849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:26.257849Z digest=sha256:a31d0d3407f259c5f05b67c71fcaba8a33afe9c137d9b69f786c9bc2030e4e17

Observation 432496d1-5fcf-4220-99ca-6085f0d7f763 · outbound

This paper cites The effect of using a large language model to respond to patient messages.The Lancet Digital Health, 6(6):e379–e381, 2024.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation The effect of using a large language model to respond to patient messages.The Lancet Digital Health, 6(6):e379–e381, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:26.413588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:26.413588Z digest=sha256:ea9f22a5301be698beabcbafff44783737578affe4d81b47c5dbea1a79a607b5

Observation 181fbc43-a92c-4a1f-afac-30ed9ebc589b · outbound

This paper cites Sara Mahdavi, Sushant Prakash, Anupam Pathak, Christopher Semturs, Shwetak Patel, Dale R.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Sara Mahdavi, Sushant Prakash, Anupam Pathak, Christopher Semturs, Shwetak Patel, Dale R

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:26.523483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:26.523483Z digest=sha256:043f54b81476a461094d02fddbb9809f1d02ecacdb67520ba96fa53a16c1e6c5

Observation 0918ed26-4579-4a86-a3b7-4623dd1e138a · outbound

This paper cites PubMedQA: A Dataset for Biomedical Research Question Answering.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation PubMedQA: A Dataset for Biomedical Research Question Answering

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:26.664483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:26.664483Z digest=sha256:14ca094fa2a8673ca2ddfe8354fe4909de1db4cf285fbb9c27c909c97177e6b6

Observation e177c21c-8e1e-4394-a07f-0a66afc3b90b · outbound

This paper cites What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:26.823891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:26.823891Z digest=sha256:3680f5e21fe12004a39e8635ce7bdddc80c0d06b84fee34b4a6ca57e1d27f1b5

Observation 183fe46d-a841-4df4-9fc0-51fcf547f4ae · outbound

This paper cites MedMCQA : A Large-scale Multi-Subject Multi-Choice Dataset for Medical domain Question Answering.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedMCQA : A Large-scale Multi-Subject Multi-Choice Dataset for Medical domain Question Answering

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:26.937401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:26.937401Z digest=sha256:81ccbe9cc3784eb7846e965215c4611b2a0806f33cf092854eec87de63d6571d

Observation f508afda-9752-4fa1-aa52-4d8214ec7c63 · outbound

This paper cites Large Language Models Encode Clinical Knowledge.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Large Language Models Encode Clinical Knowledge

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:27.075346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:27.075346Z digest=sha256:43d9b78da314839701266c4b1bd7756c8f274e9833b9bd08b45ef98d871bf821

Observation 1870a552-4732-4b07-9ad8-5c25893fe155 · outbound

This paper cites Towards Expert-Level Medical Question Answering with Large Language Models.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Towards Expert-Level Medical Question Answering with Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:27.233202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:27.233202Z digest=sha256:242538df696e7a247657181233e8a032aab602b6948268ab99f460c7fde41692

Observation fbb0743a-4cbf-4d02-9609-232b224dcf38 · outbound

This paper cites CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:27.362450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:27.362450Z digest=sha256:e25c9282960ff7320fabbf56892394847310e1a7057aa94998c845e4a7e76158

Observation c0434527-5519-4fc2-b6b4-94de581d78a6 · outbound

This paper cites PromptCBLUE: A Chinese Prompt Tuning Benchmark for the Medical Domain.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation PromptCBLUE: A Chinese Prompt Tuning Benchmark for the Medical Domain

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:27.511156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:27.511156Z digest=sha256:11c6a728f2937afde096ca86a03d6805024c7acb82914cfbd6e25fcfdbebb90b

Observation 8294e180-58e9-43c2-a5cb-53108082f788 · outbound

This paper cites Benchmarking Large Language Models on CMExam -- A Comprehensive Chinese Medical Exam Dataset.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Benchmarking Large Language Models on CMExam -- A Comprehensive Chinese Medical Exam Dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:27.681801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:27.681801Z digest=sha256:0d816657bedc328e5107448f756cdb61372f7648c3df088af8b6776ec921cfdd

Observation 0b73336c-7c26-4e88-b773-3f6ccf58c482 · outbound

This paper cites CMB: A Comprehensive Medical Benchmark in Chinese.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation CMB: A Comprehensive Medical Benchmark in Chinese

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:27.798869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:27.798869Z digest=sha256:2de0488c96672eeed5207d2789ce89616c172a3776c80d446319e87f4fff236e

Observation cde0045c-0945-4411-bed9-9f84f6eb9ca6 · outbound

This paper cites MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:27.951292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:27.951292Z digest=sha256:962e0618f0ebfc51badb4911f0d4d2ad7373ff3e1b4be54cdb307459b36fa69c

Observation 9341c942-ee09-4468-99a3-3675518a7e40 · outbound

This paper cites MedBench v4: A robust and scalable benchmark for evaluating chinese medical language models, multimodal models, and intelligent agents.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedBench v4: A robust and scalable benchmark for evaluating chinese medical language models, multimodal models, and intelligent agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:28.100984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:28.100984Z digest=sha256:a437f5b1c7170762703899e042c786e1727a8a9cdc8b0ea67e13dad940cadc3d

Observation 462b7725-9db1-45be-a48c-cc1c18e7da5d · outbound

This paper cites MedDialog: Two Large-scale Medical Dialogue Datasets.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedDialog: Two Large-scale Medical Dialogue Datasets

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:28.248603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:28.248603Z digest=sha256:28cc17bc14fca2b1fce4ec69d27bf773d85b753f37a11ee61be12d6021b82d07

Observation d7c0bf2f-8490-46d8-a4f5-388493ca024b · outbound

This paper cites MedDG: An Entity-Centric Medical Consultation Dataset for Entity-Aware Medical Dialogue Generation.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedDG: An Entity-Centric Medical Consultation Dataset for Entity-Aware Medical Dialogue Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:28.362118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:28.362118Z digest=sha256:aa10de8c053a93e7d05b60f98f0b191d9323ec0fbcd97f66fca3cb1e7b656e17

Observation ed20a62f-3db0-48d9-9903-f9a7e420320e · outbound

This paper cites MediTOD: An English Dialogue Dataset for Medical History Taking with Comprehensive Annotations.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MediTOD: An English Dialogue Dataset for Medical History Taking with Comprehensive Annotations

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:28.563988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:28.563988Z digest=sha256:bd078dc2e32fbcc73208317d9ca4d5be9b65b226a2d0ed725174f7efbffc954b

Observation e3f088dc-644e-4366-b52e-6ba78e64699b · outbound

This paper cites An Automatic Evaluation Framework for Multi-turn Medical Consultations Capabilities of Large Language Models.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation An Automatic Evaluation Framework for Multi-turn Medical Consultations Capabilities of Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:28.746413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:28.746413Z digest=sha256:f59c7fa21b780f2b046a54564258a0fda818b4b90c6f2bb0e7961d52c234edc2

Observation 310599d4-a556-4d3c-9593-20d0146ef02c · outbound

This paper cites MediQ: Question-Asking LLMs and a Benchmark for Reliable Interactive Clinical Reasoning.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MediQ: Question-Asking LLMs and a Benchmark for Reliable Interactive Clinical Reasoning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:28.926838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:28.926838Z digest=sha256:c85617559317d9c6dac2b2def89181107bbbd7dcf762bb5955ac2dfc1d8d81e6

Observation 3979da11-9bc2-4c55-992e-7e00b5a2271c · outbound

This paper cites AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:29.119519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:29.119519Z digest=sha256:924dd28e1da229159e201dbf6e34b28e00234a44b341237b53ec006763456664

Observation e5d4cf8c-894c-4615-ac3a-d8e70632fee0 · outbound

This paper cites The dialogue that heals: A comprehensive evaluation of doctor agents’ inquiry capability.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation The dialogue that heals: A comprehensive evaluation of doctor agents’ inquiry capability

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:29.236433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:29.236433Z digest=sha256:bcf9d1a077308eca42d9759f7cae2766a9d609a65b39adab93c49b2ba1adb537

Observation ad75dbaa-267b-446c-91a2-c19939d35b20 · outbound

This paper cites MedConsultBench: A full-cycle, fine-grained, process-aware benchmark for medical consultation agents.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedConsultBench: A full-cycle, fine-grained, process-aware benchmark for medical consultation agents

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:29.366837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:29.366837Z digest=sha256:2949b313cb3a790aa2133a5511400c34d8b4ae5aa7e5603aeaa11bc96f87d12d

Observation b918b6c6-5b1a-485e-a607-543c27840528 · outbound

This paper cites MedDialBench: Benchmarking LLM Diagnostic Robustness under Parametric Adversarial Patient Behaviors.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedDialBench: Benchmarking LLM Diagnostic Robustness under Parametric Adversarial Patient Behaviors

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:29.543299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:29.543299Z digest=sha256:f9a4b4576b0f447a3a22888e2e6a978a848b93494e2363d52900d310d0102889

Observation 5c28d1b3-da97-4557-b02f-1ee05f199559 · outbound

This paper cites HealthBench: Evaluating Large Language Models Towards Improved Human Health.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation HealthBench: Evaluating Large Language Models Towards Improved Human Health

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:29.669791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:29.669791Z digest=sha256:0b84a771bbef98cb7e6fdee8d7d34451dc90eb507b375f715c94317df84b00d4

Observation a3b41ff3-cf2a-498a-841c-860de381de38 · outbound

This paper cites MedDialogRubrics: A comprehensive benchmark and evaluation framework for multi-turn medical consultations in large language models.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedDialogRubrics: A comprehensive benchmark and evaluation framework for multi-turn medical consultations in large language models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:29.782033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:29.782033Z digest=sha256:71d8ab1509636abaa3f2f12a7fafba3d69f4df0d595a633b7604491a1897de84

Observation 03862263-cc3a-4a61-8712-b34e585cd437 · outbound

This paper cites HealthBench Professional: Evaluating Large Language Models on Real Clinician Chats.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation HealthBench Professional: Evaluating Large Language Models on Real Clinician Chats

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:29.880526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:29.880526Z digest=sha256:fc0e07f4b13aed5432f147eddd01660225dcbafb9cda62592b70308be0389239

Observation e3c44edf-094f-4393-b43d-1117baa87024 · outbound

This paper cites LiveMedBench: A contamination-free medical benchmark for llms with automated rubric evaluation.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation LiveMedBench: A contamination-free medical benchmark for llms with automated rubric evaluation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:29.950426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:29.950426Z digest=sha256:7420e4edb1a690feae2ead6b599c3bd5da5762d48f30387f25ccc47b5803fd2f

Observation ba18ef32-8a30-44b3-8a64-c2b74b4e3b4e · outbound

This paper cites Pathological Visual Question Answering.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Pathological Visual Question Answering

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:30.054286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:30.054286Z digest=sha256:4cbc506d94ccbc662365cb83469311901ed085fc69f2d10efcc73bb7afb84a40

Observation 369af46b-79f9-4318-a36d-a07bcb501226 · outbound

This paper cites SLAKE: A Semantically-Labeled Knowledge-Enhanced Dataset for Medical Visual Question Answering.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation SLAKE: A Semantically-Labeled Knowledge-Enhanced Dataset for Medical Visual Question Answering

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:30.225146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:30.225146Z digest=sha256:87bf016316d54c68edfbe6e17e03bfb1845031bec1dfcf7ad660200405b0679a

Observation f258a7b8-4738-47c0-b000-21122bbb8a47 · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:30.424951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:30.424951Z digest=sha256:2110dc65e74dbcbc1d99551bf38dd55e8b91e84b78196bb3988e75dfa60a9f33

Observation b3bd1d15-802d-4ec8-9978-b111521d6969 · outbound

This paper cites OmniMedVQA: A New Large-Scale Comprehensive Evaluation Benchmark for Medical LVLM.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation OmniMedVQA: A New Large-Scale Comprehensive Evaluation Benchmark for Medical LVLM

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:30.666324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:30.666324Z digest=sha256:b9b2aeb5a2acc04713a4f51ffb086f660f25d95929eca90763fcfa727ed9ff51

Observation 3021863a-813b-46d5-95f9-a8515395b7e7 · outbound

This paper cites GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:30.865915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:30.865915Z digest=sha256:397a6c2557af052c667009323dfe81c24df047269a07f55b66bb424c793587c1

Observation 3b39bde9-9457-408e-8b54-358d48810fcd · outbound

This paper cites 3MDBench: Medical multimodal multi-agent dialogue benchmark.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation 3MDBench: Medical multimodal multi-agent dialogue benchmark

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:31.006709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:31.006709Z digest=sha256:63acd47baac1eef5fac583995ec65fc7a98b344519712efd7bae4f5e852968e9

Observation 354b8dcf-1fa4-42c8-97f1-151abf94012e · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Bleu: a method for automatic evaluation of machine translation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:31.110964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:31.110964Z digest=sha256:fd542775dd77d34564843013bd845b3b77d46c06562468a78a17fb933335fbe8

Observation d18b794b-8000-48aa-b3d7-4d279717bb7b · outbound

This paper cites ROUGE: A package for automatic evaluation of summaries.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation ROUGE: A package for automatic evaluation of summaries

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:31.410371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:31.410371Z digest=sha256:a70c6476120bf69cbe87b1d043c75dd6770875261d61e968eac27cf4486f980d

Observation 06ee9cf1-7695-493b-bb5d-11eacaead115 · outbound

This paper cites Automated Rubrics for Reliable Evaluation of Medical Dialogue Systems.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Automated Rubrics for Reliable Evaluation of Medical Dialogue Systems

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:31.569388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:31.569388Z digest=sha256:1dfa03b9f1d0113ad103dd92676ae3c8cabbb05b9f4a89ff018723c152b64a6f

Observation c2fb55a3-c897-4da8-a337-a77d5b0a6f0d · outbound

This paper cites Deid-gpt: Zero-shot medical text de-identification by gpt-4.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Deid-gpt: Zero-shot medical text de-identification by gpt-4

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:31.765489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:31.765489Z digest=sha256:05c1d34f8f49e4dc2f20ad3b16fc80728dea56321c2cc29bf7cdf1bfd35f2bbd

Observation c416349b-22aa-4f58-9327-06a7940f1254 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:31.943930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:31.943930Z digest=sha256:fe1df6f4c47258df2934b0def4a102f2c6c5c6518ddd82c473157d826e65d9c0

Observation b4e77de7-b0ea-483e-b1da-89d4c3ce8286 · outbound

This paper cites Claude 4 Opus (versions 4.6 and 4.7).

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Claude 4 Opus (versions 4.6 and 4.7)

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:32.184738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:32.184738Z digest=sha256:31472deb62a7683406f3bc70bd04368fe3f38ffdde3689ad3be0816a9766385d

Observation f335b87c-5989-4339-99ad-8c944f2be03d · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Gemini: A Family of Highly Capable Multimodal Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:32.274199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:32.274199Z digest=sha256:6c8b0a3a87aa9958bf4dde3e147f10e692853eccd99693394c96c37682452d58

Observation 4477c2c2-fcd2-49a8-b58b-b394231ec60b · outbound

This paper cites Evaluation and mitigation of the limitations of large language models in clinical decision-making.Nature Medicine, 30(9):2613–2622, 2024.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Evaluation and mitigation of the limitations of large language models in clinical decision-making.Nature Medicine, 30(9):2613–2622, 2024

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:32.428880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:32.428880Z digest=sha256:38548aabc216c1e43e8bb770d0ceb21442f7ba19b7366d8e42e5d15820c9ae75

Observation 8e2217cf-6a56-44d0-a323-8ebf1191b520 · outbound

This paper cites Computing inter-rater reliability and its variance in the presence of high agreement.The British Journal of Mathematical and Statistical Psychology, 61(Pt 1):29–48, May 2008.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Computing inter-rater reliability and its variance in the presence of high agreement.The British Journal of Mathematical and Statistical Psychology, 61(Pt 1):29–48, May 2008

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:32.524263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:32.524263Z digest=sha256:03553be6f036bc46176a2eed2096b035a27323f413be2a8013b59acbee1036fa

Observation 4483cad1-d5f7-449f-9481-80bd33ee5731 · outbound

This paper cites OpenAI GPT-5 System Card.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation OpenAI GPT-5 System Card

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:32.654513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:32.654513Z digest=sha256:2dc35da8e66ab341b5092f7720a7a1e9a7ed3fd6ac32f52dc09c595c8bfb8105

Observation e264408a-9830-4863-818a-3dbf2d558523 · outbound

This paper cites Kimi K2.6: Open-weight trillion-parameter MoE agent model.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Kimi K2.6: Open-weight trillion-parameter MoE agent model

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:32.828616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:32.828616Z digest=sha256:b6ca41f7e6d612b607f14baa8a5d1cacfb3861484b7d58d1a7420df134a04a3a

Observation afe5a4f5-1d7d-4b54-98e0-74245ae93649 · outbound

This paper cites Kimi K2.5: Visual Agentic Intelligence.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Kimi K2.5: Visual Agentic Intelligence

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:32.919875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:32.919875Z digest=sha256:dc578f59b8c6b03a9c5ad95c56573f09e27df1ca93d656c27340a531eed875f2

Observation 970a2151-1b35-43bd-97e3-5f929883a890 · outbound

This paper cites Qwen3.6-27B and Qwen3.6-35B-A3B.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Qwen3.6-27B and Qwen3.6-35B-A3B

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:33.082950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:33.082950Z digest=sha256:aede8b79e258bae0db42a0b350d2d75e62cc652efd6dfeb4f46482a1c217fa32

Observation 0082b4cb-43a1-4ea4-952d-5adb8e7403fc · outbound

This paper cites GLM-5: from Vibe Coding to Agentic Engineering.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation GLM-5: from Vibe Coding to Agentic Engineering

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:33.163310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:33.163310Z digest=sha256:bf6eb5e8f1e62d27e1589de6c6d69540aa3a139675b7293304fa8786da53a200

Observation b5a2e2d9-9b8f-414f-9671-8936850fe550 · outbound

This paper cites Deepseek-v4: Towards highly efficient million-token context intelligence.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Deepseek-v4: Towards highly efficient million-token context intelligence

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:33.294868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:33.294868Z digest=sha256:a54b253c96c92407fc3309f6a2e3f57c7577e0d8b44df087587bcf662bdb7228

Observation 6913a623-f7c6-4faa-a95e-6d8849a9926b · outbound

This paper cites MedGemma Technical Report.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedGemma Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:33.376539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:33.376539Z digest=sha256:816394d94944e5875fdd28f5542c1325d3bc986d40c6a2c1a69df397eee21ce8

Observation 6a4e171b-9759-4461-9e21-9b52c520e620 · outbound

This paper cites Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:33.468367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:33.468367Z digest=sha256:2cc2fa7782ff01db390bf294e04676fa361f337d93f1897658e4007f6c5cd92d

Observation 36c637cd-853b-4147-8cb6-3d6cbb375d55 · outbound

This paper cites HuatuoGPT-3-32B large language model repository.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation HuatuoGPT-3-32B large language model repository

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:33.592310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:33.592310Z digest=sha256:08e966c9a892e1ec230af0665a60f44ec938f05604523b71a982a2b3e5e6c139

Observation d7d60289-1b8a-44f2-825d-fb59cdb324b6 · outbound

This paper cites AntAngelMed medical large language model repository.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation AntAngelMed medical large language model repository

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:33.700922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:33.700922Z digest=sha256:e97a45ec0afcf6da44faee0181b9fb3cac28588e6827ed581014ee8b22bb9131

Observation b90d72ee-6280-46d1-ae02-f1cfaf8a4bf7 · outbound

This paper cites Baichuan-m3: Modeling clinical inquiry for reliable medical decision-making.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Baichuan-m3: Modeling clinical inquiry for reliable medical decision-making

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:33.813874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:33.813874Z digest=sha256:3469e0ef4b10e00c25850ab09c509c9b491fd725b64a2a7c40588f0240dc868c

Observation 43ace38b-6942-4848-ac4b-abb59379089c · outbound

This paper cites Baichuan-M2: Scaling Medical Capability with Large Verifier System.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Baichuan-M2: Scaling Medical Capability with Large Verifier System

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:33.938358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:33.938358Z digest=sha256:b60248467eea827754a8087082f386e035f9dae0135ec1d7094c147484c6b629

Observation d307073a-6f3b-4733-b603-53b7b3bc5a33 · outbound

This paper cites an unresolved cited work.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Unresolved cited work

Reference 2002

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:31.271033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:31.271033Z digest=sha256:5c92be86d2a3762427e6e12163fc2ab0932db9203aee50983f3a8c7e287c6f06

Observation a08c4839-7f50-40b1-8dca-8061cac79deb · outbound

This paper cites an unresolved cited work.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation Unresolved cited work

Reference 2026

Resolution
parse uncertain
no resolver link, observed 2026-08-02T07:43:26.014553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:26.014553Z digest=sha256:4bea1d1d7d991841c28a53759ed9b3e24aae7bcc595ad668d5b0c0daf7440ca6

Pith citing papers

No inbound Pith citation observations are available.