Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T16:10:24.939025Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 1 inbound Pith citation observation for arXiv:2502.01243.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T16:10:24.939025Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:28:47.559357Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T15:28:49.752882Z
67 of 67 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7e99fa3b-888c-4a5c-90f2-199ff7fc5cd1 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Language models are few-shot learners
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1de9be11-abb6-45fd-949e-25ebf9fae69d · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Gpt-4 technical report
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a41e2ddc-1d59-44ec-bd1e-246b0d7f4905 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Large language models encode clinical knowledge
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0d3b57c8-360b-4980-8481-7f8ab45a301a · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Empowering biomedical discovery with ai agents
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e97a4fec-78c5-490b-9977-a1b61d53dbdd · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology A Survey on Large Language Models for Critical Societal Domains: Finance, Healthcare, and Law
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e62f7f1-09b0-4069-89fb-fe5ad06c5a24 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Integrated image- based deep learning and language models for primary diabetes care
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8d44da20-e9a1-42d2-9b23-0aee28cfa1b5 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Evaluating large language models on medical evidence summarization
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bc5c8d60-4222-432e-b65e-4f1038fa2f4b · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Adapted large language models can outperform medical experts in clinical text summarization
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5fe34d09-1522-47b6-80be-77a2e93fb158 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology A strategy for cost-effective large language model use at health system-scale
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 182b487d-c709-4e94-8f91-c865bc0549c7 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Matching patients to clinical trials with large language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a4da407d-a80b-4065-b7fb-d86764577b6f · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Scaling clinical trial matching using large language models: a case study in oncology
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation da2f6e5f-dd5a-4797-8de1-143aa2b10568 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Chatgpt and other large language models are double-edged swords, 2023
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bf522576-4202-40ae-b728-d052f4c49912 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Ethics of large language models in medicine and medical research
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c874e2de-d98c-4048-bf06-dc518f7af110 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology The ethics of chatgpt in medicine and healthcare: a systematic review on large language models (llms)
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a4f1d673-17c7-49cb-a5a8-2ab014bf55a3 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Clinical large language models with misplaced focus
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d3b74ea7-6568-4de2-9581-ade055a3e330 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b76d961a-4776-4dd0-b14d-602468740e1c · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology MedCalc-Bench: Evaluating Large Language Models for Medical Calculations
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8cc993e-f3a7-42a1-b109-dd93db1cb612 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology LongHealth: A Question Answering Benchmark with Long Clinical Documents
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 923368a5-8d70-4a75-a643-63a5321f5411 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52a7410b-c299-43fc-8f7b-65288a1214c1 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology case of the month
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 08b4cb75-e5d1-4176-a539-5258376ce987 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Accuracy and reliability of chatbot responses to physician questions
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d295919a-1296-4748-8b26-6cdc5da31ebe · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Capabilities of gpt-4 in ophthalmology: an analysis of model entropy and progress towards human-level medical question answering
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0f6d3176-eabc-423b-85c6-d0d35b2d3257 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Comparison of ophthalmologist and large language model chatbot responses to online patient eye care questions
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2a944e9d-4e16-4d65-b16e-455948925008 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Eval- uation and mitigation of the limitations of large language models in clinical decision-making
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation aa21e586-7bd3-457d-8355-710588fe5e0d · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Assessing the utility of chatgpt throughout the entire clinical workflow: development and usability study.J
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3abfeb09-357f-4ed7-a9ed-cd5a0c6b1ec4 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Benchmarking large language models’ performances for myopia care: a comparative analysis of chatgpt- 3.5, chatgpt-4.0, and google bard
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 10b8a1cc-80fb-48e3-bee7-a67abab142e6 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 19d09383-f60a-4b06-91c2-6e860c2edb0c · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Performance of large language models on medical oncology examination questions
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 04624a96-4206-4d55-9dfa-f084ee9eb189 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Assessment of a large lan- guage model’s responses to questions and cases about glaucoma and retina management
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5d08a686-870b-49d0-8346-787cd400db66 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ceff7c40-0281-4790-aa80-bf67584a4be2 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2311a256-556e-4f71-98ca-186de61b45c2 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Benchmarking large language models on cmexam-a comprehensive chinese medical exam dataset
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 617b143a-964f-4a98-9b4a-5eacc8cc4831 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Medbench: A large- scale chinese benchmark for evaluating medical large language models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f673d7e5-430d-48c9-94d8-88119444602f · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology JMedBench: A Benchmark for Evaluating Japanese Biomedical Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc0ad3c2-dcb6-41d1-9372-a96d7608da06 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c57837f2-3a88-4656-b7a3-39b439c99c3a · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology TCMBench: A Comprehensive Benchmark for Evaluating Large Language Models in Traditional Chinese Medicine
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d482c786-74a1-4a91-b16e-9883d91b5654 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology ProSA: Assessing and Understanding the Prompt Sensitivity of LLMs
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a71d522-7843-4352-9b75-b8e446b01855 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Generalization or Memorization: Data Contamination and Trustworthy Evaluation for Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e820c545-33ae-423f-92a6-4eb234ebc7d3 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Promptcblue: A chinese prompt tuning benchmark for the medical domain
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 799dc24d-aa5c-4314-b38e-520bde0bade2 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology CMB: A Comprehensive Medical Benchmark in Chinese
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 542fd470-72f3-4c92-bc67-fa278e2454c2 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology LLM-as-a-Judge & Reward Model: What They Can and Cannot Do
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f0190b6-7b45-4a20-8858-e051a840d010 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology CompassJudger-1: All-in-one Judge Model Helps Model Evaluation and Evolution
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0238ef16-b694-4f66-b618-6eee003b9da5 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0332e790-72a3-4434-ac38-195cd9d7d6db · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology PediaBench: A Comprehensive Chinese Pediatric Dataset for Benchmarking Large Language Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29c9148e-3065-47be-b73e-47d86c0a9184 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Diagnostic reasoning prompts reveal the potential for large language model interpretability in medicine.NPJ Digital Medicine, 7(1):20, 2024
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a03514cf-7089-4836-a45a-84fffe257f78 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Large language models and their impact in ophthalmology
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 95626846-fcc5-4360-afec-46667e25a53b · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology How i won singapore’s gpt-4 prompt engineering competition
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e2adcad7-6b1f-4bca-811c-b457c87c2e4e · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Alignbench: Benchmarking chinese alignment of large language models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7867d21b-fc2c-4486-8ce2-cd684a6d88ef · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Baichuan 2: Open Large-scale Language Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c66560ef-6cd1-4834-9e99-49d37cb6c9f6 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology HuatuoGPT, towards Taming Language Model to Be a Doctor
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49c815b8-5cbb-4294-b0d1-7533aea3b265 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38af74d0-459f-4dcf-8bda-8b5ab6c35937 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Tulu 3: Pushing Frontiers in Open Language Model Post-Training
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6d5d9ff-993a-41a4-b20b-8023d6e0f0cd · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology The Llama 3 Herd of Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04f9f1ef-dd9e-4ef3-b304-73b2124f4664 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Mistral 7B
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92b88e27-3741-40ac-a3f7-066036cc0196 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Pulse: Pretrained and unified language service engine
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0ffc44dc-3f9e-4072-ad93-5047c867595e · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5089a42f-448a-4ab3-b802-ee31256a8587 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Qwen2.5 Technical Report
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ba5ce7a-fd94-459e-8e69-a54ec54bef7f · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Sunsimiao: Chinese medicine llm
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d0c82835-991b-431f-be27-acdbc325c613 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Yi: Open Foundation Models by 01.AI
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 787bc067-66e8-493d-aa49-09e26bf4665c · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63b8aace-8179-418c-a570-6e8ba37b641e · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Gemma 2: Improving Open Language Models at a Practical Size
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c39933b6-4b92-4ea4-a584-cef701fc2252 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Granite 3.0 language models, 2024
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c080ce6c-7d1b-468a-801a-5a99007e2d3f · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33d9a44c-321b-4024-a45e-3c6d7aea7a3f · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology InternLM2 Technical Report
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7071e68a-a6a1-400c-b327-0df162bf3b31 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a68a059-4596-4287-89b4-b991232b8963 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology Hunyuan-Large: An Open-Source MoE Model with 52 Billion Activated Parameters by Tencent
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cda63fa-7b6c-4bb5-9f13-9021b92ca758 · outbound
OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology DeepSeek-V3 Technical Report
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71bd979c-52ef-432c-96cc-8e3fa3c7fbc4 · inbound
BEnchmarking LLMs for Ophthalmology (BELO) for Ophthalmological Knowledge and Reasoning OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.