Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T04:22:52.084114Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 10 inbound Pith citation observations for arXiv:2412.18947.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T04:22:52.084114Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T10:18:08.734883Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T14:38:28.969559Z
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation aa473d65-0605-4dbe-a5a6-5ddd948fca16 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models , " * write output.state after.block = add.period write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e2d7b8c-28af-46c5-8145-37fbd8c9f3d8 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 500e1f1e-5a0b-4ba9-87fe-ac0fbf925fdd · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ec220f59-a8ca-4b08-8204-cdfecc4eb85e · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84b98a02-031a-4cd0-9c35-525da82b8b75 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d33dbae8-1d57-4a8c-9878-8b92c2244cfa · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b026341-da33-4779-b408-0f5b920fc94b · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Few-shot learning for medical text: A systematic review
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dbe6d4fc-a215-4b5e-90eb-560cb5121b17 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb28b985-06af-43cb-ad36-9e8cd08959cb · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Data Collection and Labeling Techniques for Machine Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 009a210e-faa4-4f71-b2de-e818d5809a10 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models MedExQA: Medical Question Answering Benchmark with Multiple Explanations
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff6f7e78-af5a-453f-a981-a9400c54ea1c · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models H.; Cheatham, M.; Medenilla, A.; Sillos, C.; De Leon, L.; Elepa \ n o, C.; Madriaga, M.; Aggabao, R.; Diaz-Candido, G.; Maningo, J.; et al
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e0740159-2352-467e-bfa5-415517da60a8 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ec1325cb-2eba-4ec2-9bee-2ce55d14d874 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93120263-676b-446c-af1e-99649fc578ef · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models HaluEval: A Large-Scale Hallucination Evaluation Benchmark for Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23d7430d-49bc-493f-abd4-d122d1f16c93 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e09cb56-0dbd-4929-bd9e-494d9019eaac · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8aedec1f-4b93-45d6-99f9-6f01747208bc · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Visual Instruction Tuning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf49ac5d-345c-4a9a-9b4b-06db37eedd45 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models MMBench: Is Your Multi-modal Model an All-around Player?
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2659a164-5efb-4615-ab3b-0f86140a454e · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Med-HALT: Medical Domain Hallucination Test for Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d915b980-c5b4-4253-9080-01c451c0300a · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0c1a9fce-fe9c-484c-b358-6ea9838da3fb · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models MKRAG: Medical Knowledge Retrieval Augmented Generation for Medical Question Answering
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d59d1ffd-dc4a-4f67-9487-d3858e1caa42 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f0942674-01e6-4697-910d-a0735ac08bd0 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Aligning Large Multimodal Models with Factually Augmented RLHF
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b6e40cc-3132-49b5-9cdd-03338c580724 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Large Language Models for Data Annotation and Synthesis: A Survey
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c76bf2c-91e1-4a44-9d14-6747bbf1158f · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f046121-36bb-4d1f-b667-284807b560fc · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ef7a7b2b-55bd-437f-881a-d19cb8b48ede · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models ChiMed-GPT: A Chinese Medical Large Language Model with Full Training Regime and Better Alignment to Human Preferences
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d88da37-d0c8-454a-9ef4-82eb1dcab07f · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Safety challenges of AI in medicine in the era of large language models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b14a0c86-8da6-427c-bd95-33d68067aaf9 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 647a93f0-1ab9-4354-af3c-7d1b07101f26 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models LVLM-eHub: A Comprehensive Evaluation Benchmark for Large Vision-Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 007c30ff-fe4a-4b10-82ac-af765790fc9c · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9a21b54-6351-4374-9809-2a06294e0c8b · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27021c6f-b395-4f26-8783-31b444739187 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models A Survey of Large Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f90c3d7-d9e1-49c8-8113-7c6addc2a085 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b8b2766-b253-4acd-87cd-8976f0bdebc5 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models KG4Diagnosis: A Hierarchical Multi-Agent LLM Framework with Knowledge Graph Enhancement for Medical Diagnosis
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71201a5c-b555-44dc-a4b9-1a45bd895829 · outbound
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models Satisfactory Medical Consultation based on Terminology-Enhanced Information Retrieval and Emotional In-Context Learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7ad24faa-0347-4752-9904-db3e9fdf834a · inbound
KG4Diagnosis: A Hierarchical Multi-Agent LLM Framework with Knowledge Graph Enhancement for Medical Diagnosis MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd44847f-7dc3-4363-a261-9311cbbb39d0 · inbound
Loki's Dance of Illusions: A Comprehensive Survey of Hallucination in Large Language Models MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
Reference 200
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04ce2e7e-d737-4d09-b44d-e3a85a5e8824 · inbound
A comprehensive taxonomy of hallucinations in Large Language Models MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
Reference 109
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 038607a1-13e2-41aa-a690-ab5f2d7550b7 · inbound
TerraMAE: Learning Spatial-Spectral Representations from Hyperspectral Earth Observation Data via Adaptive Masked Autoencoders MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation baa9a7b1-57c2-4848-91ed-e9d0ce1365f3 · inbound
Trustworthy Medical Imaging with Large Language Models: A Study of Hallucinations Across Modalities MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 456e0b79-1c9f-4d9e-b161-5bfb1f2679aa · inbound
A Multi-Task Evaluation of LLMs' Processing of Academic Text Input MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f75f4bb8-a881-4007-bea9-ad58fdb4dc7d · inbound
PARALLAX: Separating Genuine Hallucination Detection from Benchmark Construction Artifacts MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1e92785f-c4af-4572-bda8-1688b7c71bfa · inbound
ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be9bd705-9c52-4a50-8ad7-ff9c51334350 · inbound
Hallucination in Medical Imaging AI: A Cross-Modality Analytical Framework for Taxonomy, Detection, and Mitigation under Regulatory Constraints MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation eff9e784-fec0-4680-a789-70fef4a19c29 · inbound
KnowHal: A Knowledge-Driven Benchmark for Comprehensive Multimodal Hallucination Evaluation MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.