Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T11:49:37.685430Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2509.02198.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T11:49:37.685430Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
40 of 40 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 109d96ab-84f0-4fc1-bcf1-e423f1aa045e · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Ev2R: Evaluating Evidence Retrieval in Automated Fact-Checking
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c55697eb-8024-4ea7-8d76-a9b4c5544369 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cf22791-b8f9-49a4-b5a3-b78f766f1075 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89acc67d-aaa3-4d0d-ba06-e4f43973457e · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain FactCHD: Benchmarking Fact-Conflicting Hallucination Detection
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68371e48-20fe-4e1a-a34d-326ac0395d43 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain FacTool: Factuality Detection in Generative AI -- A Tool Augmented Framework for Multi-Task and Multi-Domain Scenarios
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0c1ee87-9bf6-4424-a615-57d0512487d9 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b3a475c5-a31b-4fd7-9cc7-9137347f404c · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c036fda2-b86b-4aa4-84ea-06fbc66fa91c · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 602bd9f0-618c-4457-8fcf-f1944754698f · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain The Llama 3 Herd of Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9e258d2-00b9-4f0e-a265-e31f45e21cb2 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad3ac3ca-8043-47d5-b69e-d6dbfc28d21e · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d9e27792-0ba8-4f49-bff3-31d07d9d3f31 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d098079-9250-4727-9125-e7506cc82d56 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain DeBERTa: Decoding-enhanced BERT with Disentangled Attention
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78ebf89e-34ef-4c75-9d1a-48fc246e8c60 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f04cc43b-da8b-43fa-88e5-e06f5ea9d419 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bf0e5e9-c923-4f07-b974-3793a9662b2d · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Mistral 7B
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afe80fc9-18c5-4ca6-a73e-9cf387fe92aa · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Mixtral of Experts
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dfd6a38-2465-4980-9454-c172f52d4c03 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ed74b3f3-41cc-44f0-94cf-5c9356a862ed · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55be45d1-2543-4182-ba7a-7e656b0bf16b · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fbb1223-b585-4545-bcb6-5a0e0dc69c65 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Bennett, and Marti A
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a46d527-1d48-42a6-8e28-22273445200d · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation df2c9356-8962-4469-8670-2fed404fed8f · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c211d761-f6e4-43a7-bc2e-909d8eeef4fa · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6491c496-5aaa-4c8b-b457-cd3258f09b22 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01f73d1e-db95-4daf-87c9-5ba6db86c608 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0045d9d1-f571-43d0-9668-0d88b2029069 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4949615b-9cea-43ee-bfef-756ad0e042f8 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9a684d4c-e6a6-410d-9b40-27e0eed2c725 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain tasksource: A Dataset Harmonization Framework for Streamlined NLP Multi-Task Learning and Evaluation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b47bd57c-a7fd-4074-bdf3-3f2e47d79df7 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b16c0f2-4f28-4824-9fad-5818b7befcc4 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Gemma 2: Improving Open Language Models at a Practical Size
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c93e9956-2cb3-4c82-96c1-5e7664a3d549 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5dde900f-37ac-4974-952f-e7eefb06b8d7 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain A Stitch in Time Saves Nine: Detecting and Mitigating Hallucinations of LLMs by Validating Low-Confidence Generation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b92d26f-39a4-4eba-a3f7-10ad453697b9 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Aligning Large Language Models with Human: A Survey
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03e2cf4e-f5f7-4d2a-947a-a3bddc6ff438 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd1ec3dc-f4dc-4819-a6eb-07a5e3120fbf · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91931a1c-e096-456c-8196-40b97d5cd33e · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bb0d0631-73d4-433d-bde2-b62b47e6c5aa · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 36dbd05f-db23-4302-ba29-f83677e14729 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain online" 'onlinestring :=
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6635d21b-b61e-4919-87ec-5348f7cc4266 · outbound
FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain write newline
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.