Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:51:47.218026Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 1 inbound Pith citation observation for arXiv:2506.12538.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:51:47.218026Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:05:38.978358Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T18:05:39.131213Z
51 of 51 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1ebe1ed1-fe71-4da3-bdfd-4f1947a347ad · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f5ff4f8-941e-43e8-9bd9-5a3c22785194 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ed3e911e-6cd9-4d32-9fae-a0d4d30fb72e · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a7b0e061-c952-46fd-8fed-3e2ee8fadd8b · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation fecbfde6-64c8-4a5c-a74a-50763f2f8156 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4a8178e1-aa29-4acf-b8d6-6a48295682d9 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 03152f90-928d-4a02-92fb-20a726f0fc2a · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Qwen Technical Report
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4db6d88c-7dee-4e8a-8540-66d535483505 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking FactBench: A Dynamic Benchmark for In-the-Wild Language Model Factuality Evaluation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7f4c8b79-16f6-4a57-8b36-7ffde87eebe9 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Can LLMs Improve Multimodal Fact-Checking by Asking Relevant Questions?
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9effccd2-fcf4-47a0-b456-8353b6438c20 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking FacTool: Factuality Detection in Generative AI -- A Tool Augmented Framework for Multi-Task and Multi-Domain Scenarios
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01044271-72ea-43be-9808-46f3930ec56d · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8773e3da-a978-4d6e-809d-4078cd58bac6 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking 2025.Gemini 2.5 Flash Preview: Model Card
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ba5495fc-d1f6-404f-9f5f-d8aacd23fc5e · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af39b30c-8488-4bb9-b6d7-11ec16d841e2 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e762a968-8aab-41e9-8631-9aac2b895e7c · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Language Models as Knowledge Bases: On Entity Representations, Storage Capacity, and Paraphrased Queries
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation db4837fb-4eb5-4022-a794-11af203f19c0 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Do Large Language Models Know about Facts?
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84ea9e25-994c-45d7-9015-adac41299a73 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f610a3ac-c0fc-48d8-aac6-a4165239cd22 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 917d0631-5793-4d33-90d9-0b2db20f746f · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking HaluEval: A Large-Scale Hallucination Evaluation Benchmark for Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10f686e2-d920-4614-930e-dde8cc2174c3 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Large Language Model Agent for Fake News Detection
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee039324-2d8b-4905-8b73-38aa1edaf1a8 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 72406eee-512b-4b2c-aa2f-9e4a42c1586d · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking FACT-AUDIT: An Adaptive Multi-Agent Framework for Dynamic Fact-Checking Evaluation of Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4e661b3-2b1f-4456-b006-c615f13f0bd0 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2ca970c1-5790-4c33-b1c0-b29899c8f10a · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking DeepSeek-V3 Technical Report
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbff2460-3f35-417b-ba84-73776ce93e96 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking MMFakeBench: A Mixed-Source Multimodal Misinformation Detection Benchmark for LVLMs
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9d81fe3-d0d6-4517-b34d-bb411f23b1da · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 88d49459-49a6-49fd-8b71-fb02da16a45d · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f3c83d7d-0a54-40b5-baf3-226ead4a6ad8 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7e05e39f-e9df-4fee-bd6c-80051ade0dfe · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e6fed9c-f282-4d93-ba01-429ca1af880f · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 5903a929-8588-481a-ac8d-c18c4c1d95e6 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Fact-Checking Complex Claims with Program-Guided Reasoning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee2ae949-e5a1-4ad1-a9b5-41827b4881e4 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ea86b96-ac27-4a07-a49d-aa5c417f8cdb · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Exploring the Deceptive Power of LLM-Generated Fake News: A Study of Real-World Detection Challenges
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 642b8417-7f85-4486-99d0-fe5cbac9e395 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39fadc18-64c3-4d2c-927f-ba34ab1e0838 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Generative Large Language Models in Automated Fact-Checking: A Survey
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a92184fa-c109-4484-af52-e2d653f3f284 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 37945801-953c-4c50-92af-5f990f8ff8f5 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking MFC-Bench: Benchmarking Multimodal Fact-Checking with Large Vision-Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4431c997-95cc-48d9-ae55-51ed4a4ee657 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 89bd2d61-5881-4935-b719-bf0f7a1b9c9d · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1b197b68-f82b-41fa-9cc2-dac98002e09e · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1d85d0e5-e388-4c72-83fa-14526e3f3203 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking TrendFact: A Benchmark Towards Hotspot Perception in Automatic Fact-Checking
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 55b1da68-1329-4b2c-9e49-16c209071c30 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c2dcf01b-4e35-4276-95f5-0e034b3270a4 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Long-form factuality in large language models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18cbcaba-bdae-42d8-8590-20fcb4368dff · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking evidence
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e1b56da4-bb4f-4d8c-836d-13dca2e299f2 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 70b84f8e-3b88-40c4-9cc7-963e2c335519 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1f61ce84-a089-40df-a453-a816ff074c51 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking score":
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b8f259b8-967c-4317-8271-57f9938daa2a · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking FEVER: a large-scale dataset for Fact Extraction and VERification
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 224e85fe-de3e-4af1-85d4-b18479babf27 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4fd1cae4-7959-454d-ab18-00f0f4dc4a2a · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 60f0bdc5-daa9-4dd0-bb8a-309591832c35 · outbound
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Unresolved cited work
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1b0e443b-8ea8-47b9-9062-b65c83b5b731 · inbound
RAMA: Retrieval-Augmented Multi-Agent Framework for Misinformation Detection in Multimodal Fact-Checking RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.