Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T10:29:50.456781Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2502.08109.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T10:29:50.456781Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:24:26.408324Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T10:24:26.761572Z
30 of 30 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7439b2b1-eb43-4d43-b941-d7752a5faf27 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Sparks of artificial general intelligence: Early experiments with gpt-4, 2023
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4974afd8-de96-4fae-a460-284eac60156d · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Felm: Benchmarking factuality evaluation of large language models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6766fac3-796a-474a-9100-9d65997f3927 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Truthfulqa: Measuring how models mimic human falsehoods
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e709f5d9-619d-4e4b-be0d-c5110381e3e4 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Qafacteval: Improved qa-based factual con- sistency evaluation for summarization
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 87d031f6-8b72-4c58-9bb5-751c51546132 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Attention is all you need
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 48b2cba7-265e-4d83-a368-941320f703b6 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses A survey of large language models, 2023
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4668f260-6b78-4ed9-8321-edffffb89546 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Emergent abilities of large language models, 2022
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 943a29de-779a-4149-9704-60f500daa3b8 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d1f9aa0e-f755-4d53-ad0a-06d7e9025df8 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Fingpt: Open-source financial large language models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 38eca8a1-e12d-446a-9fdf-1435dd75c8a9 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Survey of hallucination in natural language generation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0d1ca2b3-eb19-490e-b90f-5eaa0a7e4b45 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses On faithfulness and factuality in abstractive summarization
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 50615555-ffad-46b7-9ed0-ea969511caeb · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Neural path hunter: Reducing hallucination in dialogue systems via path grounding
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9d6bad46-5ab8-4a8b-a442-acc9d7df7227 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses The factual inconsistency problem in abstractive text summarization: A survey, 2023
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation fe5f218e-b98f-4f00-9024-9448422b47b3 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions, 2023
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cc4c15f-ff02-4c90-9d99-2b1805d396cf · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Realtime qa: what’s the answer right now? In Proceedings of the 37th International Conference on Neural Information Processing Systems
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ecbf9e88-a255-48e1-bba2-f9216d1b18bb · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bc74f69b-b7f2-4435-b276-d3fd2d69434e · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Overcoming a theoretical limitation of self-attention, 2022
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c9561055-1de7-4d2d-9d58-3634982c2832 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Batgpt: A bidirectional autoregessive talker from generative pre-trained transformer, 2023
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ccde37e0-afaf-472a-9076-2e07aa69e894 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses A multitask, multilingual, multimodal evaluation of chatgpt on reasoning, hallucination, and interactivity
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ea860bc0-8d40-47c9-8206-f869debb2322 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Judging llm-as-a-judge with mt- bench and chatbot arena
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 74002951-328b-4d46-ba47-9351ae72bc6f · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Benchmarking foundation models with language-model-as-an- examiner, 2023
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 04d179d8-10e8-4eea-9d7e-331af6059cae · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses G-eval: Nlg evaluation using gpt-4 with better human alignment
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 639c456c-c1bc-42e9-ab26-14bf3da7109a · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Prd: Peer rank and discussion improve large language model based evaluations, 2023
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 75ac7803-b933-4e24-bad9-c41a17305c18 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Halueval: A large-scale hallucination evaluation benchmark for large language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 00323eb2-230c-4577-9c7e-ec8d29a86b16 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Factchd: Benchmarking fact-conflicting hallucination detection, 2024
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4e5a8544-d032-46e5-ac8b-297d9d1bc22a · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Faithdial: A faithful benchmark for information-seeking dialogue
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6c4352c5-a1fe-4f84-9f9a-a00a1751c7d4 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Evaluating attribution in dialogue systems: The begin benchmark
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 82dfc91a-115a-4819-849f-6777901491ee · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses The llama 3 herd of models, 2024
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a2915ead-6b6c-4331-80f3-5ad278e06a65 · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Lora: Low-rank adaptation of large language models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2daa342e-62ae-4126-beb9-2b9e0e261a7c · outbound
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses Gpt-4 technical report, 2024
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f0a5476d-6ef2-4063-9cb9-57abc06527db · inbound
Interpretation Meets Safety: A Survey on Interpretation Methods and Tools for Improving LLM Safety HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.