Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T21:47:17.085295Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 6 inbound Pith citation observations for arXiv:2412.04141.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T21:47:17.085295Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T19:53:26.084902Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T16:53:40.545243Z
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8b1d9f05-3a87-4c2a-93b3-85e12f0e9bac · outbound
Reducing Tool Hallucination via Reliability Alignment GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dd6fba7-7f00-49fa-9c4a-fe47e01ecb83 · outbound
Reducing Tool Hallucination via Reliability Alignment The Internal State of an LLM Knows When It's Lying
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 601c09ed-71a7-47ae-8048-9ae9aea9196a · outbound
Reducing Tool Hallucination via Reliability Alignment Large Language Models as Tool Makers
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d03f298a-26a4-4588-bc43-6b2c18c64fd2 · outbound
Reducing Tool Hallucination via Reliability Alignment AlpaGasus: Training A Better Alpaca with Fewer Data
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0afa8f4-f3d4-40e8-a2ed-e27880181b6d · outbound
Reducing Tool Hallucination via Reliability Alignment T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cc2ba72-73c6-40a8-a58b-bfa7f366c7ef · outbound
Reducing Tool Hallucination via Reliability Alignment The Llama 3 Herd of Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2082df96-eea7-47d3-96df-3daef9635b9f · outbound
Reducing Tool Hallucination via Reliability Alignment Retrieval-Augmented Generation for Large Language Models: A Survey
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec8361ee-2b0d-4ac6-bca9-63dcbf0b97f2 · outbound
Reducing Tool Hallucination via Reliability Alignment Gemini: A Family of Highly Capable Multimodal Models , 2023
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b2c20101-c8f8-47ab-8512-082cb6e7b28a · outbound
Reducing Tool Hallucination via Reliability Alignment M., Alves, D
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 42254d59-7427-4d35-8a4b-b170e10061f9 · outbound
Reducing Tool Hallucination via Reliability Alignment StableToolBench: Towards Stable Large-Scale Benchmarking on Tool Learning of Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e95b84db-9567-4dd0-813f-b50ba7648480 · outbound
Reducing Tool Hallucination via Reliability Alignment and Kembhavi, A
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation aca268e8-e79a-4e1b-8916-2bf739095be8 · outbound
Reducing Tool Hallucination via Reliability Alignment ToolkenGPT: Augmenting Frozen Language Models with Massive Tools via Tool Embeddings
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a709d7dc-2dac-4bb6-92e2-e96cd6b5f06d · outbound
Reducing Tool Hallucination via Reliability Alignment Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f73952d-bf56-485f-abb2-87369d3abaa8 · outbound
Reducing Tool Hallucination via Reliability Alignment Tool Documentation Enables Zero-Shot Tool-Usage with Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 883f2f3f-6a0e-49f7-8172-146583f1ea1a · outbound
Reducing Tool Hallucination via Reliability Alignment On Mitigating Code LLM Hallucinations with API Documentation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93c6a6fd-83bb-4887-b858-9585e6126a0d · outbound
Reducing Tool Hallucination via Reliability Alignment J., Madotto, A., and Fung, P
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8376b122-32f5-4d73-986b-8f6df5d279b0 · outbound
Reducing Tool Hallucination via Reliability Alignment GeneGPT: augmenting large language models with domain tools for improved access to biomedical information
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 547486c5-8e80-4616-ac2c-f9a91c7ecdcb · outbound
Reducing Tool Hallucination via Reliability Alignment Crafting papers on machine learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e21cd200-3579-4915-9d7b-07eff3b8b5b7 · outbound
Reducing Tool Hallucination via Reliability Alignment FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed12755a-886b-458a-b006-0db34eb03de8 · outbound
Reducing Tool Hallucination via Reliability Alignment Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f7b2f93-87cf-489e-acc2-c791a3f64201 · outbound
Reducing Tool Hallucination via Reliability Alignment Gorilla: Large Language Model Connected with Massive APIs
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8834674d-56b9-417d-aa1f-190df9c96539 · outbound
Reducing Tool Hallucination via Reliability Alignment The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 312c704c-7bc3-4c3c-8c2f-60c334d3bb2e · outbound
Reducing Tool Hallucination via Reliability Alignment WebCPM: Interactive Web Search for Chinese Long-form Question Answering
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91bb3e31-7ff4-41f2-aa5c-372922666d3d · outbound
Reducing Tool Hallucination via Reliability Alignment ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs , 2023 b
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 845f9018-31ce-4e61-85ea-c165325176c7 · outbound
Reducing Tool Hallucination via Reliability Alignment Tool Learning with Foundation Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7308864-700f-439f-b489-0efc2dd0042b · outbound
Reducing Tool Hallucination via Reliability Alignment D., Ermon, S., and Finn, C
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8969e862-97c1-4edf-b80e-cd7fc7e00b66 · outbound
Reducing Tool Hallucination via Reliability Alignment Toolformer: Language Models Can Teach Themselves to Use Tools
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7142ddb2-c30f-4148-9ebc-d4004ebd596a · outbound
Reducing Tool Hallucination via Reliability Alignment Trusting Your Evidence: Hallucinate Less with Context-aware Decoding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75b9f10c-2e33-4a2e-aa39-a931779fcda8 · outbound
Reducing Tool Hallucination via Reliability Alignment ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases , 2023
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a0914f47-78bd-42f1-8807-3b652bc1b636 · outbound
Reducing Tool Hallucination via Reliability Alignment Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 485c4e8a-9211-4c21-bd0e-4ba147db3a32 · outbound
Reducing Tool Hallucination via Reliability Alignment A Stitch in Time Saves Nine: Detecting and Mitigating Hallucinations of LLMs by Validating Low-Confidence Generation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ee8a57e-5458-4eea-ac76-330775e4aa7b · outbound
Reducing Tool Hallucination via Reliability Alignment Can LLMs Express Their Uncertainty? An Empirical Evaluation of Confidence Elicitation in LLMs
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2b5f1c6-d734-4c3a-837e-7738d504c989 · outbound
Reducing Tool Hallucination via Reliability Alignment Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47d9299b-f252-4d96-9ac6-4598e20d4fe4 · outbound
Reducing Tool Hallucination via Reliability Alignment Alignment for Efficient Tool Calling of Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f95929d-6260-409c-a06d-a22f4e9c41b8 · outbound
Reducing Tool Hallucination via Reliability Alignment Qwen2.5 Technical Report
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9e33d0c-17f9-4403-a547-e7af18eefd77 · outbound
Reducing Tool Hallucination via Reliability Alignment Give Us the Facts: Enhancing Large Language Models with Knowledge Graphs for Fact-aware Language Modeling
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 660c7d82-e4e1-4892-beb7-f3a2a94941fc · outbound
Reducing Tool Hallucination via Reliability Alignment WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 94888ca9-52c9-4ee2-b4a7-f5f3c3aabfd3 · outbound
Reducing Tool Hallucination via Reliability Alignment StepTool: Enhancing Multi-Step Tool Usage in LLMs via Step-Grained Reinforcement Learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2c53a25-da9f-4e2e-8afe-802eb59bcd5a · outbound
Reducing Tool Hallucination via Reliability Alignment Instruction tuning for large language models: A survey
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a162cb5d-ab0e-473a-9ff9-4827dec9ee5e · outbound
Reducing Tool Hallucination via Reliability Alignment ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0139cb11-cd07-4bfa-9736-35407c23b7d1 · outbound
Reducing Tool Hallucination via Reliability Alignment Verify-and-Edit: A Knowledge-Enhanced Chain-of-Thought Framework
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1e0a239-501e-4450-9804-0fdc96898cbf · outbound
Reducing Tool Hallucination via Reliability Alignment Enhancing llm reliability via explicit knowledge boundary modeling
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cba75c2f-91f0-4a24-af84-0c48f34332a2 · outbound
Reducing Tool Hallucination via Reliability Alignment LIMA: Less Is More for Alignment
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56fd37ba-2f00-48d8-b6e5-a51d06708945 · outbound
Reducing Tool Hallucination via Reliability Alignment ToolQA: A Dataset for LLM Question Answering with External Tools
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3baddb05-2c0a-4ee4-8e31-6d2488696285 · inbound
Enhancing Tool Learning in Large Language Models with Hierarchical Error Checklists Reducing Tool Hallucination via Reliability Alignment
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62d9a328-53ab-4567-bcc8-a755495ae583 · inbound
The Reasoning Trap: How Enhancing LLM Reasoning Amplifies Tool Hallucination Reducing Tool Hallucination via Reliability Alignment
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8d103b0d-6cc0-4c65-b181-8adff8013200 · inbound
Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments Reducing Tool Hallucination via Reliability Alignment
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7983dae6-b61d-4e63-859f-e7c643a02d78 · inbound
Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making Reducing Tool Hallucination via Reliability Alignment
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1028e53-6f80-4a14-af3b-5d0589a0e7ac · inbound
PredAct-Bench: Benchmarking Tool-Augmented Dialogue under Controlled Tool Noise Reducing Tool Hallucination via Reliability Alignment
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2972291e-fd12-47ea-83e2-d0fd5007ea01 · inbound
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning Reducing Tool Hallucination via Reliability Alignment
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.