Pith. sign in

Paper Citation Record · LEDGER

Reducing Tool Hallucination via Reliability Alignment

As of 20 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 6 inbound Pith citation observations for arXiv:2412.04141.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.04141 v3

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T21:47:17.085295Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T19:53:26.084902Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T16:53:40.545243Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8b1d9f05-3a87-4c2a-93b3-85e12f0e9bac · outbound

This paper cites GPT-4 Technical Report.

Reducing Tool Hallucination via Reliability Alignment GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.881457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.881457Z digest=sha256:5d0027802f5c774c5fda5ef3b8412580c191213fe5354882663034cbd8b16e79

Observation 1dd6fba7-7f00-49fa-9c4a-fe47e01ecb83 · outbound

This paper cites The Internal State of an LLM Knows When It's Lying.

Reducing Tool Hallucination via Reliability Alignment The Internal State of an LLM Knows When It's Lying

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.886934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.886934Z digest=sha256:9ea58bf0fbb407b30c02424a653f9a268ce9acf86e669c4b18b6b8f7470c73d0

Observation 601c09ed-71a7-47ae-8048-9ae9aea9196a · outbound

This paper cites Large Language Models as Tool Makers.

Reducing Tool Hallucination via Reliability Alignment Large Language Models as Tool Makers

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T21:47:17.903937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T21:47:16.892067Z digest=sha256:b4c9c6ed9317018d31bde2170793bb11a135b9b271796d8226ee6e0ae02d449c

Observation d03f298a-26a4-4588-bc43-6b2c18c64fd2 · outbound

This paper cites AlpaGasus: Training A Better Alpaca with Fewer Data.

Reducing Tool Hallucination via Reliability Alignment AlpaGasus: Training A Better Alpaca with Fewer Data

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.897213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.897213Z digest=sha256:76ebbaaacfb95ca33867011af58b3206f371be747af98b636b0eca22e6099ff6

Observation e0afa8f4-f3d4-40e8-a2ed-e27880181b6d · outbound

This paper cites T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step.

Reducing Tool Hallucination via Reliability Alignment T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.902502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.902502Z digest=sha256:9f9d8ec1baef21447e87d3b9ba015684f156fda34de4d81d4d8cca207734db8b

Observation 2cc2ba72-73c6-40a8-a58b-bfa7f366c7ef · outbound

This paper cites The Llama 3 Herd of Models.

Reducing Tool Hallucination via Reliability Alignment The Llama 3 Herd of Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.907887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.907887Z digest=sha256:99c4775110dd07c6d6521a9d76773a235349ad297d55e20abb81f61f6910cdb0

Observation 2082df96-eea7-47d3-96df-3daef9635b9f · outbound

This paper cites Retrieval-Augmented Generation for Large Language Models: A Survey.

Reducing Tool Hallucination via Reliability Alignment Retrieval-Augmented Generation for Large Language Models: A Survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.913515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.913515Z digest=sha256:bd60942e0dc7b853e18c5d88a43b89656b16dc7ecd7f062e9f678c5429637ffd

Observation ec8361ee-2b0d-4ac6-bca9-63dcbf0b97f2 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models , 2023.

Reducing Tool Hallucination via Reliability Alignment Gemini: A Family of Highly Capable Multimodal Models , 2023

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T21:47:17.889426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T21:47:16.918273Z digest=sha256:3d341c1a88cc1a9724700be01adaa4410da6a067ea28e5730989d1d94307fc2d

Observation b2c20101-c8f8-47ab-8512-082cb6e7b28a · outbound

This paper cites M., Alves, D.

Reducing Tool Hallucination via Reliability Alignment M., Alves, D

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T21:47:17.875105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T21:47:16.923066Z digest=sha256:f554d00f26c5b1f405269cf65e9dbaae50805a75e5eb3a271c201fb5b4361ffa

Observation 42254d59-7427-4d35-8a4b-b170e10061f9 · outbound

This paper cites StableToolBench: Towards Stable Large-Scale Benchmarking on Tool Learning of Large Language Models.

Reducing Tool Hallucination via Reliability Alignment StableToolBench: Towards Stable Large-Scale Benchmarking on Tool Learning of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.927708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.927708Z digest=sha256:4270ea4610c8f0a84228beb5bd766bd0910b7ee17dfc73ab986e4e15d61ee0fc

Observation e95b84db-9567-4dd0-813f-b50ba7648480 · outbound

This paper cites and Kembhavi, A.

Reducing Tool Hallucination via Reliability Alignment and Kembhavi, A

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T21:47:17.860965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T21:47:16.932692Z digest=sha256:7ff6f2f70d34faae05270a8a6beaeeab54afa89c55d6804c03c43a0ab08387b3

Observation aca268e8-e79a-4e1b-8916-2bf739095be8 · outbound

This paper cites ToolkenGPT: Augmenting Frozen Language Models with Massive Tools via Tool Embeddings.

Reducing Tool Hallucination via Reliability Alignment ToolkenGPT: Augmenting Frozen Language Models with Massive Tools via Tool Embeddings

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.937036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.937036Z digest=sha256:a82e3a3b1a570a439890ee346c77a835394f83fb2a6ab3c8f8fd13080c637c54

Observation a709d7dc-2dac-4bb6-92e2-e96cd6b5f06d · outbound

This paper cites an unresolved cited work.

Reducing Tool Hallucination via Reliability Alignment Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.941533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.941533Z digest=sha256:1588e1390f40cbaf898d3a87c1caa7e52b523d152f2111cd485e318cd46c03d1

Observation 8f73952d-bf56-485f-abb2-87369d3abaa8 · outbound

This paper cites Tool Documentation Enables Zero-Shot Tool-Usage with Large Language Models.

Reducing Tool Hallucination via Reliability Alignment Tool Documentation Enables Zero-Shot Tool-Usage with Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.945817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.945817Z digest=sha256:024b0ad6d383b93c7111feec9622f543e40fd64022685bbc4f942f53278c23e4

Observation 883f2f3f-6a0e-49f7-8172-146583f1ea1a · outbound

This paper cites On Mitigating Code LLM Hallucinations with API Documentation.

Reducing Tool Hallucination via Reliability Alignment On Mitigating Code LLM Hallucinations with API Documentation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.950405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.950405Z digest=sha256:ab7a6540c9015d0a5148cad352e810ac5870c0edb523675ade219cdccf240409

Observation 93c6a6fd-83bb-4887-b858-9585e6126a0d · outbound

This paper cites J., Madotto, A., and Fung, P.

Reducing Tool Hallucination via Reliability Alignment J., Madotto, A., and Fung, P

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.954966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.954966Z digest=sha256:4d6859bb00ba75b6a8167c26f89a33af3ebb2a079535889cc78ba5e5126317bd

Observation 8376b122-32f5-4d73-986b-8f6df5d279b0 · outbound

This paper cites GeneGPT: augmenting large language models with domain tools for improved access to biomedical information.

Reducing Tool Hallucination via Reliability Alignment GeneGPT: augmenting large language models with domain tools for improved access to biomedical information

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.959156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.959156Z digest=sha256:50b342aaf69dfb29227441f193092ac39c73bea8dab97406888893cc9a753935

Observation 547486c5-8e80-4616-ac2c-f9a91c7ecdcb · outbound

This paper cites Crafting papers on machine learning.

Reducing Tool Hallucination via Reliability Alignment Crafting papers on machine learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T21:47:17.828369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T21:47:16.963725Z digest=sha256:2c56e504a3862ede23c34f973c49c3f4ba6c8c26333331e447588c83d8d8591b

Observation e21cd200-3579-4915-9d7b-07eff3b8b5b7 · outbound

This paper cites FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation.

Reducing Tool Hallucination via Reliability Alignment FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.968271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.968271Z digest=sha256:1f2f79656564393550b4eed5500fdda3bcfd1759125ca83820bc5d67480590ad

Observation ed12755a-886b-458a-b006-0db34eb03de8 · outbound

This paper cites Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation.

Reducing Tool Hallucination via Reliability Alignment Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.972775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.972775Z digest=sha256:b0ca450fa47f39fe7869a3fc7d99cd57c471274fcb75c2511956cb99d4a61759

Observation 2f7b2f93-87cf-489e-acc2-c791a3f64201 · outbound

This paper cites Gorilla: Large Language Model Connected with Massive APIs.

Reducing Tool Hallucination via Reliability Alignment Gorilla: Large Language Model Connected with Massive APIs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.977314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.977314Z digest=sha256:ee8bcebb02348dee114d93df707f512ec2adf55328121c6af5d14f556d49256f

Observation 8834674d-56b9-417d-aa1f-190df9c96539 · outbound

This paper cites The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only.

Reducing Tool Hallucination via Reliability Alignment The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.982221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.982221Z digest=sha256:c0ab93d9deea24c39695d8a08541ff24f01f53a7980c2485bf0f521d1e42d473

Observation 312c704c-7bc3-4c3c-8c2f-60c334d3bb2e · outbound

This paper cites WebCPM: Interactive Web Search for Chinese Long-form Question Answering.

Reducing Tool Hallucination via Reliability Alignment WebCPM: Interactive Web Search for Chinese Long-form Question Answering

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.987237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.987237Z digest=sha256:9e41403afc34b56997df6e7390d151d99b021ec88d4ecd15f6119cb40abcf88e

Observation 91bb3e31-7ff4-41f2-aa5c-372922666d3d · outbound

This paper cites ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs , 2023 b.

Reducing Tool Hallucination via Reliability Alignment ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs , 2023 b

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T21:47:17.813437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T21:47:16.992285Z digest=sha256:b911b9fdcb724be09c283245d5d3955875904b4cbf677ee62a650603d432dc0f

Observation 845f9018-31ce-4e61-85ea-c165325176c7 · outbound

This paper cites Tool Learning with Foundation Models.

Reducing Tool Hallucination via Reliability Alignment Tool Learning with Foundation Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:16.996711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:16.996711Z digest=sha256:b3f05988a699d8e53fd358e5f914c3c21372256d6b2ec796bb09eaab1a2eea09

Observation d7308864-700f-439f-b489-0efc2dd0042b · outbound

This paper cites D., Ermon, S., and Finn, C.

Reducing Tool Hallucination via Reliability Alignment D., Ermon, S., and Finn, C

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.001677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.001677Z digest=sha256:14be8ba3469ed6205196e35e85df88d65f996da78c85fce64b224cc131df6640

Observation 8969e862-97c1-4edf-b80e-cd7fc7e00b66 · outbound

This paper cites Toolformer: Language Models Can Teach Themselves to Use Tools.

Reducing Tool Hallucination via Reliability Alignment Toolformer: Language Models Can Teach Themselves to Use Tools

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.006563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.006563Z digest=sha256:4c21e4086ff9ddcb80e04385552e91b3d0d5b482e79e32cddfeb1b06b498e90b

Observation 7142ddb2-c30f-4148-9ebc-d4004ebd596a · outbound

This paper cites Trusting Your Evidence: Hallucinate Less with Context-aware Decoding.

Reducing Tool Hallucination via Reliability Alignment Trusting Your Evidence: Hallucinate Less with Context-aware Decoding

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.011819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.011819Z digest=sha256:d33df5a366156a4e6d89b594ff18d575d1cb3331c9e94ecdcc302c5b7d259820

Observation 75b9f10c-2e33-4a2e-aa39-a931779fcda8 · outbound

This paper cites ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases , 2023.

Reducing Tool Hallucination via Reliability Alignment ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases , 2023

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T21:47:17.789095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T21:47:17.016619Z digest=sha256:1a67554955d7ec355f56e8337d6ee9c7a3d06fe14c423be6f5702fe36e0f86fe

Observation a0914f47-78bd-42f1-8807-3b652bc1b636 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Reducing Tool Hallucination via Reliability Alignment Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.021017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.021017Z digest=sha256:0a1871d71fd5bf59775b01d02deef8a67e323805cf7691349cba6c0ba95ddd76

Observation 485c4e8a-9211-4c21-bd0e-4ba147db3a32 · outbound

This paper cites A Stitch in Time Saves Nine: Detecting and Mitigating Hallucinations of LLMs by Validating Low-Confidence Generation.

Reducing Tool Hallucination via Reliability Alignment A Stitch in Time Saves Nine: Detecting and Mitigating Hallucinations of LLMs by Validating Low-Confidence Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.025661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.025661Z digest=sha256:03fb868fbae60ce25dadca1e3212977a9b18d3c9ac9f0ca0d8863aca6e9bcc33

Observation 2ee8a57e-5458-4eea-ac76-330775e4aa7b · outbound

This paper cites Can LLMs Express Their Uncertainty? An Empirical Evaluation of Confidence Elicitation in LLMs.

Reducing Tool Hallucination via Reliability Alignment Can LLMs Express Their Uncertainty? An Empirical Evaluation of Confidence Elicitation in LLMs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.030332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.030332Z digest=sha256:d5f7a3bd5c560247cf2e969f273de242ccc2665534bac4afa4dcfc5ed3df8133

Observation e2b5f1c6-d734-4c3a-837e-7738d504c989 · outbound

This paper cites Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback.

Reducing Tool Hallucination via Reliability Alignment Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.035095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.035095Z digest=sha256:50da2fde8557fe163a94c2c345b04dbebaa0ddbc125b1c67952a9caff8a05eb4

Observation 47d9299b-f252-4d96-9ac6-4598e20d4fe4 · outbound

This paper cites Alignment for Efficient Tool Calling of Large Language Models.

Reducing Tool Hallucination via Reliability Alignment Alignment for Efficient Tool Calling of Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.039661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.039661Z digest=sha256:b48a8a542ae4ba24acca40fce5e3dd7246832a1cc1a51d3d7d93834a3541dc55

Observation 2f95929d-6260-409c-a06d-a22f4e9c41b8 · outbound

This paper cites Qwen2.5 Technical Report.

Reducing Tool Hallucination via Reliability Alignment Qwen2.5 Technical Report

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.044367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.044367Z digest=sha256:5e8473642bcc45b0d51b15443e59b07cbd0cee5481b8cea98aec8335a31b3caf

Observation f9e33d0c-17f9-4403-a547-e7af18eefd77 · outbound

This paper cites Give Us the Facts: Enhancing Large Language Models with Knowledge Graphs for Fact-aware Language Modeling.

Reducing Tool Hallucination via Reliability Alignment Give Us the Facts: Enhancing Large Language Models with Knowledge Graphs for Fact-aware Language Modeling

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.048885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.048885Z digest=sha256:6ff1c113aa90318e4a2e88944af57811db6c7f9c2903016cca2277ce0f6d05e3

Observation 660c7d82-e4e1-4892-beb7-f3a2a94941fc · outbound

This paper cites WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents.

Reducing Tool Hallucination via Reliability Alignment WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T21:47:17.773493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T21:47:17.053419Z digest=sha256:30f286da3a21b8248ce3ead507b163894e237ce3fa7711be6117230de9f60ed6

Observation 94888ca9-52c9-4ee2-b4a7-f5f3c3aabfd3 · outbound

This paper cites StepTool: Enhancing Multi-Step Tool Usage in LLMs via Step-Grained Reinforcement Learning.

Reducing Tool Hallucination via Reliability Alignment StepTool: Enhancing Multi-Step Tool Usage in LLMs via Step-Grained Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.057484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.057484Z digest=sha256:f5aa8435a88d3603b24c7b95620e174b10b910436865f063bcdfbf638359d800

Observation f2c53a25-da9f-4e2e-8afe-802eb59bcd5a · outbound

This paper cites Instruction tuning for large language models: A survey.

Reducing Tool Hallucination via Reliability Alignment Instruction tuning for large language models: A survey

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.062005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.062005Z digest=sha256:481e1ca22bca2c87b13380a2d318ef41224339d430e694c6a1dce0ef00fcdb91

Observation a162cb5d-ab0e-473a-9ff9-4827dec9ee5e · outbound

This paper cites ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models.

Reducing Tool Hallucination via Reliability Alignment ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.066252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.066252Z digest=sha256:f36f12c8bca486e652336f3a86748036aff8e7fce78d202625a06c8eacb1d723

Observation 0139cb11-cd07-4bfa-9736-35407c23b7d1 · outbound

This paper cites Verify-and-Edit: A Knowledge-Enhanced Chain-of-Thought Framework.

Reducing Tool Hallucination via Reliability Alignment Verify-and-Edit: A Knowledge-Enhanced Chain-of-Thought Framework

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.070929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.070929Z digest=sha256:d4ee0fffc490d8282f23fecb2bb2647818c3c82139c066eda3f2a31cc0327526

Observation f1e0a239-501e-4450-9804-0fdc96898cbf · outbound

This paper cites Enhancing llm reliability via explicit knowledge boundary modeling.

Reducing Tool Hallucination via Reliability Alignment Enhancing llm reliability via explicit knowledge boundary modeling

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.076073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.076073Z digest=sha256:d43998e52c6af7d0fbd5e6861c4163bb13e18ebbea2c82010a048bfdb646033b

Observation cba75c2f-91f0-4a24-af84-0c48f34332a2 · outbound

This paper cites LIMA: Less Is More for Alignment.

Reducing Tool Hallucination via Reliability Alignment LIMA: Less Is More for Alignment

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.080503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.080503Z digest=sha256:1f3aa589470fd6c27030eb26331ccc24de55352f91f70471de43bdb532c47652

Observation 56fd37ba-2f00-48d8-b6e5-a51d06708945 · outbound

This paper cites ToolQA: A Dataset for LLM Question Answering with External Tools.

Reducing Tool Hallucination via Reliability Alignment ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.085295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.085295Z digest=sha256:51085aac4cbd92e63815288764d29523fde435bf58729edc00bd580f0f598333

Pith citing papers

Observation 3baddb05-2c0a-4ee4-8e31-6d2488696285 · inbound

Enhancing Tool Learning in Large Language Models with Hierarchical Error Checklists cites this paper.

Enhancing Tool Learning in Large Language Models with Hierarchical Error Checklists Reducing Tool Hallucination via Reliability Alignment

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:21:43.349220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:21:43.349220Z digest=sha256:e298df176ff39d5e9de6c45772cc150eed932a42cbe4be7c20448f4eeb1021f5

Observation 62d9a328-53ab-4567-bcc8-a755495ae583 · inbound

The Reasoning Trap: How Enhancing LLM Reasoning Amplifies Tool Hallucination cites this paper.

The Reasoning Trap: How Enhancing LLM Reasoning Amplifies Tool Hallucination Reducing Tool Hallucination via Reliability Alignment

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:10:51.316095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T04:09:50.183494Z digest=sha256:f145dedf1068d15df80be09a7ab0ed9f01cd61d1615d144326fea533485d109a

Observation 8d103b0d-6cc0-4c65-b181-8adff8013200 · inbound

Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments cites this paper.

Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments Reducing Tool Hallucination via Reliability Alignment

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-06-29T16:53:40.546847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T16:51:36.524194Z digest=sha256:a9becd70b746bf4564380ee497d84cf5ef8f5ee08356a493327c09d186fd7d06

Observation 7983dae6-b61d-4e63-859f-e7c643a02d78 · inbound

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making cites this paper.

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making Reducing Tool Hallucination via Reliability Alignment

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T02:42:07.036102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:42:07.036102Z digest=sha256:bb64542541ccf14a8583b6686a24e606c71d5fee59c2fd7a40a7a47b6e286369

Observation a1028e53-6f80-4a14-af3b-5d0589a0e7ac · inbound

PredAct-Bench: Benchmarking Tool-Augmented Dialogue under Controlled Tool Noise cites this paper.

PredAct-Bench: Benchmarking Tool-Augmented Dialogue under Controlled Tool Noise Reducing Tool Hallucination via Reliability Alignment

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T08:32:48.011938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:32:48.011938Z digest=sha256:8a699cea36a23959de53cdf17b649bd22690a79839355e38bb5446922638ebb3

Observation 2972291e-fd12-47ea-83e2-d0fd5007ea01 · inbound

AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning cites this paper.

AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning Reducing Tool Hallucination via Reliability Alignment

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T19:53:26.084902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T19:53:26.084902Z digest=sha256:90e94488009b36a523e9f7e70c749b8192d486f21debc36780e0be0923d6bc3a