Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T17:31:59.951364Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 1 inbound Pith citation observation for arXiv:2502.05911.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T17:31:59.951364Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-18T01:23:01.921132Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T01:25:34.979210Z
59 of 59 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 31cd75f1-2fa6-4cda-93b0-b13659aa9c99 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation RAT: Adversarial Attacks on Deep Reinforcement Agents for Targeted Behaviors
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b4b5c76b-2b33-428f-baf1-ef36367a6c38 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Efficient Model-agnostic Alignment via Bayesian Persuasion
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97cbe8de-681f-4aee-8d13-936e02b17731 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 18e2ba56-6341-44c8-b3d0-589e66b514d9 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Safety-Tuned LLaMAs: Lessons From Improving the Safety of Large Language Models that Follow Instructions
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4ac8346-95d8-49b3-be49-5d588b7dcc2e · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Teaching Large Language Models to Express Knowledge Boundary from Their Own Signals
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3359a1b-a55b-4c0e-9335-cf2814405e51 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Can AI Assistants Know What They Don't Know?
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac8b7d7f-9149-49d4-9e40-d617761dbb7f · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c26d678-0727-4566-88da-af77e191c3f6 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Knowledge Neurons in Pretrained Transformers
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3420928a-3a50-460c-b3a7-1425a68ea300 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e144b74d-ab38-48a2-b0b8-32305bcabc2c · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation The Llama 3 Herd of Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1b452a8-a336-425d-a42a-a9cbfff375d0 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Don't Hallucinate, Abstain: Identifying LLM Knowledge Gaps via Multi-LLM Collaboration
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56032f7f-3a9f-4bf5-8f60-5a264f06f9f6 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Successor Heads: Recurring, Interpretable Attention Heads In The Wild
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8282045c-f268-41fc-8588-fc6203cc46c3 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Measuring Massive Multitask Language Understanding
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a7a9c2d-3f2f-4731-af89-5c472a7fa147 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation LoRA: Low-Rank Adaptation of Large Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abf18524-daa3-4e47-a935-8b6e08c0640e · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Lora: Low-rank adaptation of large language models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8f0cb3b-5383-4bae-a944-53268445c686 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 637c8f11-9eae-419c-80e1-cf83d3a797a0 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation TC-RAG:Turing-Complete RAG's Case study on Medical LLM Systems
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd7b17a7-b4e3-43e7-9c07-18ca3bd834b9 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation HyKGE: A Hypothesis Knowledge Graph Enhanced Framework for Accurate and Reliable Medical LLMs Responses
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fea0eba-5208-4f40-bbe4-ade36c48fcad · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Johnson and Joram Lindenstrauss
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7886acf6-e0f5-4cd5-a930-50db5c6a6460 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e847c900-f1ff-4b67-adb6-c42e79d94777 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unfamiliar Finetuning Examples Control How Language Models Hallucinate
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d274182-0347-4f98-a148-6fc5013bfcfb · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Scaling Laws for Neural Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a48ee927-a83c-4218-8823-ba6bcaf698fb · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation GRAD-MATCH: Gradient Matching based Data Subset Selection for Efficient Deep Model Training
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74d034ec-774d-4d25-876f-f0999b45b9ea · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f69519e-fed4-4bc7-b23a-0895dcdfcb86 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation A Survey on the Honesty of Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ef876ad-e0c4-4a39-a5bd-2aebf392653c · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d0e44cc3-ed0f-4761-8386-0274115767e1 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Kuaiji: the First Chinese Accounting Large Language Model
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 25960b7c-3f1e-4010-b469-52d0a7dfbc28 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 397fd5d1-786b-4a54-bb40-3aa001d159e3 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation GPT-4 Technical Report
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c1cc5c3-ed39-405e-9f49-0550080fb395 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Training language models to follow instructions with human feedback
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cc7f3b43-1bf4-409a-9bdb-0e9295edce31 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7750d40b-369a-4ba5-b095-73d5dae974dc · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f8865962-6189-4098-a9de-b200789aee2a · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Identifying Semantic Induction Heads to Understand In-Context Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4d1cefa-0a74-4185-9aad-9a37d29062d8 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Learning or Self-aligning? Rethinking Instruction Fine-tuning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65f3cb4c-1b59-4f7e-b1bc-d7bbeb5920f2 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Learning Dynamics of LLM Finetuning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7066f076-15fe-4743-a4dc-7c2835b23a56 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfe3773f-abb7-4a73-9c47-d90c3ee23680 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54c2efaf-0d3d-4cb5-985e-80c3d9a9097b · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation GPTVoiceTasker: Advancing Multi-step Mobile Task Efficiency Through Dynamic Interface Exploration and Learning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aafbbff8-f7fc-4085-b8bb-09e499b67b8c · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Knowledge Verification to Nip Hallucination in the Bud
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f22eb664-662e-4ebd-97df-30ff5b35a66c · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Uncertainty Aware Learning for Language Model Alignment
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a11969d5-99cc-4bb6-bc41-ced495ba74b8 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Know your limits: A survey of abstention in large language models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7203753f-3149-4d15-bb44-1ec0e10bbe41 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Do Llamas Work in English? On the Latent Language of Multilingual Transformers
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f22ed801-3d0a-4800-ab1f-77523a0e600e · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11ccd4aa-1365-4999-b946-c001122c4f81 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation LESS: Selecting Influential Data for Targeted Instruction Tuning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7664ad69-dbd1-4a3c-b4f9-f1cd953e8518 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a99bae8a-575c-4940-ac76-c709c416c0ae · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc5bfc07-519f-404d-9dbc-fabd1755de94 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Parenting: Optimizing Knowledge Selection of Retrieval-Augmented Language Models with Parameter Decoupling and Tailored Tuning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 12886857-c64a-46bf-bb63-0f7df68c5661 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation SmallToLarge (S2L): Scalable Data Selection for Fine-tuning Large Language Models by Summarizing Training Trajectories of Small Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27cadcd8-87f4-4fdb-8557-e5daf82216dd · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Alignment for Honesty
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93b4e325-2e03-4e26-87a0-b2cd9d336945 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation xFinder: Large Language Models as Automated Evaluators for Reliable Evaluation
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 873a2a21-b63a-4e57-9382-109f4d6f1500 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Neuron-Level Knowledge Attribution in Large Language Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67d847bb-b7db-4638-bd13-7edf1ba7dc52 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f993a29d-7435-4afa-8b80-77269f7066e6 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 447ebd98-c33a-494a-9478-a25c9a8450c5 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab1fb5ae-d453-4cbb-b173-f9e0cd0da1c6 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Dataset Condensation with Gradient Matching
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27e868bd-7e2a-4925-ad71-684eed03445d · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 407f7128-254f-49bd-9c51-ffdb39be292e · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Utilize the Flow before Stepping into the Same River Twice: Certainty Represented Knowledge Flow for Refusal-Aware Instruction Tuning
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c254be1d-af1e-49de-b343-610620c637ca · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation online" 'onlinestring :=
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 902c6faa-59bb-49be-b97e-97d82bd793b9 · outbound
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation write newline
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e66bb30-e92b-406d-84a9-65e7d7589ac1 · inbound
Understanding New-Knowledge-Induced Factual Hallucinations in LLMs: Analysis and Interpretation GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.