Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T17:20:47.964154Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2411.12701.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T17:20:47.964154Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
45 of 45 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f1eb6100-59c3-4b36-98b8-2cc0e7f6c2af · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Understanding intermediate layers using linear classifier probes
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f178567c-0611-4702-a76d-4077d70ba11e · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Eliciting Latent Predictions from Transformers with the Tuned Lens
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27c905c8-6ad6-47fb-bc92-6c226359ecde · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 637fa55d-1d2e-45ba-8087-bb5981de2d11 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53106c5f-ed7d-47c4-8f45-8f0b456594a3 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7e84de85-9ca3-4042-b9c6-24a429937849 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Lookback Lens: Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3d0e54c-e0b5-4894-b46f-9b10b8d31218 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c36ca47f-55d3-468f-98c9-1f6a47776e4f · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aca82476-07ea-4f10-bc94-08849207c958 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Triggerless Backdoor Attack for NLP Tasks with Clean Labels
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fadf73e0-49b4-453e-8f93-68d04a6cc602 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d89cbbe-6302-4101-b536-5a4952ea89be · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 186b9804-67c8-4b61-ac49-354c8b15a793 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 325ab4dd-55eb-461d-9f05-2c73828076b5 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Weight Poisoning Attacks on Pre-trained Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 102dd777-73dd-40ee-abd7-3faa44deb3e8 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a6dc7e3-981d-44e7-8413-5d21954d5860 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1d46b3f7-bf71-44a1-8d75-00b08198cff8 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 581408bd-5971-44b6-9da1-418806bb18b6 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 267bf8d8-1047-4a5f-89be-98078480cae2 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dd48b56c-81b7-4730-a59f-60255ab08680 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 32301e36-bc1c-41bd-89f0-00daf1912226 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b347b450-3b4f-4704-a495-d75e18890df3 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations WT5?! Training Text-to-Text Models to Explain their Predictions
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48aaefce-1dec-4c5c-a9bc-06121786601c · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9ccc6ebd-8851-44d4-a026-ced087df5e60 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations GPT-4o System Card
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5ce55ae-c121-4750-91e1-6592de525067 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fc59da8b-0f4d-4237-95c6-d73502ddc931 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Mind the Style of Text! Adversarial and Backdoor Attacks Based on Text Style Transfer
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84750a0f-f502-48d2-995a-8129036479c5 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bfe9e137-fcb2-4fae-9a75-dd4f03e51e0d · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Explain Yourself! Leveraging Language Models for Commonsense Reasoning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cabb687a-784b-4f96-b0a6-53419eabc550 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c6b94e1e-7282-434b-aa34-c34e4c7b6d8e · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Manning, Andrew Ng, and Christopher Potts
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 416d7128-de65-417d-8066-dffb55ce40ab · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 270d12a4-f5bf-4bad-9d1a-44d7d41ea3db · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Large Language Models Can be Lazy Learners: Analyze Shortcuts in In-Context Learning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d758a62-24bc-4551-a0c3-1eb27c043fd2 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b37e0629-3f99-435c-9ab2-518a2d8f542a · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85933582-f49c-4345-a080-874ae0a43d74 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations LLaMA: Open and Efficient Foundation Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2179665d-6bd6-4ce5-8ef2-53b70a7e7ddc · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Concealed Data Poisoning Attacks on NLP Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0282666-d307-4c75-906f-0a5beda1308a · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63f11b28-45ac-4144-844b-1c0a5b001817 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Trojan Activation Attack: Red-Teaming Large Language Models using Activation Steering for Safety-Alignment
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4d7a69f-6c88-4de7-b71a-c5d8bd329d9a · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Usable XAI: 10 Strategies Towards Exploiting Explainability in the LLM Era
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 191b7ac9-7b76-4796-97dd-639d1dc7d929 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6eac4be4-6f81-44cd-a736-92449211f721 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations BITE: Textual Backdoor Attacks with Iterative Trigger Injection
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a5e70c6-d9b6-48fa-aceb-4e392b373cde · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations The Unreliability of Explanations in Few-shot Prompting for Textual Reasoning
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae37f722-f861-4f1c-8a3f-6666fd1f5e07 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4aa54054-1c99-4f75-ae2c-2baaaf0a278e · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68f934ec-9df8-4eb0-96a6-1bab8c4dac12 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations online" 'onlinestring :=
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a609f6b-27fc-40a2-b100-184d6934f132 · outbound
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations write newline
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.