Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T23:52:40.912038Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 9 of 9 outbound references and 1 inbound Pith citation observation for arXiv:2602.12418.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T23:52:40.912038Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-30T20:47:40.813521Z
A source-named dated measurement, never combined with another source.
Source: cited_works
9 of 9 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 582e6c86-985f-4869-aa50-b70660a12d98 · outbound
Sparse Autoencoders are Capable LLM Jailbreak Mitigators Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acfae772-0729-4c05-a39e-eeaf0f77960c · outbound
Sparse Autoencoders are Capable LLM Jailbreak Mitigators rights around free speech and freedom of assembly
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1aaf292f-8844-46e5-8042-bc758f4b19ea · outbound
Sparse Autoencoders are Capable LLM Jailbreak Mitigators Intriguing properties of neural networks
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7df0063e-b2b4-458e-96d4-16550bff531e · outbound
Sparse Autoencoders are Capable LLM Jailbreak Mitigators Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27cefc0c-5eb9-40a4-90cb-8cacecfa42e3 · outbound
Sparse Autoencoders are Capable LLM Jailbreak Mitigators — "You should be a responsible Language Model and should not generate harmful or misleading content! Please answer the following user query in a responsible way. {}
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 191d38ee-d039-4b79-aec7-4800c10e19da · outbound
Sparse Autoencoders are Capable LLM Jailbreak Mitigators AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c76d597-7b76-4d26-ae18-5bad78f5d46c · outbound
Sparse Autoencoders are Capable LLM Jailbreak Mitigators RobustBench: a standardized adversarial robustness benchmark
Reference 844
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51025924-d1e2-4e81-bb09-9efbfae16cef · outbound
Sparse Autoencoders are Capable LLM Jailbreak Mitigators Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a677b74-7079-44f6-a909-3ad05bc53d35 · outbound
Sparse Autoencoders are Capable LLM Jailbreak Mitigators Steering Large Language Model Activations in Sparse Spaces
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f51b1d39-94d5-439d-874f-899be739798d · inbound
Do LLMs Know Their Vulnerable Scenarios? Sparse Autoencoders are Capable LLM Jailbreak Mitigators
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.