Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 39 inbound Pith citation observations for arXiv:2308.09662.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:41:10.294985Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
16
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation cde53581-e803-4c34-a549-6c02d378077d · inbound
Jailbreak Attacks and Defenses Against Large Language Models: A Survey Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 62c434ea-a5db-40fe-9784-2908dbadbd76 · inbound
AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dae4e495-d378-4606-8300-f0e20fe5a30b · inbound
LLM-Safety Evaluations Lack Robustness Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6c695795-7c4c-4210-9327-852985dffe04 · inbound
Phonetic Perturbations Reveal Tokenizer-Rooted Safety Gaps in LLMs Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e7cfc789-826b-4fe4-8f1e-0e0814606f5e · inbound
Exploring Jailbreak Attacks on LLMs through Intent Concealment and Diversion Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 634fa79b-81ca-4dbe-9899-d1b732efd882 · inbound
MTSA: Multi-turn Safety Alignment for LLMs through Multi-round Red-teaming Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46f8839c-2343-4faa-a81c-bff59ebf220b · inbound
Mitigating Deceptive Alignment via Self-Monitoring Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf7d8278-25e5-430d-8210-761a9ce1c0b0 · inbound
Security Concerns for Large Language Models: A Survey Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d45dbd4-4bca-4e1b-ba73-570d344eedba · inbound
Token-level Accept or Reject: A Micro Alignment Approach for Large Language Models Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be60ffa7-ae5f-4129-a599-cc034230db8d · inbound
Risk-aware Direct Preference Optimization under Nested Risk Measure Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbcb2b21-ad46-4d62-a163-76dcfe37b233 · inbound
LLM Agents Should Employ Security Principles Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78c1836b-0f38-4111-aa6c-981b6d6ae984 · inbound
Adversarial Preference Learning for Robust LLM Alignment Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ece555ab-f3a1-4a6a-89d4-87c87c5497b5 · inbound
LLMs Caught in the Crossfire: Malware Requests and Jailbreak Challenges Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73959c27-4331-4ad2-ac15-f280e44a1f53 · inbound
The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2c388da-08b8-437f-9576-e73b9fee28e6 · inbound
Hatevolution: What Static Benchmarks Don't Tell Us Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e1182d2-c42d-48c9-8d81-b7028e4a16c1 · inbound
PL-Guard: Benchmarking Language Model Safety for Polish Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fe0b9fa-78e1-43e1-ac5d-954235d63b18 · inbound
MGC: A Compiler Framework Exploiting Compositional Blindness in Aligned LLMs for Malware Generation Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dca1a89b-0b75-4bdf-9f69-555eb9834939 · inbound
CAVGAN: Unifying Jailbreak and Defense of LLMs via Generative Adversarial Attacks on their Internal Representations Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48420487-1b02-4ebc-8055-a103043207ed · inbound
Internal Value Alignment in Large Language Models through Controlled Value Vector Activation Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc893af8-465b-4303-8687-e6e0658b8567 · inbound
Guiding LLM Decision-Making with Fairness Reward Models Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e31d8f53-d82f-4ae8-8d80-108c0eabed2d · inbound
SATORI: Static Test Oracle Generation for REST APIs Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b6e49a8-3215-48d8-8cea-84d2b7846095 · inbound
Learning to Conceal Risk: Controllable Multi-turn Red Teaming for LLMs in the Financial Domain Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bae77025-96af-4599-a740-cfa343fa042e · inbound
Evalet: Evaluating Large Language Models through Functional Fragmentation Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d4b5a365-d421-46b2-b351-7573e17cf87c · inbound
When Search Goes Wrong: Red-Teaming Web-Augmented Large Language Models Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e41f6a8a-e9f9-4042-9641-5bcc6b2df7a1 · inbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c6a59d3-cb9a-4631-be2b-e88e2783613e · inbound
ToxSearch: Evolving Prompts for Toxicity Search in Large Language Models Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e1d4702-9bb7-43e7-bb08-5e519c59d389 · inbound
When Personalization Legitimizes Risks: Uncovering Safety Vulnerabilities in Personalized Dialogue Agents Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ea431d66-5646-44b9-8f16-ae81d0743128 · inbound
Diversifying Toxicity Search in Large Language Models Through Speciation Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8e386a28-aaa2-4d51-9187-1f729c03705b · inbound
MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 65f6e4cf-e496-4a14-98f8-2f9057019a7c · inbound
Misrouter: Exploiting Routing Mechanisms for Input-Only Attacks on Mixture-of-Experts LLMs Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c7778730-89ca-470d-b42a-0ffd009ed8b2 · inbound
You Snooze, You Lose: Automatic Safety Alignment Restoration through Neural Weight Translation Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6d3603b4-cad3-4497-ae07-151bd157bf72 · inbound
Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6991985d-2de9-49f6-bd3c-a95fd28a1d48 · inbound
Dialectics of Alignment: Harnessing Unsafe Knowledge for Dynamic Safety Routing Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 25061803-f922-4013-9920-04f853853a18 · inbound
SafeSteer: Localized On-Policy Distillation for Efficient Safety Alignment Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2dba2dfc-6332-40b8-9f6a-2f3f983906c9 · inbound
Defending Jailbreak Attacks on Large Language Models via Manifold Trajectory Kinetics Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 587eacad-b34c-4eed-ae8e-424a3e580995 · inbound
Distributed Quality-Diversity Search for Toxicity in Large Language Models Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5aa005cc-1aac-45f5-b27f-407bd964ef24 · inbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8e4ed37-5070-4455-b867-09c47f93ea5b · inbound
No Single Neuron of Failure: Distributed Safety Alignment Against White-Box Attacks Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51ee87da-f0b5-42b3-9169-f8cfd2021453 · inbound
Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.