Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 37 inbound Pith citation observations for arXiv:2403.04783.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:10:31.093405Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-09T10:26:11.076263Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 2dcc7f64-d806-45fe-8dd4-bcce51bdd1fc · inbound
Jailbreak Attacks and Defenses Against Large Language Models: A Survey AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 110
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a0d05ba6-b761-453e-b545-d5a28031f29f · inbound
Preventing Jailbreak Prompts as Malicious Tools for Cybercriminals: A Cyber Defense Perspective AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d13057ea-efe0-42db-9eed-43781c6f70b3 · inbound
Boundless Socratic Learning with Language Games AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 098a5ccf-bdb8-4385-a41f-96c0fc9ef065 · inbound
Look Before You Leap: Enhancing Attention and Vigilance Regarding Harmful Content with GuidelineLLM AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f76712e0-a05a-4230-92d8-9c76d534f4a9 · inbound
Latent-space adversarial training with post-aware calibration for defending large language models against jailbreak attacks AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77d6c430-4514-4969-9bfa-f424bdae5aa0 · inbound
Large Language Model Agent: A Survey on Methodology, Applications and Challenges AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 188
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e0e1ef83-954b-4e3d-baf3-6fe8801ae300 · inbound
DETAM: Defending LLMs Against Jailbreak Attacks via Targeted Attention Modification AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9f61786-26dd-48ea-83e1-d3002b0515cc · inbound
T2VShield: Model-Agnostic Jailbreak Defense for Text-to-Video Models AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c252686f-dd82-4de3-9377-8ba484070689 · inbound
Attack and defense techniques in large language models: A survey and new perspectives AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3dfc389-8db7-4198-a553-00d97d2002e2 · inbound
LLM Security: Vulnerabilities, Attacks, Defenses, and Countermeasures AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 165
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e12043b5-dbbc-450a-a216-c17e72b6f728 · inbound
Three Minds, One Legend: Jailbreak Large Reasoning Model with Adaptive Stacked Ciphers AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35c4c82a-a13b-483a-8835-afb609c51de4 · inbound
Revisiting Multi-Agent Debate as Test-Time Scaling: A Systematic Study of Conditional Effectiveness AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25943294-4e36-4a9f-9906-151cb48001dd · inbound
A Red Teaming Roadmap Towards System-Level Safety AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 105
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20c6fe58-3282-48ed-ad79-6b3d4083e531 · inbound
The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55fb3ae3-8746-46bf-a422-91928754ac8c · inbound
SoK: The Privacy Paradox of Large Language Models: Advancements, Privacy Risks, and Mitigation AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 138
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96fd0665-3084-4a7e-82af-daa38dcc1b51 · inbound
SecurityLingua: Efficient Defense of LLM Jailbreak Attacks via Security-Aware Prompt Compression AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 661181d5-bd27-468d-bc24-180d16b4f5f6 · inbound
Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e17440da-2c71-4720-b4c8-554849d03e38 · inbound
SV-LLM: An Agentic Approach for SoC Security Verification using Large Language Models AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f41a7bbc-2b48-4cb1-b593-8e1176907018 · inbound
Evaluating Multi-Agent Defences Against Jailbreaking Attacks on Large Language Models AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7041c1e2-d6ea-4233-86d3-70ca0cd1eaec · inbound
MIND: A Multi-agent Framework for Zero-shot Harmful Meme Detection AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c5a84dd-3a68-45a3-b162-74353007e0d2 · inbound
Multi-Actor Generative Artificial Intelligence as a Game Engine AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 331f0a6e-6396-4d20-a2e5-9d131f820e8d · inbound
ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f41039d3-9d9d-411b-99f0-e8eda572e6bf · inbound
A Real-Time, Self-Tuning Moderator Framework for Adversarial Prompt Detection AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a78fe27-3580-4bdf-95e6-6f4d72bcf790 · inbound
Evaluating the Robustness of Retrieval-Augmented Generation to Adversarial Evidence in the Health Domain AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c31e2c17-d4e9-45b0-98b1-b8c18bbedec8 · inbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 228
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e636f8af-8f02-4e4f-b995-b87b5480e453 · inbound
Sentra-Guard: A Real-Time Multilingual Defense Against Adversarial LLM Prompts AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ff1a4df8-4f67-4ee6-a596-00bcdeae6c8b · inbound
From Evidence to Verdict: An Agent-Based Forensic Framework for AI-Generated Image Detection AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 191d38ee-d039-4b79-aec7-4800c10e19da · inbound
Sparse Autoencoders are Capable LLM Jailbreak Mitigators AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33a197dd-4980-4503-ace5-b3db170b185e · inbound
GAMMAF: A Common Framework for Graph-Based Anomaly Monitoring Benchmarking in LLM Multi-Agent Systems AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e7e31ae5-7b21-4ec2-b77b-92152c3bf1b1 · inbound
SoK: Robustness in Large Language Models against Jailbreak Attacks AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7d4cad6a-e56d-413b-b317-5a8a43f7c4f9 · inbound
Jailbreak susceptibility prediction and mitigation via the behavioral geometry of models AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8334a1b7-c402-49a2-8f98-bc7ff2b7b8b5 · inbound
Cognitive Firewall: A Proactive, Zero-Trust, Multi-Gate Framework for LLM Safety AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 820b58bb-20bb-41c9-82a6-f5c7487c9d40 · inbound
Mitigating Taint-Style Vulnerabilities in MCP Servers via Security-Aware Tool Descriptions AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation df871845-a3d4-4bf0-a2d3-dd431ad15d7c · inbound
SafeFlow: Semantic Information-Flow Control for Blocking Malicious Propagation in Multi-Agent Systems AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60b68910-3e3b-4683-ad27-bd0fb2648e82 · inbound
Adversarial Attacks in Multi-Agent LLM Pipelines: Unveiling Structural Vulnerabilities in Agentic AI Architectures AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 335b71bd-8f04-476c-afdb-45d55c71fb06 · inbound
When Collaboration Becomes a Trigger: Collective Evidence-Threshold Backdoors in Multi-Agent Systems AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db108649-8481-473c-8799-452690224ebe · inbound
On Understanding, Identifying, and Mitigating Vulnerabilities in Agentic Large Language Models AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Reference 154
Source-reported events for the cited work
Unavailable: canonical work link unavailable.