Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 80 inbound Pith citation observations for arXiv:2310.06387.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T18:53:15.305158Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
21
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 117d23d0-55a9-4849-8078-a4f74db8fd9c · inbound
Jailbreak Attacks and Defenses Against Large Language Models: A Survey Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9ca19ab0-3dea-4262-820a-f10d6891b276 · inbound
FlexLLM: Exploring LLM Customization for Moving Target Defense on Black-Box LLMs Against Jailbreak Attacks Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c95fb98b-469c-4d12-a165-66b54dc5fc35 · inbound
Look Before You Leap: Enhancing Attention and Vigilance Regarding Harmful Content with GuidelineLLM Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fedc0ca-9974-49fb-8a0f-8024c1a93efe · inbound
No Free Lunch for Defending Against Prefilling Attack by In-Context Learning Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d24f0f6a-2382-49ba-b029-bfaf1f3e4515 · inbound
Towards Responsible Governing AI Proliferation Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cfd5c5e-551a-4de1-8066-5a1b12d2d127 · inbound
Token Highlighter: Inspecting and Mitigating Jailbreak Prompts for Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cc1274f-7c2e-475b-99c0-af3b7e4c34ac · inbound
LLM-Virus: Evolutionary Jailbreak Attack on Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ef6edf3-7144-427a-8a56-67f480a69451 · inbound
SaLoRA: Safety-Alignment Preserved Low-Rank Adaptation Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 962b564d-6ac9-41c1-8db9-b5cc36a03762 · inbound
Auto-RT: Automatic Jailbreak Strategy Exploration for Red-Teaming Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6ef1e05-0fd2-4d45-bd4d-a8d6458c5f61 · inbound
Safeguarding Large Language Models in Real-time with Tunable Safety-Performance Trade-offs Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d605eed-3151-4592-a9e4-07be3a451d65 · inbound
Layer-Level Self-Exposure and Patch: Affirmative Token Mitigation for Jailbreak Attack Defense Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed581a41-2e19-4c3a-93a0-50d01a279b19 · inbound
Self-Instruct Few-Shot Jailbreaking: Decompose the Attack into Pattern and Behavior Learning Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19cf3e0d-9de7-48fe-b988-f87b47517819 · inbound
Latent-space adversarial training with post-aware calibration for defending large language models against jailbreak attacks Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb412c3f-1d20-49c5-8853-766be59699f9 · inbound
Episodic memory in AI agents poses risks that should be studied and mitigated Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0175743d-88fe-4eb6-9851-69008c436cdf · inbound
PromptShield: Deployable Detection for Prompt Injection Attacks Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f26dd78-4c35-4556-8f79-98f7892de47d · inbound
Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a928bc4c-1dd2-43bd-862a-186b3966b4c8 · inbound
PANDAS: Improving Many-shot Jailbreaking via Positive Affirmation, Negative Demonstration, and Adaptive Sampling Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e184b199-3148-4634-bc49-02b558a134dc · inbound
JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2a488f1-08c4-43a6-a2a4-94b88f78a080 · inbound
MetaSC: Test-Time Safety Specification Optimization for Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cba2e2e6-8940-425b-aecc-e698ad06b693 · inbound
Advancing LLM Safe Alignment with Safety Representation Ranking Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bdd8198-b371-4ea6-b84c-bdc93100f652 · inbound
Scalable Defense against In-the-wild Jailbreaking Attacks with Safety Context Retrieval Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 380184e7-86ed-4d1d-acad-bd528028193c · inbound
Implicit Jailbreak Attacks via Cross-Modal Information Concealment on Vision-Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f79f5c5-dadf-42b5-871f-460fbc8e94b7 · inbound
Secure LLM Fine-Tuning via Safety-Aware Probing Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 50dc41cf-1536-4e7b-869a-64ad6d1d0615 · inbound
Exploring the Vulnerability of the Content Moderation Guardrail in Large Language Models via Intent Manipulation Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 315c024d-59f5-4a55-8fe7-56370521a7b3 · inbound
Benign-to-Toxic Jailbreaking: Inducing Harmful Responses from Harmless Prompts Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73d2de4a-8bd0-44b0-b6e4-7b1dec2610c8 · inbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5faeabe1-ec16-46fd-8fbe-ec8a76ffd52e · inbound
Bootstrapping LLM Robustness for VLM Safety via Reducing the Pretraining Modality Gap Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75792518-7d97-4fe7-917b-fcd454ccace3 · inbound
ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 60f67dd0-5532-4871-9215-078dcc65b23a · inbound
Adversarial Attacks on Robotic Vision Language Action Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abd6921c-f91f-4c0e-b1af-47bf8baa381a · inbound
TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de60991d-b09f-4bc5-815e-e386c2095b22 · inbound
Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b255213c-9252-42fc-8015-5501bcc45f5a · inbound
LLMs Caught in the Crossfire: Malware Requests and Jailbreak Challenges Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bfa0aee-8356-4b43-a9f4-82adba2f8a86 · inbound
The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 133
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55f36b96-93fa-4d92-8919-e1cc47c74088 · inbound
InfoFlood: Jailbreaking Large Language Models with Information Overload Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acd11a48-ed3f-4c4d-be94-32473105ddd4 · inbound
SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8564b34-057e-407c-bf81-6ed1135155b4 · inbound
Linearly Decoding Refused Knowledge in Aligned Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9e12c66-942d-4f13-b196-686c4bf8756d · inbound
Defending Against Prompt Injection With a Few DefensiveTokens Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78c43e03-bdbb-458f-83aa-5413134e33b9 · inbound
Circumventing Safety Alignment in Large Language Models Through Embedding Space Toxicity Attenuation Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf2e2b86-141e-493b-a58e-0e75f00be4e4 · inbound
Innocence in the Crossfire: Roles of Skip Connections in Jailbreaking Visual Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0d19bb1-5cea-4a7d-928b-271ecf66ead2 · inbound
MOCHA: Are Code Language Models Robust Against Multi-Turn Malicious Coding Prompts? Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 218856ef-8017-4313-b7d8-3c4288aa9d6f · inbound
PUZZLED: Jailbreaking LLMs through Word-Based Puzzles Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b53d180c-a509-4bc2-8d06-8efb4e8c110b · inbound
Automatic LLM Red Teaming Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 1685
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11d2042c-6acd-4c2f-8ba3-21a2845c9a02 · inbound
A Real-Time, Self-Tuning Moderator Framework for Adversarial Prompt Detection Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bbdb906-2757-4b09-abf5-0001b09d1ebc · inbound
A Survey on Training-free Alignment of Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e071d78-c297-4a80-8084-43d8cac74c4f · inbound
Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 103a5d3e-866c-42b7-a0b0-5c77d922d988 · inbound
SafeLLM: Unlearning Harmful Outputs from Large Language Models against Jailbreak Attacks Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe642ff8-1d94-4870-a75b-bc2df70e9cc0 · inbound
On Surjectivity of Neural Networks: Can you elicit any behavior from your model? Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 921a1417-8384-499b-97e9-4bdae34b8f08 · inbound
GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 262a77df-7593-46d2-b169-d8fe64bae7d3 · inbound
Baichuan-M2: Scaling Medical Capability with Large Verifier System Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcf06fb2-f84c-486c-9e38-4c9f4c0ca189 · inbound
A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 264
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 881ed3e0-8875-48cf-a485-a295d796bf27 · inbound
MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb991391-33fa-4d40-b1c4-1646f578c6a8 · inbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 200
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b9af65f-eef9-4e39-b148-2c7c4b2f4287 · inbound
SAID: Safety-Aware Intent Defense via Prefix Probing for Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5187e3f8-ddbc-49ef-b93f-23f7013e2313 · inbound
Measuring the Security of Mobile LLM Agents under Adversarial Prompts from Untrusted Third-Party Channels Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e241d81d-697b-4658-86dc-2532dc2b8cd0 · inbound
ASTRA: An Automated Framework for Strategy Discovery, Retrieval, and Evolution for Jailbreaking LLMs Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1c861905-1bde-4f80-b222-a54f7756b3e8 · inbound
GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b6802b60-e140-435f-8723-813ea5ddc511 · inbound
ContextCov: Deriving and Enforcing Executable Constraints from Agent Instruction Files Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b6f4ff22-d27a-4aa8-b762-1b9e0680ce30 · inbound
TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3d58642f-7d0f-43f0-9c67-768094314bbb · inbound
GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 25c6cce4-a38e-4317-b3ab-66456ab89088 · inbound
TEMPLATEFUZZ: Fine-Grained Chat Template Fuzzing for Jailbreaking and Red Teaming LLMs Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7b10e68e-4c58-4538-949f-52ec49ce67bb · inbound
A Synonymous Variational Perspective on the Rate-Distortion-Perception Tradeoff Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2448a1f-80b8-4bbe-b1be-faeaf618d404 · inbound
Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 75fea871-0b2c-455e-b68a-69d3a924bc36 · inbound
A Systematic Study of Training-Free Methods for Trustworthy Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c0804e6e-5974-4d59-9fc7-fcc348e25e30 · inbound
Jailbreaking Large Language Models with Morality Attacks Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 654cd810-3283-4979-be35-e00290ea9fe6 · inbound
SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c698fb39-f359-4ccf-bc76-66c16fef678f · inbound
Automation-Exploit: A Multi-Agent LLM Framework for Adaptive Offensive Security with Digital Twin-Based Risk-Mitigated Exploitation Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 63e4218e-7582-4fa1-9f5b-30946ea95025 · inbound
A Systematic Survey of Security Threats and Defenses in LLM-Based AI Agents: A Layered Attack Surface Framework Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5d7afb9e-d275-447f-9f23-c01aedbff02b · inbound
Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 598a05e0-a41c-42a9-858b-69653353566e · inbound
Latent Personality Alignment: Improving Harmlessness Without Mentioning Harms Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 833c9c08-b805-4c2d-828c-8bd4f095fd6c · inbound
PQR: A Framework to Generate Diverse and Realistic User Queries that Elicit QA Agent Failures Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c766b4e8-3c80-4acc-acc4-12cc8b8205e3 · inbound
Jailbreak susceptibility prediction and mitigation via the behavioral geometry of models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 440422c9-2f00-4359-a058-80339a4cc457 · inbound
THRD: A Training-Free Multi-Turn Defense Framework for Jailbreak Attacks on Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0806e1bd-bc88-4890-a644-04a84df5052a · inbound
SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7ad5939a-73f8-40b2-993c-5fe2b2a0181f · inbound
A Layered Security Framework Against Prompt Injection in RAG-Based Chatbots Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 89595de1-b9ad-4080-a147-1ed5fd276557 · inbound
Investigating The Security of Modern AI and Cloud Infrastructure Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 137
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e938f8b3-7d40-4f1f-9857-d41e30c67f32 · inbound
ToxiREX: A Dataset on Toxic REasoning in ConteXt Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 223
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ccf75546-7565-452c-b4bb-2286973270ba · inbound
Position: Preventing AI-Generated CSAM Necessitates New Approaches to AI Safety Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 104
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 121b3da8-53a6-49bf-a5ae-cf67e5e61b74 · inbound
Mitigating Taint-Style Vulnerabilities in MCP Servers via Security-Aware Tool Descriptions Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dde8f715-429a-4a7b-a486-671e594e1b93 · inbound
How Jailbreak Attacks Inform Safety Alignment: A Defender-Centric, Shapley-Based Evaluation of Jailbreak Contributions Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a179e9a-bb3b-445a-b683-f7a1f9e23786 · inbound
When LLM Defenses Backfire: Characterizing Safety, Performance, and Cost Trade-offs Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.