Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T01:16:15.489997Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 0 inbound Pith citation observations for arXiv:2608.00134.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T01:16:15.489997Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
66 of 66 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 798a21b2-7217-4fa7-b796-04582796ca4b · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e9400a6-d005-41c6-add6-9971e971b577 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Gemini: A Family of Highly Capable Multimodal Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 197cda93-1f6e-451b-ba6b-589f4480435f · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks LLaMA: Open and Efficient Foundation Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1afd53bc-d838-4a72-8bd4-2cb3df258454 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks A Survey of Large Language Models in Medicine: Progress, Application, and Challenge
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 155f7a77-09a2-4a6b-b4de-b421612b25a3 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Exploring Recommendation Capabilities of GPT-4V(ision): A Preliminary Case Study
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1e15201-e534-458f-aad3-c61eaa37143d · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks {LLM-Fuzzer}: Scaling assessment of large language model jailbreaks,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb0cadd4-0fc1-4fa6-ba87-396222896a5e · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Exploiting the Index Gradients for Optimization-Based Jailbreaking on Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab6b3cef-02ed-44c9-aa4b-6c404eb1e212 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Making them ask and answer: Jailbreaking large language models in few queries via disguise and reconstruction,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 986c3a0b-3af2-4f0f-9beb-c9e6bdff0a39 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Helping big language models protect themselves: An enhanced filtering and summarization system,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 404c6bc3-e694-4d71-8157-9785c1c560ff · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42419e43-111b-46b5-987d-646f2b11d314 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37abdc89-91e0-4d96-b872-5b08253372c3 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Datasentinel: A game-theoretic detection of prompt injection attacks,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9fddcb5-449b-45cc-b026-56190ae5a93b · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61e51d5b-99a8-4d04-bdca-02bb060bf374 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Fight back against jailbreaking via prompt adversarial tuning,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b501988-70d3-46de-a021-4fa32eacd20a · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Safety-Tuned LLaMAs: Lessons From Improving the Safety of Large Language Models that Follow Instructions
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a4e43e2-0e5d-4de7-bdcb-23599c9c7c44 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Distributional preference learning: Understanding and accounting for hidden context in rlhf,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5232ba9-5422-459d-ba88-940486fce466 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Attacking Large Language Models with Projected Gradient Descent
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa11cee1-adf0-4aaf-9489-059349dd9982 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f7201d3-4551-4f8f-a646-d595f553792f · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a36c6875-97c0-47a9-8bd8-25e0b0f97ecd · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a37eead2-fe61-4568-bb8d-e8cebb9cacac · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15e36a56-ffb3-45ed-9880-b7a8ba38d7b5 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b286444e-f098-4bb3-affc-415bf51b6227 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Honeytrap: Deceiving large language model attackers to honeypot traps with resilient multi-agent defense,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75924873-2eeb-426e-8578-a3e07ec5c661 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Great, Now Write an Article About That: The Crescendo Multi-Turn LLM Jailbreak Attack
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2c38328-d87e-4d06-bd39-e9b7c027bb6d · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Tree of attacks: Jailbreaking black-box llms automatically,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 698e5121-2960-4aa4-a4cb-2b58e6633457 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks PAPILLON: Efficient and Stealthy Fuzz Testing-Powered Jailbreaks for LLMs
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b57bb15-a2bf-48a0-aead-4011b5f951a5 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks ” do anything now
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e0a7775-d09c-445a-b87c-59792a2412ed · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Bag of Tricks: Benchmarking of Jailbreak Attacks on LLMs
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed4c2d9f-cf68-4431-816e-f8271dc0c47a · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks A Survey on Trustworthy LLM Agents: Threats and Countermeasures
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b701a1c7-75e3-470a-8697-8a2126d352c8 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Masterkey: Automated jailbreaking of large language model chatbots,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9043d682-2db1-45fc-bc51-2d569acad0a4 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Jailbreakzoo: Survey, landscapes, and horizons in jailbreaking large lan- guage and vision-language models,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1292fb00-17c5-4c13-b6d4-8fd53a9db6bc · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Gpt- 4 is too smart to be safe: Stealthy chat with llms via cipher,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e08f1a45-122a-409c-b3b1-99c7d60ad914 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43d83782-7851-4cf1-97a9-6032f993231a · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Boosting Jailbreak Transferability for Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be428ec5-f14b-4543-a2cc-a2b6aab04e02 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Jailbreaking Black Box Large Language Models in Twenty Queries
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9481f41b-60c7-4599-9b37-10358127434f · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks DeepInception: Hypnotize Large Language Model to Be Jailbreaker
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bec4a2cb-5209-4cbe-9ebe-e9c849fb47d7 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Tempest: Autonomous Multi-Turn Jailbreaking of Large Language Models with Tree Search
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b40f7fa-4ed7-4e8a-ad95-f43854b50ad1 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Defending LLMs against Jailbreaking Attacks via Backtranslation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d520fa31-9bdf-4ace-9186-f1ddc6454d72 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Jailbroken: How does llm safety training fail?
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5aa005cc-1aac-45f5-b27f-407bd964ef24 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a38e1ad-70e8-40d5-be3d-d16c2fb0d05d · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Attack prompt generation for red teaming and defending large language mod- els,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48601139-73e1-473a-ad66-da94405e14d2 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks From Theft to Bomb-Making: The Ripple Effect of Unlearning in Defending Against Jailbreak Attacks
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc595a05-2b83-4101-a1c1-46c4445fa5de · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Baseline Defenses for Adversarial Attacks Against Aligned Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 771bd9f5-e888-44a8-a10d-77ae5df37a46 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Certifying LLM Safety against Adversarial Prompting
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81cb2df2-f37f-457d-a8d0-16e043624665 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba8886ca-b654-4e61-8022-fc96553d7efc · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Defending Against Alignment-Breaking Attacks via Robustly Aligned LLM
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9744ff32-f663-4470-a466-7f4a778c6c67 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Defending chatgpt against jailbreak attack via self-reminders,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0ef8a60-ca68-4930-a283-77ed33a62d93 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Steering dialogue dynamics for robustness against multi-turn jailbreaking attacks,
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9921ae4a-ca99-4243-9263-0365000a2123 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks X-boundary: Establishing exact safety boundary to shield llms from multi-turn jailbreaks without compromising usability,
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b15ac3db-aece-42e1-aa4a-759eb576099d · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks RED QUEEN: Safeguarding Large Language Models against Concealed Multi-Turn Jailbreaking
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e43c751-4a75-44ae-8c4a-12950bd73985 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Generative agents: Interactive simulacra of human behavior,
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bb1203a-fa28-45f2-88ae-bd7312b242a7 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Training Socially Aligned Language Models on Simulated Social Interactions
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac613f5c-476e-4f49-936a-dd1cf9cfb4f0 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Camel: Communicative agents for
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a872541-14d7-4d58-8867-cd96d55ade56 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3c7f28f-25fa-4bd5-9330-81387c2e42df · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Metagpt: Meta programming for a multi-agent collaborative framework,
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66ee22d6-3d85-4913-9552-7bfac0e5507c · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks ChatDev: Communicative Agents for Software Development
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8892c0c9-d9e9-4521-b8d7-f7b5f9749de4 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Improving Factuality and Reasoning in Language Models through Multiagent Debate
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c713df20-a4d7-48cd-9945-5ffdffa5331d · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Encouraging Divergent Thinking in Large Language Models through Multi-Agent Debate
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccbcfc1e-1960-4ffa-af48-ee77962be44e · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Defending large language models against jailbreaking attacks through goal priori- tization,
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3de10f62-1386-4ca4-9a38-e8c914f54690 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Robust prompt optimization for defend- ing language models against jailbreaking attacks,
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e49c6208-7c87-4509-8609-7c19ad7b1c60 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks SecurityLingua: Efficient Defense of LLM Jailbreak Attacks via Security-Aware Prompt Compression
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad2bd6db-372d-4528-bcb9-99a440857d4b · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Fine-tuning aligned language models compromises safety, even when users do not intend to!
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfe5471f-fd88-46ed-a81b-d8486bf54dc1 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Jailbreaking black box large language models in twenty queries,
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2829aa8-a185-4f8d-bd83-8e7ce496b17e · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks M2s: Multi-turn to single-turn jailbreak in red teaming for llms,
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a22fc85-a5fe-456f-b43a-4d89dc2add11 · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks MT-Bench-101: A Fine-Grained Benchmark for Evaluating Large Language Models in Multi-Turn Dialogues
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1702fe08-30d1-48ab-94b9-b7ac696391da · outbound
Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Cosafe: Evaluating large language model safety in multi-turn dialogue coreference,
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.