Pith. sign in

Paper Citation Record · LEDGER

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

As of 20 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 9 inbound Pith citation observations for arXiv:2506.00782.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00782 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:01:11.473485Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:37:16.783441Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T04:16:34.709256Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy13
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9a1e72f4-36bd-4f13-ba4c-6fb71b2c8699 · outbound

This paper cites Claude-3.5-sonnet, 2024.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Claude-3.5-sonnet, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:15.888902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:08.174595Z digest=sha256:a5cd7a47481a82dbdfd10ca0c005bb05cabcec64d1592b6a7400e0968f8f17f2

Observation a57cad7f-0a66-442d-ba67-07d698debdec · outbound

This paper cites Diverse and Effective Red Teaming with Auto-generated Rewards and Multi-step Reinforcement Learning.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Diverse and Effective Red Teaming with Auto-generated Rewards and Multi-step Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.215841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.215841Z digest=sha256:d3f1236dd1016ef351b7d272f8fb17440c8653260f8db2aa06b42358ac295af1

Observation 0c83958e-51f1-4715-841a-38e611cdb273 · outbound

This paper cites Bhardwaj, D.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Bhardwaj, D

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:15.779030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:08.265880Z digest=sha256:8cdb38f82bb26277bc7dd39ebe1742d0305f4bcc11beebb891753486c7966cb7

Observation bd7a3d64-e8c8-44d7-8978-302a0a69abae · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning On the Opportunities and Risks of Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.323501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.323501Z digest=sha256:eb5d0c766342bcfc24dde466555b2eb91aa91cc9d1704f35a0323418d0be3906

Observation f393ed59-e8d9-45e9-8660-c820a41aa1e5 · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.370972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.370972Z digest=sha256:f465fd7d431b06acdf54f29e9ebf30a28f055bb8f2fdc9fd5fd009291ecfbad6

Observation b364c890-f672-4494-8c4e-0c3ca41f3b99 · outbound

This paper cites Chiang, L.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Chiang, L

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:15.678396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:08.431421Z digest=sha256:9de514e615d51c148022850291384280dfe4ec50b5c5767f658643c9fdf8be5f

Observation 3a3d76af-8b39-4115-a880-59665351a6f2 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:15.529872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:08.467412Z digest=sha256:baa5431e03cb0b65fd989a97d821656b43460c51cee11739217e0c3acc9bcb41

Observation 673ba450-cba6-4eeb-b443-628662efb98e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.506447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.506447Z digest=sha256:a8988dace0797063d3f68863e4c71369de4eedf5f717e9e68c955723ab142824

Observation c18d2fbe-eeb3-4382-a8db-52aef4f2b1d5 · outbound

This paper cites The Llama 3 Herd of Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.542610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.542610Z digest=sha256:16078bb8d72d81bb566dfdcd8afe19f4167a29f76c4e427afde37a871f54284b

Observation d985b050-1d04-41e4-95bc-c5fb7d70cc38 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:15.437821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:08.618088Z digest=sha256:75aedd175035052cc8358eb3b30cda958048590fc82640979db534273e7e06a6

Observation 072fc78b-14de-47d8-9eb1-8f1452568d74 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:15.329250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:08.659593Z digest=sha256:faa53051c57caf345c1069889ac8bfea366b8bad1f96182b77dfa2b7c7875a75

Observation ca9c6f7a-6ddd-4ef9-9e2c-e833b8540891 · outbound

This paper cites Best-of-N Jailbreaking.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Best-of-N Jailbreaking

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.712974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.712974Z digest=sha256:a011ef6aac2fb3b61758b5ae56829e5e20a69dcabc990c5ea773dcb1f9b5730c

Observation eefbd577-524a-4cc1-a8d0-3faa2c8147af · outbound

This paper cites PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.749164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.749164Z digest=sha256:3d7b4e1ef1980752bf5b9fedb19ad5be17ab9bbef18cb71a0baa0a1900fb00e5

Observation f96588aa-b5a8-47c9-a3a3-77cb239334db · outbound

This paper cites Jiang, K.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Jiang, K

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:15.216923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:08.790854Z digest=sha256:483a324f70ccaf2726d596a506c39ed58cbe862ce4bacb9c293d027499ae95a3

Observation 8b7839bf-1144-4e59-9590-e49c9f2bbb38 · outbound

This paper cites Learning diverse attacks on large language models for robust red-teaming and safety tuning.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Learning diverse attacks on large language models for robust red-teaming and safety tuning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.855205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.855205Z digest=sha256:14763627f07c92be97ea300ce4c105343c0ca909279c0ec2cb784946eb2e016d

Observation d9d7b985-ec8f-438f-b860-4a5163b078c8 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:15.081114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:08.904306Z digest=sha256:98e5e59afc32e4954063454c8c6332f1f142ec53dc74e78eac6e5bfd4a2ae8da

Observation caa11580-6c71-4128-ad4f-253b6c47ab5d · outbound

This paper cites AmpleGCG: Learning a Universal and Transferable Generative Model of Adversarial Suffixes for Jailbreaking Both Open and Closed LLMs.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AmpleGCG: Learning a Universal and Transferable Generative Model of Adversarial Suffixes for Jailbreaking Both Open and Closed LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.949311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.949311Z digest=sha256:34aabee09d508881afc7466d2334b5cf0bb30aac95256845f88e6357d6bad90a

Observation 33b04a5e-5201-4103-a505-f126f201da8a · outbound

This paper cites AutoDAN-Turbo: A Lifelong Agent for Strategy Self-Exploration to Jailbreak LLMs.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AutoDAN-Turbo: A Lifelong Agent for Strategy Self-Exploration to Jailbreak LLMs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.010611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.010611Z digest=sha256:b52bf221d89813696c449e73b638140d5779175ddb49dbd6fa21db1d9dc5eabd

Observation 79503cc3-a8ae-4757-97cc-c009a116ff06 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:14.920262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:09.049209Z digest=sha256:635f09afeb6cfa0b3d6b9b5991b6ddc4acdda0ff5bc9aa494e3391a53506c6d3

Observation 7ef22d1b-236d-4d76-b9a5-70094e50e563 · outbound

This paper cites Auto-RT: Automatic Jailbreak Strategy Exploration for Red-Teaming Large Language Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Auto-RT: Automatic Jailbreak Strategy Exploration for Red-Teaming Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.082636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.082636Z digest=sha256:80ae065b0eb3ccb708d563a96ed14a75a1b3b2fcd95a42dba5995f1464d1fff9

Observation 9f76d43c-b567-41d5-8d90-ffe9b0a7e40a · outbound

This paper cites Mazeika, L.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Mazeika, L

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.799153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:09.159736Z digest=sha256:4777b0f05b6d1188425abdf53e6bed4c2907f32455ca27723d40544ebfeac25a

Observation b1e15a24-cb77-4191-9ab0-92c61ff84c58 · outbound

This paper cites Mehrotra, M.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Mehrotra, M

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.626933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:09.219879Z digest=sha256:d61cfba7027e4f4f120cc13bb299b7583884c5689307aad5770244b777ee7f67

Observation 566053bd-80f3-4d37-853f-df3a49293b28 · outbound

This paper cites Gpt-3.5 turbo, 2023.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Gpt-3.5 turbo, 2023

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.498579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:09.278456Z digest=sha256:379936a8aae91d439a0d0b5ebfa95afad2612d566df33173958d319c7e4f8869

Observation 626e1ff1-468b-4aaa-bfb7-e71a93d30c8f · outbound

This paper cites Gpt-4o system card, 2024a.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Gpt-4o system card, 2024a

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.309406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:09.326484Z digest=sha256:0568bbd9e9687f6ba221fa20e3ba04c9dd428eb1afd05d239490d99f2b08da25

Observation 08782bc4-4c05-447a-8f7a-e6054cf6024a · outbound

This paper cites AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.355215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.355215Z digest=sha256:194820b4fac9137025b4f2bd8e3ff3536e3fb7c0e43f8408d5a551556a937279

Observation e8db39e9-155c-47b6-9fc9-841594d08b36 · outbound

This paper cites Perez, S.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Perez, S

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.151394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:09.433892Z digest=sha256:4a1b179c8e144dc4add83c6ed5462d0c7f7a96582103a26164e7e570aa54ba64

Observation 8608203b-ab0d-4b7d-9e6d-5f5952b7c7aa · outbound

This paper cites ToolRL: Reward is All Tool Learning Needs.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning ToolRL: Reward is All Tool Learning Needs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.512168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.512168Z digest=sha256:7b0bde8a614951d5de62eb442205d8e661583cf9896c23db958943b651b570a7

Observation a61ec825-55bf-40eb-92e3-e9999b760f4e · outbound

This paper cites Samvelyan, S.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Samvelyan, S

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:13.886647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:09.600327Z digest=sha256:839a67d1b4d6cc456c343c8782b3c27aa2693f94186cb6396af78a6422fbcba9

Observation 2fe3d659-1d9e-4572-83d3-d16129c3df13 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.657922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.657922Z digest=sha256:f49dd8e28d4cd714447f3320bc0a6cb4e41143fe4203bbcff9d2ca94a5e0bc3c

Observation 337364ca-5d04-4c9d-aa9e-140249eaeb30 · outbound

This paper cites Shaikh, H.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Shaikh, H

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:13.639232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:09.746566Z digest=sha256:498d8eb8ac0e71d7450a283a42aefda953332470f5f9cb9c762b7cb306adffdb

Observation 33a0d048-074e-4b90-8607-1d344d8b45c1 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.833907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.833907Z digest=sha256:3aa81614d29bcbdb4cded59cfe43672813c72da87147eb2fcb4670d6b69df37c

Observation 1da0f343-f8de-4521-ae3e-d911e0b7fecd · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.906691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.906691Z digest=sha256:58959c9b9b5bb1c612b5edc67a1c5e721403822cab7995f59f43ad84d07232c1

Observation c80dcfb9-4c3d-4a0d-b3d0-46c2a6f5d20a · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.034542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.034542Z digest=sha256:8b1eb1b3d59e667b8252dd15c394b1d41a96d2a4af516f7b70cb6dd57bbf2551

Observation 20bb5f9f-9282-4eea-9926-6fe1ae134672 · outbound

This paper cites Wang and K.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Wang and K

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:13.402890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:10.174828Z digest=sha256:92db8373fc9b52dfc30c292eef941a5f25504eaba9e53468720ab9965b6bac78

Observation 9116b090-aef8-4210-95f5-bf89a5cdba32 · outbound

This paper cites Qwen2 Technical Report.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Qwen2 Technical Report

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.230800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.230800Z digest=sha256:41101026607d61136769076e5b2a10008df62f1e7158f028308869addde2b644

Observation 964183fa-aa52-49f2-927f-c5173047915a · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:13.185963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:10.304646Z digest=sha256:464d5810161926ff4e402e26b0e99a8520e7ccf9b691eef0ed62d486c5397031

Observation d46e43ee-3901-486a-9347-9043cb9922c4 · outbound

This paper cites STAIR: Improving Safety Alignment with Introspective Reasoning.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning STAIR: Improving Safety Alignment with Introspective Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.437566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.437566Z digest=sha256:d1b32b369ddb254a9a8e49a31be247e3bba7785a29ea2a55b5b87e108eb52431

Observation 9be8f1ce-bb31-4409-9aea-d7454a72728f · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:13.013773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:10.548907Z digest=sha256:70b1a7854fea3e1ff0f5bbec7930050c87d66e4acc569cb0e5faa6101f33e56d

Observation 0aeb746e-80aa-4229-83f4-d714ffbf2ddf · outbound

This paper cites Toward Optimal LLM Alignments Using Two-Player Games.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Toward Optimal LLM Alignments Using Two-Player Games

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.671950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.671950Z digest=sha256:42467151924c3c60de79253474820eb8296135a0b1f3759361913e6c7b87df45

Observation 362e0e77-debc-4ddb-a8c9-f1c8227e1e3d · outbound

This paper cites AutoRedTeamer: Autonomous Red Teaming with Lifelong Attack Integration.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AutoRedTeamer: Autonomous Red Teaming with Lifelong Attack Integration

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.771523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.771523Z digest=sha256:923de5d30d01fe95dc4976eb5b5580b28faeb0d40f0b70efa67d0685b8f330cb

Observation ed3ab69c-a3f5-45fb-a68a-4206b040c247 · outbound

This paper cites Purple-teaming LLMs with Adversarial Defender Training.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Purple-teaming LLMs with Adversarial Defender Training

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.911575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.911575Z digest=sha256:7094a5c766f66350ce504afa65a29c17aa211d78a884da53a9901be4fce7403f

Observation cc43c2c8-46e6-48aa-b081-3da843160ed5 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:11.128250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:11.128250Z digest=sha256:ce12097cd4382f381fa5e3ba611cafb89f64ab352cbdc8b4a9d104c5fc6d7b4b

Observation 2d3ff433-976c-4255-ac9a-da42342f056e · outbound

This paper cites After obtaining the filtered 2k samples, we prompt the Qwen2.5-7B-Instruct model to imitate the sample attack as an example.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning After obtaining the filtered 2k samples, we prompt the Qwen2.5-7B-Instruct model to imitate the sample attack as an example

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:12.850546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:11.261804Z digest=sha256:8e1dd7c3e5f0cda923c7fe43da39f78e97c33f12c62c3581974fbcb455ddea08

Observation e1cf1b40-4758-4b2e-8d33-9e64c96a27d7 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:12.612250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:01:11.473485Z digest=sha256:0291e23989f51c610829546e4bc447330ee3286d8185b14a88c1f1bd681ad471

Pith citing papers

Observation 3caa8e0b-18b1-4a1e-a875-266a80aa3761 · inbound

Stable-GFlowNet: Toward Diverse and Robust LLM Red-Teaming via Contrastive Trajectory Balance cites this paper.

Stable-GFlowNet: Toward Diverse and Robust LLM Red-Teaming via Contrastive Trajectory Balance Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:41:16.865385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-09T19:29:36.348680Z digest=sha256:aa17527fb2ffa1c284b0a1a2e268255e16b5298605e1bccb9bba224c9245d8f7

Observation 4426f7f5-4c93-4191-8186-f553eb93f59c · inbound

Internalizing Safety Understanding in Large Reasoning Models via Verification cites this paper.

Internalizing Safety Understanding in Large Reasoning Models via Verification Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:51:14.269926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-12T01:50:59.283409Z digest=sha256:91136aff38da86e1d76dd725b627b86881d3ea9a9c67dcf8a950667c17f948a7

Observation 4645865c-d491-420f-917f-eb479a6b126e · inbound

Self-ReSET: Learning to Self-Recover from Unsafe Reasoning Trajectories cites this paper.

Self-ReSET: Learning to Self-Recover from Unsafe Reasoning Trajectories Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:11:16.194279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-12T02:07:04.364417Z digest=sha256:9b04e54f41d0fefabcb2b079d8f8486c5814993139f39b68f35e03777d2b7599

Observation a31831b8-9265-4c4e-9e19-3609eaaf51d2 · inbound

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs cites this paper.

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:16:34.711696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T09:21:57.373862Z digest=sha256:7ad323bde3b22f7407b37d2901b5a9ec599106d040d0b6caa9f54ab599d30239

Observation e777a612-105b-4bf5-bd5a-b1cd2db8b8f5 · inbound

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents cites this paper.

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-11T22:40:37.839133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T22:40:37.839133Z digest=sha256:4e60d30560b599aa93875dd0a41500f9ce715d9c10dc6e8ef89dc28c8fb73689

Observation 0b78426c-7ced-42f2-b964-4f713e454373 · inbound

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents cites this paper.

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-13T07:01:49.222325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:01:49.222325Z digest=sha256:86f44ace8f8471666a6455d661a505567617603f48c27d6eb6baf9b8796df258

Observation f8ab7de0-ef03-447a-9f60-7378374f6c54 · inbound

MJ: Multi-turn LLM Jailbreaking via Decomposed Credit Assignment cites this paper.

MJ: Multi-turn LLM Jailbreaking via Decomposed Credit Assignment Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-14T07:16:57.009797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:16:57.009797Z digest=sha256:ec3bc3dfdbfee4f4968d2a10215f8c016221e9a97d0d23cc41d66f6595a02516

Observation 6b3f8d1d-b25c-4b2c-a058-d53d4c66aeff · inbound

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs cites this paper.

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T16:21:26.984437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:21:26.984437Z digest=sha256:64069a8a256318ddc18d8470354b6aefa46c4c0380ea88ab1f980a7a527d39ec

Observation ee15abf9-9bed-41e4-ae5b-10ddc5902714 · inbound

Learning to Persuade Exposes How Easily LLMs Abandon Correct Beliefs cites this paper.

Learning to Persuade Exposes How Easily LLMs Abandon Correct Beliefs Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:16.783441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:16.783441Z digest=sha256:0c539a68e4befe58e89d32561f03501ec74ca7edac6f301bf525a3bad0646232