Pith. sign in

Paper Citation Record · LEDGER

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

As of 8 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 8 inbound Pith citation observations for arXiv:2506.00782.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00782 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:01:11.473485Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T16:21:26.984437Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T04:16:34.709256Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy13
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9a1e72f4-36bd-4f13-ba4c-6fb71b2c8699 · outbound

This paper cites Claude-3.5-sonnet, 2024.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Claude-3.5-sonnet, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:15.888902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:08.174595Z digest=sha256:58ed9c627d28c6cf7099fb06fb076fc6547daf9f922ed31fc6c09bd89da1c3c0

Observation a57cad7f-0a66-442d-ba67-07d698debdec · outbound

This paper cites Diverse and Effective Red Teaming with Auto-generated Rewards and Multi-step Reinforcement Learning.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Diverse and Effective Red Teaming with Auto-generated Rewards and Multi-step Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.215841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.215841Z digest=sha256:57f04d09a04bfabc06c959cece57cae431415061a9aac92a37faf0bc0516c0b3

Observation 0c83958e-51f1-4715-841a-38e611cdb273 · outbound

This paper cites Bhardwaj, D.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Bhardwaj, D

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:15.779030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:08.265880Z digest=sha256:a7b6eca19e8c3aaa602aaa399f6d0e1b67119ab5c72c81ee4b1594c9df642585

Observation bd7a3d64-e8c8-44d7-8978-302a0a69abae · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning On the Opportunities and Risks of Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.323501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.323501Z digest=sha256:c4a00ec8caffd912006c9dfcd48950bd003546da9c1ea6124fdc6f4015eb7946

Observation f393ed59-e8d9-45e9-8660-c820a41aa1e5 · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.370972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.370972Z digest=sha256:cf8086f7447d34e3e0bb1570aa858431b391f8894b6717564dbec32ac31f0d96

Observation b364c890-f672-4494-8c4e-0c3ca41f3b99 · outbound

This paper cites Chiang, L.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Chiang, L

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:15.678396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:08.431421Z digest=sha256:553a8f63b69cc765ac9acf1338ff1d7ad52ec21e84ae760d62a92688232bef86

Observation 3a3d76af-8b39-4115-a880-59665351a6f2 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:15.529872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:08.467412Z digest=sha256:bb0ca3c28a29899a47d1fcb878a013556fe5dd3c2824671da8da9ce4d1f4d0c0

Observation 673ba450-cba6-4eeb-b443-628662efb98e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.506447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.506447Z digest=sha256:bd5ebd9f015354a40f7700291cf2b9a4369750ebdb4c8944c1bdc9211d4ff22e

Observation c18d2fbe-eeb3-4382-a8db-52aef4f2b1d5 · outbound

This paper cites The Llama 3 Herd of Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.542610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.542610Z digest=sha256:eda3ee810569b79f618078b126acb6de35ed713b2f4349df842f1af0f723f53c

Observation d985b050-1d04-41e4-95bc-c5fb7d70cc38 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:15.437821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:08.618088Z digest=sha256:41bfa59c1e0393f40cef55dfaa72020fe17086339346aea66e1999a963e52365

Observation 072fc78b-14de-47d8-9eb1-8f1452568d74 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:15.329250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:08.659593Z digest=sha256:64d283eeee763bda8b74d12ded3066e7d5503d88d4f7310c999bc5fc608cfd91

Observation ca9c6f7a-6ddd-4ef9-9e2c-e833b8540891 · outbound

This paper cites Best-of-N Jailbreaking.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Best-of-N Jailbreaking

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.712974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.712974Z digest=sha256:43f18ff776a797a13c933de7793b14895aeb5ff677d98498d5dd4a878d2d010d

Observation eefbd577-524a-4cc1-a8d0-3faa2c8147af · outbound

This paper cites PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.749164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.749164Z digest=sha256:d59f0cb81fad590d49febc38864da5a8e64cb0d71166e39071065727266127cb

Observation f96588aa-b5a8-47c9-a3a3-77cb239334db · outbound

This paper cites Jiang, K.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Jiang, K

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:15.216923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:08.790854Z digest=sha256:d13018a9131c3ce346ef11f2ba75330161503191b0d59902aa7a457730445ed1

Observation 8b7839bf-1144-4e59-9590-e49c9f2bbb38 · outbound

This paper cites Learning diverse attacks on large language models for robust red-teaming and safety tuning.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Learning diverse attacks on large language models for robust red-teaming and safety tuning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.855205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.855205Z digest=sha256:bc2931ff3be69fd095ce6d1b34780e817699ac7e53c0418b0449bb7afba652bd

Observation d9d7b985-ec8f-438f-b860-4a5163b078c8 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:15.081114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:08.904306Z digest=sha256:c7bf3dcd9e4bfa634e122349e60a1893cd16edcff33d47d6fb438d4fb0223bca

Observation caa11580-6c71-4128-ad4f-253b6c47ab5d · outbound

This paper cites AmpleGCG: Learning a Universal and Transferable Generative Model of Adversarial Suffixes for Jailbreaking Both Open and Closed LLMs.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AmpleGCG: Learning a Universal and Transferable Generative Model of Adversarial Suffixes for Jailbreaking Both Open and Closed LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.949311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.949311Z digest=sha256:d274f76b7a5c0e1ea3b7e1b965e272e7b16805580403fff1c08cb34b5b66b1a9

Observation 33b04a5e-5201-4103-a505-f126f201da8a · outbound

This paper cites AutoDAN-Turbo: A Lifelong Agent for Strategy Self-Exploration to Jailbreak LLMs.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AutoDAN-Turbo: A Lifelong Agent for Strategy Self-Exploration to Jailbreak LLMs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.010611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.010611Z digest=sha256:84e1b4d275030cd5edabd0c1c823008abf17baa36a63c007b1b9801d12b3acfc

Observation 79503cc3-a8ae-4757-97cc-c009a116ff06 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:14.920262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:09.049209Z digest=sha256:a774ce680afd152d6125949b79cda24b2608713bdda4a6979dc9df87ba07c3c8

Observation 7ef22d1b-236d-4d76-b9a5-70094e50e563 · outbound

This paper cites Auto-RT: Automatic Jailbreak Strategy Exploration for Red-Teaming Large Language Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Auto-RT: Automatic Jailbreak Strategy Exploration for Red-Teaming Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.082636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.082636Z digest=sha256:43e781ae256da919eb9cf7fba08f191dc52ddb08f69cfd4be38e0b6dd552b3f7

Observation 9f76d43c-b567-41d5-8d90-ffe9b0a7e40a · outbound

This paper cites Mazeika, L.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Mazeika, L

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.799153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:09.159736Z digest=sha256:e034554e471018cb3759c3072ce053d76641a091ac0686c1ed1f1fecc1f06945

Observation b1e15a24-cb77-4191-9ab0-92c61ff84c58 · outbound

This paper cites Mehrotra, M.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Mehrotra, M

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.626933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:09.219879Z digest=sha256:2b9a162eff035ebd8ed2fc4f7c4cf163d87187806c853b3f257cdd8d278f7b88

Observation 566053bd-80f3-4d37-853f-df3a49293b28 · outbound

This paper cites Gpt-3.5 turbo, 2023.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Gpt-3.5 turbo, 2023

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.498579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:09.278456Z digest=sha256:75e0a3c76eb6bf738706bbd0bc679106df804594c44e7ee6d23120b0246742be

Observation 626e1ff1-468b-4aaa-bfb7-e71a93d30c8f · outbound

This paper cites Gpt-4o system card, 2024a.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Gpt-4o system card, 2024a

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.309406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:09.326484Z digest=sha256:d840f56dc5644319291d6fbf6d5d4996025701fca70aa2c0e2e398ac57e138bb

Observation 08782bc4-4c05-447a-8f7a-e6054cf6024a · outbound

This paper cites AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.355215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.355215Z digest=sha256:63580e8ad345f8d217802241e968ce477f06b272bc42930b2d9cf27dbd78e938

Observation e8db39e9-155c-47b6-9fc9-841594d08b36 · outbound

This paper cites Perez, S.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Perez, S

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.151394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:09.433892Z digest=sha256:a61944f5893e1236e9e2df285365c65ec89e0b7be5d15645f3250bee0fcd9b0d

Observation 8608203b-ab0d-4b7d-9e6d-5f5952b7c7aa · outbound

This paper cites ToolRL: Reward is All Tool Learning Needs.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning ToolRL: Reward is All Tool Learning Needs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.512168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.512168Z digest=sha256:7bf5ee6321cd516479da5eaf94f379ae54277ec573decf2fbf2409c18d88ddc1

Observation a61ec825-55bf-40eb-92e3-e9999b760f4e · outbound

This paper cites Samvelyan, S.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Samvelyan, S

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:13.886647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:09.600327Z digest=sha256:d041f3a4b427553dcdc45085c730e97ec485f1d7dd24583e0c1574513cad7179

Observation 2fe3d659-1d9e-4572-83d3-d16129c3df13 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.657922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.657922Z digest=sha256:ee1444c0d1027b20aee2f36986d3c920c6cfc74aea464092de0fdef6cc964c62

Observation 337364ca-5d04-4c9d-aa9e-140249eaeb30 · outbound

This paper cites Shaikh, H.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Shaikh, H

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:13.639232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:09.746566Z digest=sha256:24d09941befa59c6f29c0c92b0c22129bc4d89a07c284c10b6de6000e91dfa3a

Observation 33a0d048-074e-4b90-8607-1d344d8b45c1 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.833907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.833907Z digest=sha256:d7084fbb0dd05cb1fe78f775582571528e84c928ccb14b971cb3913770ffad4e

Observation 1da0f343-f8de-4521-ae3e-d911e0b7fecd · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.906691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.906691Z digest=sha256:476c46cc70e1d380c5478fd8f16f84a007aaba435afc20c4a4bb144ce1ca189e

Observation c80dcfb9-4c3d-4a0d-b3d0-46c2a6f5d20a · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.034542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.034542Z digest=sha256:8a6aed3009f08c4b436f81492dae0a83c517106c160a4df692263076c2e6a08e

Observation 20bb5f9f-9282-4eea-9926-6fe1ae134672 · outbound

This paper cites Wang and K.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Wang and K

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:13.402890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:10.174828Z digest=sha256:ffc80bcf6e63a4c9f0c267da42ed7a126645ff3ba03ab838b42c85b23d3b650e

Observation 9116b090-aef8-4210-95f5-bf89a5cdba32 · outbound

This paper cites Qwen2 Technical Report.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Qwen2 Technical Report

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.230800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.230800Z digest=sha256:ae179c35ba7196bf2165361436989ddad196f3449ab9a8831b8c47823148ceef

Observation 964183fa-aa52-49f2-927f-c5173047915a · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:13.185963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:10.304646Z digest=sha256:45d4826885ef707d8db2c48c653be52367a39036255a87a1249d823f9aeaf719

Observation d46e43ee-3901-486a-9347-9043cb9922c4 · outbound

This paper cites STAIR: Improving Safety Alignment with Introspective Reasoning.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning STAIR: Improving Safety Alignment with Introspective Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.437566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.437566Z digest=sha256:a1599769b3f408ddef783fe682413bc9ce85c6566226f14542553c849e022cc5

Observation 9be8f1ce-bb31-4409-9aea-d7454a72728f · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:13.013773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:10.548907Z digest=sha256:6c28813552c1b75f9cc4f24a904f3b421c3f09deeccc65230af968e6db2c95b1

Observation 0aeb746e-80aa-4229-83f4-d714ffbf2ddf · outbound

This paper cites Toward Optimal LLM Alignments Using Two-Player Games.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Toward Optimal LLM Alignments Using Two-Player Games

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.671950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.671950Z digest=sha256:facf1a39b9dfff9eb8a909b7bfc6c6fee99668b2160eb418c559f334bfa581c0

Observation 362e0e77-debc-4ddb-a8c9-f1c8227e1e3d · outbound

This paper cites AutoRedTeamer: Autonomous Red Teaming with Lifelong Attack Integration.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AutoRedTeamer: Autonomous Red Teaming with Lifelong Attack Integration

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.771523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.771523Z digest=sha256:744a4bcc88a15e18a27e75a31c24ffc8138e741ea900e8c4d705c17c312add3f

Observation ed3ab69c-a3f5-45fb-a68a-4206b040c247 · outbound

This paper cites Purple-teaming LLMs with Adversarial Defender Training.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Purple-teaming LLMs with Adversarial Defender Training

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.911575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.911575Z digest=sha256:b74124a22bd80a4e6fe3274138aea4a093e66188b6562f1d9235c64f46d6a06f

Observation cc43c2c8-46e6-48aa-b081-3da843160ed5 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:11.128250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:11.128250Z digest=sha256:15e34c8d87865055f48d0c08c654d628873c2ff48204f9fece240b9d80817cfb

Observation 2d3ff433-976c-4255-ac9a-da42342f056e · outbound

This paper cites After obtaining the filtered 2k samples, we prompt the Qwen2.5-7B-Instruct model to imitate the sample attack as an example.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning After obtaining the filtered 2k samples, we prompt the Qwen2.5-7B-Instruct model to imitate the sample attack as an example

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:12.850546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:11.261804Z digest=sha256:5348db486c3ac1656ecad794714496489aff8ced4bfb2f500a5cf4b598aab8f1

Observation e1cf1b40-4758-4b2e-8d33-9e64c96a27d7 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:12.612250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:01:11.473485Z digest=sha256:ef84346484617f44e1c9cf83c6cc5386158ebe0567d1d88b1b9b333872ddf314

Pith citing papers

Observation 3caa8e0b-18b1-4a1e-a875-266a80aa3761 · inbound

Stable-GFlowNet: Toward Diverse and Robust LLM Red-Teaming via Contrastive Trajectory Balance cites this paper.

Stable-GFlowNet: Toward Diverse and Robust LLM Red-Teaming via Contrastive Trajectory Balance Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:41:16.865385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-09T19:29:36.348680Z digest=sha256:b3ec3fce4446a46ea6bf352faec6ad905f1f6a35471725d75354615d46731dbf

Observation 4426f7f5-4c93-4191-8186-f553eb93f59c · inbound

Internalizing Safety Understanding in Large Reasoning Models via Verification cites this paper.

Internalizing Safety Understanding in Large Reasoning Models via Verification Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:51:14.269926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T01:50:59.283409Z digest=sha256:e908c717f69c417eb44726ffbbfd166a1d278647f704a52f93d5507ff82eed45

Observation 4645865c-d491-420f-917f-eb479a6b126e · inbound

Self-ReSET: Learning to Self-Recover from Unsafe Reasoning Trajectories cites this paper.

Self-ReSET: Learning to Self-Recover from Unsafe Reasoning Trajectories Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:11:16.194279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T02:07:04.364417Z digest=sha256:bc6455c73ca289968a493b682eb13e6003bf3784881b5dbcfb09c67e9cc37900

Observation a31831b8-9265-4c4e-9e19-3609eaaf51d2 · inbound

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs cites this paper.

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:16:34.711696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T09:21:57.373862Z digest=sha256:95b338f22353b8f3fe0e7b4f607e34c29e9112ca3fbe152b1219efc80b80dccf

Observation e777a612-105b-4bf5-bd5a-b1cd2db8b8f5 · inbound

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents cites this paper.

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-11T22:40:37.839133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T22:40:37.839133Z digest=sha256:38bee52c0436e18effc331eccbe34e17d4e32bb88c993240251e14035e0c28db

Observation 0b78426c-7ced-42f2-b964-4f713e454373 · inbound

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents cites this paper.

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-13T07:01:49.222325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:01:49.222325Z digest=sha256:f59bb4dde3593cc1b9182fab4bc407e45bdb9c316c0ae8fc6b66fa9a3c23bba3

Observation f8ab7de0-ef03-447a-9f60-7378374f6c54 · inbound

MJ: Multi-turn LLM Jailbreaking via Decomposed Credit Assignment cites this paper.

MJ: Multi-turn LLM Jailbreaking via Decomposed Credit Assignment Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-14T07:16:57.009797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:16:57.009797Z digest=sha256:7a9d61a5b64aab6d327844f069daed738809b49bafc5fac490aa37e44dec20e8

Observation 6b3f8d1d-b25c-4b2c-a058-d53d4c66aeff · inbound

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs cites this paper.

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T16:21:26.984437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:21:26.984437Z digest=sha256:29d312d50a8ea16e41a92ac3a1ce95b5e61bb1424339be06cd41c91dc9d7cb0d