Pith. sign in

Paper Citation Record · LEDGER

A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2502.15806.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.15806 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:43:44.659105Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fbce5c5e-de55-46b8-be4e-f93f801f0854 · inbound

100 Days After DeepSeek-R1: A Survey on Replication Studies and More Directions for Reasoning Language Models cites this paper.

100 Days After DeepSeek-R1: A Survey on Replication Studies and More Directions for Reasoning Language Models A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 147

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:44.659105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:44.659105Z digest=sha256:2f11350d66d8dc06aaafe76d95347e318a719396c1fbaeb00912cd6b49d7b05e

Observation 5b73d6e8-2592-4f58-95f0-be3cd567a42a · inbound

Practical Reasoning Interruption Attacks on Reasoning Large Language Models cites this paper.

Practical Reasoning Interruption Attacks on Reasoning Large Language Models A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T22:42:00.160331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:42:00.160331Z digest=sha256:626d541e9461c4c190b479420cc28a2e2f0ea413547ba514ebe1d22bd5d217d3

Observation f40914cc-6009-4da9-a0cf-589d81c78e93 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 155

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:23.356745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:23.356745Z digest=sha256:0366b722e0c13eafd68754542403170ad05cf5d4398ba910a90d051a9615bb85

Observation dbbadce9-82d1-441d-89e5-0f25e9b13be0 · inbound

A Red Teaming Roadmap Towards System-Level Safety cites this paper.

A Red Teaming Roadmap Towards System-Level Safety A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:25.509233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:11:25.509233Z digest=sha256:cdee1ccc1fe259e7d4dd338cb0a141ea5349dac8b4268a55bd4bef0ffe15e20d

Observation 6800cdc0-9bb0-4750-9704-262036744b00 · inbound

Probing the Robustness of Large Language Models Safety to Latent Perturbations cites this paper.

Probing the Robustness of Large Language Models Safety to Latent Perturbations A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:18.339304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:18.339304Z digest=sha256:ae4ad9bbff80ab2293f104b352067a701283ab34b4df2d3374fe16237a0de7b5

Observation a9d9edf9-7a4c-42c1-a5b5-0a423b16bca3 · inbound

Does More Inference-Time Compute Really Help Robustness? cites this paper.

Does More Inference-Time Compute Really Help Robustness? A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:26:22.339685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:26:22.339685Z digest=sha256:94044114f581d54dd16984078d7fdb198672c5e1b3afcf17dc3d6235e6260a56

Observation b7696e2b-4587-485a-9407-7e652911b5ee · inbound

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments cites this paper.

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-19T01:02:54.839444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-19T01:02:07.088724Z digest=sha256:b7982422e5e5197cb0305eabe44c2fb7453b6fe08a6634eb7ebd4f932bb9bf26

Observation f98e7050-be1d-4313-a5f9-290799e95582 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:06.933765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:06.933765Z digest=sha256:8bbb3bbe2a913cb01bb68985c56a5746b5f73bedbbe8fcd22b429e8631e67015

Observation 110a121c-e183-434f-aef6-346ff9a4de23 · inbound

Less Diverse, Less Safe: The Indirect But Pervasive Risk of Test-Time Scaling in Large Language Models cites this paper.

Less Diverse, Less Safe: The Indirect But Pervasive Risk of Test-Time Scaling in Large Language Models A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:56:13.076463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T09:53:57.765473Z digest=sha256:7a0c6ba2c7821e69f115b379f6cc769c6019806841f3761a74feaed031955d99

Observation 8cbd3984-d77a-43b7-9ebe-b11eaa65e1a1 · inbound

Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models cites this paper.

Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-04T08:15:58.064118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:15:58.064118Z digest=sha256:2a6900cac34df90cd888bcf62c04cf1d1e30ffde38a45f13730cfd6f56abe693

Observation 4244c204-aa99-46be-8a7b-b34ba8ba1c84 · inbound

Attention-Guided Reward for Reinforcement Learning-based Jailbreak against Large Reasoning Models cites this paper.

Attention-Guided Reward for Reinforcement Learning-based Jailbreak against Large Reasoning Models A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:38:05.874153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T06:33:30.647965Z digest=sha256:0e0f0912a103b96dba1e6ebe46c1a75b4a7738ee5cff4cca26e47d1a60415d5e

Observation f2ff8299-ced0-4ffb-9695-09a5301cc154 · inbound

SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling cites this paper.

SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:59:33.713324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T17:23:06.382761Z digest=sha256:111f16864d600c1f908fa5ff08523ea35575ad1d07d8b295914f0af6745ab49e

Observation 63cfdfbd-f4ce-438b-a4fa-4507d18b7f79 · inbound

ToxiREX: A Dataset on Toxic REasoning in ConteXt cites this paper.

ToxiREX: A Dataset on Toxic REasoning in ConteXt A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 226

Resolution
verified exact
arxiv_id, observed 2026-06-29T04:43:07.170199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-29T04:33:18.794505Z digest=sha256:a4bd4c6fad09cfab591bdf03d8786c8e3ed4e3cbc6394001c67ab48c5ee4c507

Observation 3ffc37a2-3e0c-4263-b51e-d3e355932c4f · inbound

Addressing Over-Refusal in LLMs with Competing Rewards cites this paper.

Addressing Over-Refusal in LLMs with Competing Rewards A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:55:35.545009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-07-01T06:59:12.695984Z digest=sha256:12d47072c2b63f0afadf2b06a3d787723b9ed9be1573cb90da7aa41b1d665eaf