Pith. sign in

Paper Citation Record · LEDGER

A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2502.15806.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.15806 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:31:23.356745Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f40914cc-6009-4da9-a0cf-589d81c78e93 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 155

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:23.356745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:23.356745Z digest=sha256:8e69279da2800d46c696557887c79372365c4548715401e85857ee4d86264132

Observation dbbadce9-82d1-441d-89e5-0f25e9b13be0 · inbound

A Red Teaming Roadmap Towards System-Level Safety cites this paper.

A Red Teaming Roadmap Towards System-Level Safety A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:25.509233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:11:25.509233Z digest=sha256:9f70b447ebfd2e892bcf6c57d4c0421433a78d3411a76960a0aea7d7932b9848

Observation 6800cdc0-9bb0-4750-9704-262036744b00 · inbound

Probing the Robustness of Large Language Models Safety to Latent Perturbations cites this paper.

Probing the Robustness of Large Language Models Safety to Latent Perturbations A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:18.339304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:18.339304Z digest=sha256:115e4bf142f9ef426ce53bd3831461c978781c2ffa45f42293a343528894eab8

Observation a9d9edf9-7a4c-42c1-a5b5-0a423b16bca3 · inbound

Does More Inference-Time Compute Really Help Robustness? cites this paper.

Does More Inference-Time Compute Really Help Robustness? A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:26:22.339685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:26:22.339685Z digest=sha256:7d69d91f9c961f104335380ab1b77b0a41240ee4d21fdbf490c6e467ea38123e

Observation b7696e2b-4587-485a-9407-7e652911b5ee · inbound

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments cites this paper.

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-19T01:02:54.839444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T01:02:07.088724Z digest=sha256:5b167b4c29a7668e188b559a8d4d23f5aace96b050bfff79f84c529ba8f5a6d6

Observation f98e7050-be1d-4313-a5f9-290799e95582 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:06.933765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:06.933765Z digest=sha256:ec5c5b1105a9dd86ff57078b2c6030bb37073b3690f7b0eaad9e08ee29b02b3e

Observation 110a121c-e183-434f-aef6-346ff9a4de23 · inbound

Less Diverse, Less Safe: The Indirect But Pervasive Risk of Test-Time Scaling in Large Language Models cites this paper.

Less Diverse, Less Safe: The Indirect But Pervasive Risk of Test-Time Scaling in Large Language Models A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:56:13.076463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T09:53:57.765473Z digest=sha256:f6deaac08c380883ca3f3841dc27ec5763680b1230d40551eeed178533abd890

Observation 8cbd3984-d77a-43b7-9ebe-b11eaa65e1a1 · inbound

Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models cites this paper.

Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-04T08:15:58.064118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:15:58.064118Z digest=sha256:b753979616e5de2f9523aa7af712a7d40cec77d011cf8841a4b05c484332c902

Observation 4244c204-aa99-46be-8a7b-b34ba8ba1c84 · inbound

Attention-Guided Reward for Reinforcement Learning-based Jailbreak against Large Reasoning Models cites this paper.

Attention-Guided Reward for Reinforcement Learning-based Jailbreak against Large Reasoning Models A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:38:05.874153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T06:33:30.647965Z digest=sha256:a7b23df37cc4af08ab5aca364b322c645790b63109a162a46ba47a37dcac244c

Observation f2ff8299-ced0-4ffb-9695-09a5301cc154 · inbound

SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling cites this paper.

SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:59:33.713324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T17:23:06.382761Z digest=sha256:cf4a110caa36b5300c1603ad4dd3c90ffc7f3449f28b0dcbd2edf2f89c9f3a18

Observation 63cfdfbd-f4ce-438b-a4fa-4507d18b7f79 · inbound

ToxiREX: A Dataset on Toxic REasoning in ConteXt cites this paper.

ToxiREX: A Dataset on Toxic REasoning in ConteXt A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 226

Resolution
verified exact
arxiv_id, observed 2026-06-29T04:43:07.170199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T04:33:18.794505Z digest=sha256:42ea9f3b468c36e6b5c2d77226d1406e2bf627dfb1541486a06e72dc5839f1c1

Observation 3ffc37a2-3e0c-4263-b51e-d3e355932c4f · inbound

Addressing Over-Refusal in LLMs with Competing Rewards cites this paper.

Addressing Over-Refusal in LLMs with Competing Rewards A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:55:35.545009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T06:59:12.695984Z digest=sha256:0c1b9649da656a948321723e9927b092436ad319248580a8e10abc314d5e0549