Pith. sign in

Paper Citation Record · LEDGER

SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2502.12025.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.12025 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:48:29.146185Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T08:55:35.552681Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e9b87ba3-cb91-401b-ae3e-c91b1dfe7eb9 · inbound

A Survey of Scaling in Large Language Model Reasoning cites this paper.

A Survey of Scaling in Large Language Model Reasoning SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.979508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:ecd0fb3cfd3dbd980b5fd6f3f0bdf8c88b67e9593a45bc80f1b9b19a5e64ccb8

Observation 23a33d8e-2625-4bff-9419-3d0866e4a02e · inbound

Phi-4-reasoning Technical Report cites this paper.

Phi-4-reasoning Technical Report SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:40:25.790569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-17T03:40:25.706499Z digest=sha256:68ffbd9954480288fe3971c669388e683b0e4cd3172d94dfa82e8f90dc73f7ca

Observation 91a54c12-0a46-4c0a-8d3e-33d7679a8283 · inbound

R-TOFU: Unlearning in Large Reasoning Models cites this paper.

R-TOFU: Unlearning in Large Reasoning Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:08.460709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:26:08.460709Z digest=sha256:4385857c247363fe0b4aff431c5257409b5407f1f3d9b4b182e6bacbdb1d81c6

Observation 7cd9c439-37a6-4b5d-9c59-7ef9346469ae · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 154

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:23.220042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:23.220042Z digest=sha256:9a48937f673a83536f0e0073a7612bacbe8b67d001b1afccd71c071304e7fe89

Observation d66ec9ff-2935-4e0f-b5a2-1b45b0df9e93 · inbound

Mitigating Deceptive Alignment via Self-Monitoring cites this paper.

Mitigating Deceptive Alignment via Self-Monitoring SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:08.592130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:30:08.592130Z digest=sha256:56aeb68fc471faadc013067bfd4c6e089087666f01326954016095f9c6a31120

Observation 24de3f31-d729-4346-8c47-8fdb9c4bf5a4 · inbound

VisCRA: A Visual Chain Reasoning Attack for Jailbreaking Multimodal Large Language Models cites this paper.

VisCRA: A Visual Chain Reasoning Attack for Jailbreaking Multimodal Large Language Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:18.072363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:13:18.072363Z digest=sha256:2ce1a614c64c008e5368d3608b46f1b57034b4424d4a73e36585ed8f46261586

Observation cb9dbb89-0ad2-42af-a47f-2198161e37b9 · inbound

Beyond Safe Answers: A Benchmark for Evaluating True Risk Awareness in Large Reasoning Models cites this paper.

Beyond Safe Answers: A Benchmark for Evaluating True Risk Awareness in Large Reasoning Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:49.893258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:49.893258Z digest=sha256:ba7e0527560efdaff395daf482b35b9810410f0b20659e16518fd1f875b7f63b

Observation ae230cd9-f965-458c-b10e-c53bee44f1f4 · inbound

Lifelong Safety Alignment for Language Models cites this paper.

Lifelong Safety Alignment for Language Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:06.273227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:06.273227Z digest=sha256:9f32eb300c410207c6b9dbf23f7f2ced749107748aab23c35b58c15becd37451

Observation 1244253f-4154-4114-bd33-9b4ed1d2517d · inbound

We Should Identify and Mitigate Third-Party Safety Risks in MCP-Powered Agent Systems cites this paper.

We Should Identify and Mitigate Third-Party Safety Risks in MCP-Powered Agent Systems SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:36.556165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:32:36.556165Z digest=sha256:c05237527f0595374a8aeaff0cf8e5d86a70d0766fcfd17bd732680c9b2adfff

Observation c2e345ca-a9cf-4acd-ac90-7c630980ee51 · inbound

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey cites this paper.

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T17:53:55.578953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:53:55.578953Z digest=sha256:38e75aa4b60985bd056974998697cc07bc04e7ac06602cc766b97f8fcd2259e8

Observation 93e3aa59-286d-49c3-80c9-8d16f71f1a21 · inbound

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models cites this paper.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 139

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.368021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.368021Z digest=sha256:2763e0da2048e4b844a80dc29057951b8e7836a85af584c348c5085c5e3bd575

Observation 785f0dd7-84af-428d-90f6-45692600b60a · inbound

Does More Inference-Time Compute Really Help Robustness? cites this paper.

Does More Inference-Time Compute Really Help Robustness? SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:26:20.476585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:26:20.476585Z digest=sha256:a0d558d9e16e60cd3697cb674fd6f5a489282f04ffaebc05ae719f90e71239dd

Observation 0cd5faa1-545e-4019-9f1b-eac6bc1db736 · inbound

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges cites this paper.

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T14:13:06.241993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:13:06.241993Z digest=sha256:b8df17a95fafd1f5f57624b88e24a4b88eddbcc3df06355adc6f856b03fe88b6

Observation 95da808b-9f6f-439c-b054-0b6db5c73715 · inbound

R1-ACT: Efficient Reasoning Model Safety Alignment by Activating Safety Knowledge cites this paper.

R1-ACT: Efficient Reasoning Model Safety Alignment by Activating Safety Knowledge SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:24.965589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:18:24.965589Z digest=sha256:c89847c33ea113a9c3b4079ca38ec91571683d3c9030d08072711b7502f282a8

Observation cf8e90aa-d624-4460-b90f-7fe9f85e0294 · inbound

The Emotional Baby Is Truly Deadly: Does your Multimodal Large Reasoning Model Have Emotional Flattery towards Humans? cites this paper.

The Emotional Baby Is Truly Deadly: Does your Multimodal Large Reasoning Model Have Emotional Flattery towards Humans? SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T00:59:41.757081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:59:41.757081Z digest=sha256:040c3e3931d8428d5732bc2003604fee3b5b1fc894042dfcb257b72dc0b2b62c

Observation a5c0b53e-4d5c-40f3-a69e-d3a85035e10f · inbound

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments cites this paper.

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-19T01:02:54.860234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-19T01:02:07.088724Z digest=sha256:23455cc98c23db2dadacd657fb23c49c6e108afe2b0b1ad50c2874da49510db6

Observation 7d684ae2-6ab1-4786-8186-c77c6ece9bcd · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:38:58.142775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:38:58.142775Z digest=sha256:ff9925d2e686fd5d27ba533f4b62b24f35fb6e3a4eb7eb2487d12b245e31e54e

Observation 89dce95f-0d2c-4792-9e61-17f46f8792bb · inbound

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security cites this paper.

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-04T23:09:41.575944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:09:41.575944Z digest=sha256:cfc04e6dc980b109c19e5d4b897ca70bc982b402969b8ea5519337107778965c

Observation f72b30ac-d771-4576-87cf-fe9620e445e8 · inbound

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards cites this paper.

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:26:28.387214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-18T14:24:48.666197Z digest=sha256:433663939ee86350cc07427a3d7db98e3dbddba726276abdf8d434a56d5919da

Observation deff1170-7053-4616-8c1d-4b7c29da20c6 · inbound

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards cites this paper.

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T15:48:29.146185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:48:29.146185Z digest=sha256:18d059265ffe502035cbc7906c56c23d40c3ad5db03326bba6f36a8800a47657

Observation 9cca73e6-2752-4c87-8f5b-9eae2fc6848a · inbound

When Models Outthink Their Safety: Unveiling and Mitigating Self-Jailbreak in Large Reasoning Models cites this paper.

When Models Outthink Their Safety: Unveiling and Mitigating Self-Jailbreak in Large Reasoning Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T05:05:55.632098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-18T05:02:52.176629Z digest=sha256:458282a04f0a071b61c998748a37f164320208a513c51b4bffeff119b1dca2c7

Observation 2246ef72-5faa-42b5-8601-6e3ec91efdd9 · inbound

Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations cites this paper.

Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T03:47:15.222573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T03:43:18.987241Z digest=sha256:485ac1030945d939f1b7c8d9ed82c3977de40220b3deafd1376bb3019b39fc7c

Observation f6b26f5c-75ba-4d0a-9964-d8593b2000f9 · inbound

Safety Under Scaffolding: How Evaluation Conditions Shape Measured Safety cites this paper.

Safety Under Scaffolding: How Evaluation Conditions Shape Measured Safety SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-15T13:17:48.274611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:17:48.274611Z digest=sha256:31ef24a30a44a48da84c03e7b0e5a2a6e2e67fddc4aaeba6f76ba525fa1d805a

Observation e7784479-1434-4f86-955f-d66d8e53d913 · inbound

Selective Forgetting for Large Reasoning Models cites this paper.

Selective Forgetting for Large Reasoning Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-13T13:01:23.728846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:01:23.728846Z digest=sha256:3e078d4dc0d179462969054e8dae66f7ed18bc542344da6db06ffbfde0b7844a

Observation 3f893392-45e6-4b59-84e3-90445d5bc50d · inbound

Reasoning Structure Matters for Safety Alignment of Reasoning Models cites this paper.

Reasoning Structure Matters for Safety Alignment of Reasoning Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T02:38:17.415455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-10T02:36:57.093584Z digest=sha256:1474165db4d865f96c0f9c9e8d91200c328851354a9a5fab679e2a8de5506471

Observation 4279a036-820a-4619-8e3f-e68edefdcb15 · inbound

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety cites this paper.

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:31:01.221364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-10T16:00:32.413225Z digest=sha256:67157bc91a587eb9f162c6b1fe4748c3c06b5b7d34bcd77ba0dad8f516bd9d4b

Observation b7261381-028c-4847-a9d9-470fa855da20 · inbound

Addressing Over-Refusal in LLMs with Competing Rewards cites this paper.

Addressing Over-Refusal in LLMs with Competing Rewards SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 94

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:55:35.554191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-07-01T06:59:12.695984Z digest=sha256:7384d80b9a0001054dbbe4cd853db17c5bc263e9314fb842bf3e157cbe209a1f

Observation 416ce64f-faaf-4207-8b34-551df27d8ef5 · inbound

Reasoning as a Double-Edged Sword: Architecture and Cross-Stage Robustness in Vision-Language-Action Models cites this paper.

Reasoning as a Double-Edged Sword: Architecture and Cross-Stage Robustness in Vision-Language-Action Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T15:41:31.489483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:41:31.489483Z digest=sha256:382d6a3b12e070e8d801908cacaaa4c84827e73412f46956d650436208aa7f87

Observation 5cb68055-c73f-4f62-b7e8-9052c118be16 · inbound

Risky Business: Measuring The Faithfulness-Safety Tension cites this paper.

Risky Business: Measuring The Faithfulness-Safety Tension SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T13:40:55.329417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:40:55.329417Z digest=sha256:8f3391e0a8af833d33fb4bbd19b22674e86ee418291a1d3f5b7cc51383b4572b