Pith. sign in

Paper Citation Record · LEDGER

SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2502.12025.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.12025 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:26:08.460709Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T08:55:35.552681Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e9b87ba3-cb91-401b-ae3e-c91b1dfe7eb9 · inbound

A Survey of Scaling in Large Language Model Reasoning cites this paper.

A Survey of Scaling in Large Language Model Reasoning SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:08.979508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:d9e5ac7b10d1f3e29a70a644a2c0d3f466659d8f2e94e1841afcdd67d1a8cd7c

Observation 23a33d8e-2625-4bff-9419-3d0866e4a02e · inbound

Phi-4-reasoning Technical Report cites this paper.

Phi-4-reasoning Technical Report SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:40:25.790569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T03:40:25.706499Z digest=sha256:aa20c3a67f354fbeb987c14f86c12e1dd929f9f29ede5300a362b9277a3a44b1

Observation 91a54c12-0a46-4c0a-8d3e-33d7679a8283 · inbound

R-TOFU: Unlearning in Large Reasoning Models cites this paper.

R-TOFU: Unlearning in Large Reasoning Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:26:08.460709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:26:08.460709Z digest=sha256:97270ac9c60a0f7e129fe72cc6cb0acdf0c5c1126ba63009b2476ea27624ce15

Observation 7cd9c439-37a6-4b5d-9c59-7ef9346469ae · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 154

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:23.220042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:23.220042Z digest=sha256:eab4be5c034b42a07bca4fa6b0e55c9e593171713ae214dc9bd208738ef4f425

Observation d66ec9ff-2935-4e0f-b5a2-1b45b0df9e93 · inbound

Mitigating Deceptive Alignment via Self-Monitoring cites this paper.

Mitigating Deceptive Alignment via Self-Monitoring SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:08.592130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:30:08.592130Z digest=sha256:419002f3740279db120bc931fd72b95d62c599e50288f50551e374287f9b822f

Observation 24de3f31-d729-4346-8c47-8fdb9c4bf5a4 · inbound

VisCRA: A Visual Chain Reasoning Attack for Jailbreaking Multimodal Large Language Models cites this paper.

VisCRA: A Visual Chain Reasoning Attack for Jailbreaking Multimodal Large Language Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:18.072363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:13:18.072363Z digest=sha256:ec1cc8d9f11c968e39e5077f4b3bd4040b88343d1bed8fc27359a1d7627a9dea

Observation cb9dbb89-0ad2-42af-a47f-2198161e37b9 · inbound

Beyond Safe Answers: A Benchmark for Evaluating True Risk Awareness in Large Reasoning Models cites this paper.

Beyond Safe Answers: A Benchmark for Evaluating True Risk Awareness in Large Reasoning Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:49.893258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:49.893258Z digest=sha256:3be7dcecc43d707ddd545080880741ced734842fdc5b73204abc28563b0d67f2

Observation ae230cd9-f965-458c-b10e-c53bee44f1f4 · inbound

Lifelong Safety Alignment for Language Models cites this paper.

Lifelong Safety Alignment for Language Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:06.273227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:06.273227Z digest=sha256:19431e17793618904d06ced072c2b7b44f4ccd4a89a38a47c811e3aca09a919a

Observation 1244253f-4154-4114-bd33-9b4ed1d2517d · inbound

We Should Identify and Mitigate Third-Party Safety Risks in MCP-Powered Agent Systems cites this paper.

We Should Identify and Mitigate Third-Party Safety Risks in MCP-Powered Agent Systems SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:36.556165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:32:36.556165Z digest=sha256:c301b459a2f6a2fa2cd2e3dd1c783cc47941fa9c36d9b382f1e71913fe62cf61

Observation c2e345ca-a9cf-4acd-ac90-7c630980ee51 · inbound

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey cites this paper.

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T17:53:55.578953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:53:55.578953Z digest=sha256:e071e1822f45abf4ffeb25bfe5829c40af9d9772195437c15680295c11670f9a

Observation 93e3aa59-286d-49c3-80c9-8d16f71f1a21 · inbound

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models cites this paper.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 139

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.368021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.368021Z digest=sha256:2199c4727688a338bf809d7f6bb5cf600240dd2362b115c88f4d4ff21049f606

Observation 785f0dd7-84af-428d-90f6-45692600b60a · inbound

Does More Inference-Time Compute Really Help Robustness? cites this paper.

Does More Inference-Time Compute Really Help Robustness? SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:26:20.476585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:26:20.476585Z digest=sha256:8086f7e358a70e0d58ddd61b57fad4360300a9e59fe0238d84a3c7903da9fe26

Observation 0cd5faa1-545e-4019-9f1b-eac6bc1db736 · inbound

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges cites this paper.

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T14:13:06.241993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:13:06.241993Z digest=sha256:f863037750c8da3c5de6036685c4e9123ff8a7e76da9fc0c8dc1ff744b785496

Observation 95da808b-9f6f-439c-b054-0b6db5c73715 · inbound

R1-ACT: Efficient Reasoning Model Safety Alignment by Activating Safety Knowledge cites this paper.

R1-ACT: Efficient Reasoning Model Safety Alignment by Activating Safety Knowledge SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T10:18:24.965589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:18:24.965589Z digest=sha256:8a92d0e192980df250a256ef7e4a29cf92e18f85694f4d0680cac372e1854dc6

Observation cf8e90aa-d624-4460-b90f-7fe9f85e0294 · inbound

The Emotional Baby Is Truly Deadly: Does your Multimodal Large Reasoning Model Have Emotional Flattery towards Humans? cites this paper.

The Emotional Baby Is Truly Deadly: Does your Multimodal Large Reasoning Model Have Emotional Flattery towards Humans? SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T00:59:41.757081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:59:41.757081Z digest=sha256:1048a5230b8b6c57d2567e40066263decb6468b7a2312f4b338f7fdc719c072d

Observation a5c0b53e-4d5c-40f3-a69e-d3a85035e10f · inbound

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments cites this paper.

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-19T01:02:54.860234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T01:02:07.088724Z digest=sha256:846e0b1cb4cc833cbe3e857a71d78215d2905cab43acca571132994a5b818c99

Observation 7d684ae2-6ab1-4786-8186-c77c6ece9bcd · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:38:58.142775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:38:58.142775Z digest=sha256:1793184cfb8c3f1a8712fbe7aa9a01c381a9b0d88ea893ef2a908363b1333c1b

Observation 89dce95f-0d2c-4792-9e61-17f46f8792bb · inbound

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security cites this paper.

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-04T23:09:41.575944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:09:41.575944Z digest=sha256:f53438fdd4e19aa416fcd988580c14ebc752c617ff985a14e66807ce6955ad27

Observation f72b30ac-d771-4576-87cf-fe9620e445e8 · inbound

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards cites this paper.

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:26:28.387214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:24:48.666197Z digest=sha256:ef0db5da0c8bc9ed61034ccca2a992090a006fbd1c45805c0b62ca15c83447e8

Observation 9cca73e6-2752-4c87-8f5b-9eae2fc6848a · inbound

When Models Outthink Their Safety: Unveiling and Mitigating Self-Jailbreak in Large Reasoning Models cites this paper.

When Models Outthink Their Safety: Unveiling and Mitigating Self-Jailbreak in Large Reasoning Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T05:05:55.632098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T05:02:52.176629Z digest=sha256:14541824f5c67dd9a99499c77d72c36641e58bb5ef7f5e9b3e7fe203a07a12fc

Observation 2246ef72-5faa-42b5-8601-6e3ec91efdd9 · inbound

Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations cites this paper.

Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T03:47:15.222573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T03:43:18.987241Z digest=sha256:7321f8c6e0f4addc4abdb323c71723af4478c8fb4372b902dd3a58fea436cc71

Observation f6b26f5c-75ba-4d0a-9964-d8593b2000f9 · inbound

Safety Under Scaffolding: How Evaluation Conditions Shape Measured Safety cites this paper.

Safety Under Scaffolding: How Evaluation Conditions Shape Measured Safety SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-15T13:17:48.274611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:17:48.274611Z digest=sha256:7502affbf8099bbeb2014a0c055f097fa7443da48e7fa46391adb83b328a9272

Observation e7784479-1434-4f86-955f-d66d8e53d913 · inbound

Selective Forgetting for Large Reasoning Models cites this paper.

Selective Forgetting for Large Reasoning Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-13T13:01:23.728846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:01:23.728846Z digest=sha256:b0a75028832be3a6cc4c9ef5223220e777a444c5efafe39a50572bb482a3c23f

Observation 3f893392-45e6-4b59-84e3-90445d5bc50d · inbound

Reasoning Structure Matters for Safety Alignment of Reasoning Models cites this paper.

Reasoning Structure Matters for Safety Alignment of Reasoning Models SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T02:38:17.415455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T02:36:57.093584Z digest=sha256:2af6e158a2db4e15a9ec9a5fcfc05df93844b58f9b37196dc9c70922a3e6c2c7

Observation 4279a036-820a-4619-8e3f-e68edefdcb15 · inbound

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety cites this paper.

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:31:01.221364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T16:00:32.413225Z digest=sha256:cefecaefc32a17e287589e211431b77414f6061942b026ee262dd129e4278cb6

Observation b7261381-028c-4847-a9d9-470fa855da20 · inbound

Addressing Over-Refusal in LLMs with Competing Rewards cites this paper.

Addressing Over-Refusal in LLMs with Competing Rewards SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 94

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:55:35.554191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-01T06:59:12.695984Z digest=sha256:c18de273ed8cc9602baec7c69e912e0a507022b2f59f40512615f3f74ca6a2b4

Observation 5cb68055-c73f-4f62-b7e8-9052c118be16 · inbound

Risky Business: Measuring The Faithfulness-Safety Tension cites this paper.

Risky Business: Measuring The Faithfulness-Safety Tension SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T13:40:55.329417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:40:55.329417Z digest=sha256:7b350886b917080dd61a3eac3c99a2e8bc01d6cc3b2dd27ad75af5d33047944e