Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T18:15:32.936227Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2502.00669.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T18:15:32.936227Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a1c68aee-393d-41ac-aa02-6dbec66f71ab · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2adb4e64-ac68-4f96-a29f-93594ba5248d · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective GPT-4 Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ede8505-f567-4d25-8e08-42b71a4e83c4 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26357bcb-33a8-4dbc-bfdd-b147b3c2f91d · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3865526-d003-4d3a-87a4-427909f12ae2 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Concentration inequalities
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation aff7a8f5-7412-4496-aa56-6b6b0e819e8c · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective A., Jagielski, M., Gao, I., Koh, P
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b6abb9b-6002-4b1f-bc8d-23831e701bbe · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Unraveling Arithmetic in Large Language Models: The Role of Algebraic Structures
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ce9f81b-03ea-40de-822c-96b052c9ba27 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective J., and Wong, E
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f7c32822-ae7f-4f98-8649-356701b933ca · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Shifting attention to relevance: Towards the predictive uncertainty quantification of free-form large language models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6d52bdbb-6091-4a04-9b94-c9c396c1bb85 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 11c3cc77-1867-4fc3-a571-02ad49232186 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Kto: Model alignment as prospect theoretic optimization
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c03d6884-bc03-4688-93e7-c340d4c6acd5 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4483b9e3-68ac-4695-af91-6ab4b3630c89 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c1bff9c-f5c7-42b4-8bd8-962cc37df406 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Safety Arithmetic: A Framework for Test-time Safety Alignment of Language Models by Steering Parameters and Activations
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fd99f3d-829e-41e1-9868-6ead4daaa5cf · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Catastrophic jailbreak of open-source llms via exploiting generation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ed44d68c-fa37-40c4-9803-716b36f3dc31 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Exploring Group and Symmetry Principles in Large Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4031ab7e-7bf1-452d-a683-698af3fa2341 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Neural tangent kernel: Convergence and generalization in neural networks
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d32a9366-0ad5-4618-ba3d-25249dbd4d1f · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective LoRA Training in the NTK Regime has No Spurious Local Minima
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e104d32e-07ec-4af8-b8f2-3f2f257402a6 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f7824558-fb1d-4cca-a4be-5a7281e36284 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Decoupling noise and toxic parameters for language model detoxification by task vector merging
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4fb17518-3c4a-4b3e-9f8c-6a4a76fceaf7 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 843d9b8d-22a8-4cee-8ba9-2514ef6e260e · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective A kernel-based view of language model fine-tuning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 53e4cba7-1e6b-4a40-ae76-0f260f625e4b · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Training language models to follow instructions with human feedback
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49939161-8884-44cd-998d-ce38929f7f38 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Visual adversarial examples jailbreak aligned large language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c00176fb-b8bc-4b46-94f9-808b408461e6 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 086a878c-dc55-4f14-b6c0-9f6361320ccb · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4169530e-8513-4cd0-8714-d4c184d8daf7 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective D., Ermon, S., and Finn, C
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70ae71cb-2c0f-4d63-a75f-5e0711fdf588 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b4c462d-e2ae-4a19-8d82-ffe5ad6a3334 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7955819a-b393-4b30-a831-9d2e7c10ce98 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Gemini: A Family of Highly Capable Multimodal Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 059a1817-5705-4b53-b873-b6fe5261a1b9 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Understanding Linear Probing then Fine-tuning Language Models from NTK Perspective
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4bb110e-b0be-44b3-a5d7-ee315513a6c1 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 260ae4dd-9bb6-4ebd-b804-7199c082a674 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Qwen2.5 Technical Report
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47eedbd9-d5c5-4fd1-8b79-3423b730deea · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective On the vulnerability of safety alignment in open-access llms
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 051da380-7ca0-4764-b07a-2060435f37fb · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Large Language Models as Markov Chains
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c455825a-9c42-49be-b81d-cdd36d81e952 · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8951279a-4609-4e0c-a8b0-207e8a98574d · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Towards comprehensive and efficient post safety alignment of large language models via safety patching, 2024
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 97daa6df-ba0d-4f32-81f6-2bc62ef2573d · outbound
Safety Alignment Depth in Large Language Models: A Markov Chain Perspective Emulated Disalignment: Safety Alignment for Large Language Models May Backfire!
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
No inbound Pith citation observations are available.