Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T23:43:07.803719Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 6 inbound Pith citation observations for arXiv:2602.12966.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T23:43:07.803719Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-14T06:09:46.933171Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-03T14:28:31.386601Z
12 of 12 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d7db0cbe-6faa-4159-83b1-d5822ad74e8d · outbound
ProbeLLM: Automating Principled Diagnosis of LLM Failures URL https: //aclanthology.org/2025.acl-long.17/
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cf49950-f877-4f14-98ac-021460384fb4 · outbound
ProbeLLM: Automating Principled Diagnosis of LLM Failures Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02beb234-b5d3-47f7-bdcd-86b5ceee21f2 · outbound
ProbeLLM: Automating Principled Diagnosis of LLM Failures gpt-oss-120b & gpt-oss-20b Model Card
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62c7fc2c-a713-4d3c-9f1f-d8ae70d65ebf · outbound
ProbeLLM: Automating Principled Diagnosis of LLM Failures DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b38479a-5fe6-48ff-b489-5aa0bc94c903 · outbound
ProbeLLM: Automating Principled Diagnosis of LLM Failures Discovering Knowledge Deficiencies of Language Models on Massive Knowledge Base
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e203e4a-8339-481e-8f50-b0f2ab11aa5e · outbound
ProbeLLM: Automating Principled Diagnosis of LLM Failures Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2578719-f5dc-470c-9026-44b8627cdff5 · outbound
ProbeLLM: Automating Principled Diagnosis of LLM Failures org/CorpusID:15184765
Reference 2006
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 746a367d-6d75-4b97-850f-e419c3b57eb1 · outbound
ProbeLLM: Automating Principled Diagnosis of LLM Failures A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35f6cc96-e0d7-4b2e-b2c7-488aac151725 · outbound
ProbeLLM: Automating Principled Diagnosis of LLM Failures Toward Automated Robustness Evaluation of Mathematical Reasoning
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7328d18-8700-4e10-9f7c-525745f77d99 · outbound
ProbeLLM: Automating Principled Diagnosis of LLM Failures Efficient Process Reward Model Training via Active Learning
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af7a3fd0-9c46-43f7-bb56-7917466a3894 · outbound
ProbeLLM: Automating Principled Diagnosis of LLM Failures Program Synthesis with Large Language Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfca5e10-3479-4c8a-821a-b05cb32972c2 · outbound
ProbeLLM: Automating Principled Diagnosis of LLM Failures org/CorpusID:277452302
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72ccd56b-2c5d-483f-84e1-aabe6c4cb8b2 · inbound
Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs ProbeLLM: Automating Principled Diagnosis of LLM Failures
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1e1b4f2a-747c-4ddf-9282-4b6491c86229 · inbound
Chain of Risk: Safety Failures in Large Reasoning Models and Mitigation via Adaptive Multi-Principle Steering ProbeLLM: Automating Principled Diagnosis of LLM Failures
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 835d23ac-818a-45cb-b47f-0dbb2c9695c1 · inbound
FailureScope: Cross-Regime Behavioral Diagnosis of Language Model Weaknesses ProbeLLM: Automating Principled Diagnosis of LLM Failures
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1b007848-5601-4196-97d2-e454efe05a34 · inbound
Getting Better at Working With You: Compiling User Corrections into Runtime Enforcement for Coding Agents ProbeLLM: Automating Principled Diagnosis of LLM Failures
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d69426c1-14cb-4eb9-bfc8-27e268cffb13 · inbound
DeepBias: Adaptive In-depth Probing of Social Biases in LVLMs ProbeLLM: Automating Principled Diagnosis of LLM Failures
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d4d1f21-7196-433c-a1f7-36cb705dc80b · inbound
Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias ProbeLLM: Automating Principled Diagnosis of LLM Failures
Reference 141
Source-reported events for the cited work
Unavailable: canonical work link unavailable.