Pith. sign in

Paper Citation Record · LEDGER

AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2407.17436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.17436 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T11:42:07.741050Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0869e5b3-4271-45b0-9692-6b34579f48e1 · inbound

Breaking Down Bias: On The Limits of Generalizable Pruning Strategies cites this paper.

Breaking Down Bias: On The Limits of Generalizable Pruning Strategies AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-08T11:42:07.741050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:42:07.741050Z digest=sha256:ec2a41fc9f14d68eb7a9742dd7d5b10bd8aacc19bb1e8949c4c316bb376660e0

Observation d7045213-da63-4230-becc-db1ecb4d8590 · inbound

Safety Degradation in AI Agents cites this paper.

Safety Degradation in AI Agents AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:41:05.918299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:41:05.918299Z digest=sha256:cae26025cf056cc5df16e354459bd1b459185df9487206624ee66506855b9321

Observation 7adf61df-b8bb-4260-86ec-bb937963cec4 · inbound

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas cites this paper.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:27.682170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:27.682170Z digest=sha256:1714b9bfa36759cbfa6a62041f4cde1ce4c87ef50bafe4b5ec43907252abf7bb

Observation acbf7256-9e73-4d4b-8725-487d51da02a4 · inbound

MedSentry: Understanding and Mitigating Safety Risks in Medical LLM Multi-Agent Systems cites this paper.

MedSentry: Understanding and Mitigating Safety Risks in Medical LLM Multi-Agent Systems AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:52:03.610676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:52:03.610676Z digest=sha256:fec74f4f01e19fd96a19ffd3ac70d436a1c4e0c5b1f4c67fc4ef444e4feb58a1

Observation 1b7e9ba2-be7e-4c7e-bfca-b8d16364fb28 · inbound

ACCESS DENIED INC: The First Benchmark Environment for Sensitivity Awareness cites this paper.

ACCESS DENIED INC: The First Benchmark Environment for Sensitivity Awareness AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:56.078889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:56:56.078889Z digest=sha256:f6c383e4f88dee899ef39e1eaf2c57e452d26cca335cc10a10d7145a500ff76a

Observation 26835b23-1641-4939-ab42-09bdb134a4a5 · inbound

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures cites this paper.

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:40:20.722684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:40:20.722684Z digest=sha256:1d47d9e44655da23efef2e5234ee5db2b114e525555b4d15922b3ff507890355

Observation 0f2d9e24-efd8-43c9-bf37-44c75b1e2e61 · inbound

Q-resafe: Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models cites this paper.

Q-resafe: Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:19.627162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:00:19.627162Z digest=sha256:de26864b8ade7c40a06cc097bcb8bd91bfc186cbfd2014d3d21be6945411c13e

Observation 2638d8e7-5e47-43ad-9163-e84c4210e807 · inbound

PAC Bench: Do Foundation Models Understand Prerequisites for Executing Manipulation Policies? cites this paper.

PAC Bench: Do Foundation Models Understand Prerequisites for Executing Manipulation Policies? AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:38:51.524460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:38:51.524460Z digest=sha256:201b25551840416b658b78f4d011185a3debd01a165bcfc875a566dab7777ad6

Observation 32bb5059-d424-47b3-a16e-80c5ac345f4a · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 242

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:08.125748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:08.125748Z digest=sha256:dc637dbced107c64b01d2ca4eeb804cc23a8d96ccc793c346c9b1b6a571e3d50

Observation 25b02937-0332-4627-a663-341e2264d9ac · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 230

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:58.221798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:58.221798Z digest=sha256:c76ecef05817214c3833c8673096728f41f3d45ab3404af7efb9495cad7f56a7

Observation fd6579a4-a8d8-4ddf-9357-46a11e763ff8 · inbound

OmniCompliance-100K: A Multi-Domain, Rule-Grounded, Real-World Safety Compliance Dataset cites this paper.

OmniCompliance-100K: A Multi-Domain, Rule-Grounded, Real-World Safety Compliance Dataset AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T11:35:31.636611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T11:32:52.986523Z digest=sha256:9194aa6b2a5c82a5e844a7454e963cabc146bf62552eb258a2f020470bf9b7d2

Observation 93a34e3b-59e6-4fa6-92b9-1d52fe91019f · inbound

Seed1.8 Model Card: Towards Generalized Real-World Agency cites this paper.

Seed1.8 Model Card: Towards Generalized Real-World Agency AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:45:14.326403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T07:44:02.827006Z digest=sha256:655d85dd3072accff87b8b752445b4e2585cc72579376b9a1b402108f6cef22b

Observation f3818a60-a2e4-466c-80fa-cf4cc7520b4a · inbound

A Validated Prompt Bank for Malicious Code Generation: Separating Executable Weapons from Security Knowledge in 1,554 Consensus-Labeled Prompts cites this paper.

A Validated Prompt Bank for Malicious Code Generation: Separating Executable Weapons from Security Knowledge in 1,554 Consensus-Labeled Prompts AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:40:43.456003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T18:11:29.066362Z digest=sha256:518e5b792454ab880351da153dff2d1a085fee4e661991d5690d3262dfbf4d0f

Observation ad4d6a5d-b8e0-490c-9650-c701483c320c · inbound

Beyond Fixed Benchmarks and Worst-Case Attacks: Dynamic Boundary Evaluation for Language Models cites this paper.

Beyond Fixed Benchmarks and Worst-Case Attacks: Dynamic Boundary Evaluation for Language Models AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:06:13.738353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T10:13:35.777910Z digest=sha256:28f5ae05e20822ac56fba51a5b52c8e69e86d84d76d350f13d47a0d6c3bb2cf8

Observation 7ebd75bf-b23f-4b84-97de-3bd1b6993d3f · inbound

ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety cites this paper.

ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:59:45.799831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T04:55:36.069767Z digest=sha256:75be3c0c7ab6fddfc2b28256555f3f706813e9138171d431a252fa2c384f6340

Observation a768578b-b939-42c5-a201-531f761575b0 · inbound

ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety cites this paper.

ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T16:49:32.395383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T16:49:32.395383Z digest=sha256:f4378666b3e20da101353dacfe8ba934f351aea0dfeda28fdf484f041fef3c6d

Observation 361f769d-7269-4c3e-8ecf-29bfd8626531 · inbound

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety cites this paper.

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:51:07.882817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T05:50:28.114140Z digest=sha256:17e494c33b9a1fdac4c21a71a507a41ee21b086914ffc8fbbecc4440cf5a4d68

Observation 5671be64-514b-49dc-a6b5-e5b8aac7edb1 · inbound

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety cites this paper.

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:06:42.863331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T06:05:27.736494Z digest=sha256:dd280ecb2abfc4d228a2fe11b4af5615ce2ecccc8a2baf788fb819c9f89a3a67

Observation 4b7a78d5-fc87-4552-9dfc-399aa26317b1 · inbound

SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation cites this paper.

SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-06-28T17:12:25.159443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T17:05:57.685728Z digest=sha256:81ea24210362f87742ccd8fefc91f3d57309cc1fc07d6ac4aebfa48106aa7c70

Observation d8ecb8ee-7117-4d70-ba5a-7ac8bfe58c4c · inbound

Safety Measurements for Fine-tuned LLMs Should be Grounded in Capability cites this paper.

Safety Measurements for Fine-tuned LLMs Should be Grounded in Capability AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:26:29.809439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T10:00:30.904247Z digest=sha256:c39467918a9e048b0f142cf00378bb33cb525fd76f92b7c6d67fca9aecafd615

Observation a49d69b0-35de-48a1-93c7-5d5d442eaa76 · inbound

Beyond Single-Policy: Evaluating Composed Organization-Specific Policy Alignment in LLM Chatbots cites this paper.

Beyond Single-Policy: Evaluating Composed Organization-Specific Policy Alignment in LLM Chatbots AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-28T05:41:40.158587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T05:41:16.033862Z digest=sha256:d2789ed3ed964cef7223156bedeee8e05576e208d7b09dea6eaf3b467181aeaa

Observation 821c38c6-b39a-4790-b855-e6e2e115e6e1 · inbound

RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-dimensional annotations cites this paper.

RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-dimensional annotations AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:47:26.405643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T18:37:24.395892Z digest=sha256:2db39a884da0d065a71cae5fd3dfd0986c55f8fd08fbcfd86f5d8f0d1b797c63

Observation cd4ad914-2d01-4bd9-8315-73e47c8324af · inbound

Culturally-Adapted Red-Teaming Across East and Southeast Asian Contexts: A Methodological and Comparative Analysis cites this paper.

Culturally-Adapted Red-Teaming Across East and Southeast Asian Contexts: A Methodological and Comparative Analysis AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T01:07:30.328565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T16:48:54.802860Z digest=sha256:294b6d7c2ce5a63b199474581a2037c39e87d3c5bd49e762e7b8fc2e977cf093

Observation b466eda9-ecf2-4ae4-9d68-1757f587d565 · inbound

FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming cites this paper.

FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:09:34.777211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T17:11:40.088809Z digest=sha256:238cef4d40dd715aad7757a0501810dc2b446a0cee5a8ae46b31e531e7db41f3

Observation 3eccc86e-7846-4ba5-a478-6d7a2ec44d1e · inbound

Efficient Safety Benchmarking via Item Response Theory cites this paper.

Efficient Safety Benchmarking via Item Response Theory AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:55:48.974164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T15:51:00.484343Z digest=sha256:feae39887ab55c52f33fe6ebeea8e59e3e81d96cbb58557f3b8e966d018433c4

Observation 7c3bdb28-b686-421f-ac4d-5bb66a412136 · inbound

How Jailbreak Attacks Inform Safety Alignment: A Defender-Centric, Shapley-Based Evaluation of Jailbreak Contributions cites this paper.

How Jailbreak Attacks Inform Safety Alignment: A Defender-Centric, Shapley-Based Evaluation of Jailbreak Contributions AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T18:54:52.928929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:54:52.928929Z digest=sha256:a18b49c4bd4abe9062c819911f1020c778883cb5cd9f8000454dc2939ed28d02

Observation 76cd7933-fe44-421e-b840-c1cef0f98e17 · inbound

AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models cites this paper.

AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T08:30:01.720511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:30:01.720511Z digest=sha256:9566d91e749488b2854c0667beb833bc716f828f56c1e6a17cbe693009b1f942

Observation 2a2bfb26-0fbe-4c53-955a-2c7da7224d28 · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:22.163420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:22.163420Z digest=sha256:adac8345b326f78d2bfdcb64772f350886c63be1da2d91e1c1199e8790a24ab7