Pith. sign in

Paper Citation Record · LEDGER

Expert Survey: AI Reliability & Security Research Priorities

As of 20 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2505.21664.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.21664 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:31:01.571483Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a452e9a7-4d48-4b32-a886-11989e9da89f · outbound

This paper cites RedCode: Risky Code Execution and Generation Benchmark for Code Agents.

Expert Survey: AI Reliability & Security Research Priorities RedCode: Risky Code Execution and Generation Benchmark for Code Agents

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:00.221217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:00.221217Z digest=sha256:a637e94e36b65f502490ec144c61e9b76ab662b9f7503572227451c6ad9d910a

Observation 045969ea-fc04-4026-a3a6-b0924cf06f25 · outbound

This paper cites Self-Destructing Models: Increasing the Costs of Harmful Dual Uses of Foundation Models.

Expert Survey: AI Reliability & Security Research Priorities Self-Destructing Models: Increasing the Costs of Harmful Dual Uses of Foundation Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:00.326166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:00.326166Z digest=sha256:41ec34880a85a8de7408da0f06d7c63450fd5bc96e7a037fe38ab7ac8564a0df

Observation 8054dbf1-e20a-4a50-9820-b18105b5baf2 · outbound

This paper cites Datamodels: Predicting Predictions from Training Data.

Expert Survey: AI Reliability & Security Research Priorities Datamodels: Predicting Predictions from Training Data

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:00.440577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:00.440577Z digest=sha256:d84d6527da3acd36ba9a66d570c235d4017e487cadd0d74b643ab495ce071fcc

Observation f964876c-f82f-49f1-bbce-54718f8a861a · outbound

This paper cites A Watermark for Large Language Models.

Expert Survey: AI Reliability & Security Research Priorities A Watermark for Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:00.570046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:00.570046Z digest=sha256:cc970a065f5a77506925b02271f2b975c0fd21ecceba34778a9883317e08a7d4

Observation be7dc243-3e19-40ce-afd4-da1b88830317 · outbound

This paper cites Towards A Proactive ML Approach for Detecting Backdoor Poison Samples.

Expert Survey: AI Reliability & Security Research Priorities Towards A Proactive ML Approach for Detecting Backdoor Poison Samples

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:31:02.251024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T13:31:00.925304Z digest=sha256:5b2485caf300e3b30c85632b0ea4dcdeff0ae13507776ed8a08f2a4fed71f461

Observation 5d9ed15b-8e2b-4af5-95ef-8d6314215e61 · outbound

This paper cites Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!.

Expert Survey: AI Reliability & Security Research Priorities Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:01.051341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:01.051341Z digest=sha256:4ef66f1205fbb6cdfacf890a572f76cf97c9d895033bbc563a029b795ea1805a

Observation a44dc603-4a85-406b-9742-f1b0b0604356 · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

Expert Survey: AI Reliability & Security Research Priorities Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:01.134755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:01.134755Z digest=sha256:4a85525208f2d2ae96ed27504a611193b79ef8a159a4c662eb4855f2814fdd03

Observation 6a661662-760e-4a70-a8c8-5e48479c992f · outbound

This paper cites Instructional Fingerprinting of Large Language Models.

Expert Survey: AI Reliability & Security Research Priorities Instructional Fingerprinting of Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:01.257973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:01.257973Z digest=sha256:34f975a409553387b4311461bb7da79cf0c5480d9bc9b558f7362f673af8e990

Observation 9aaf56e4-1414-4bb6-8438-5260f479e868 · outbound

This paper cites Low-Resource Languages Jailbreak GPT-4.

Expert Survey: AI Reliability & Security Research Priorities Low-Resource Languages Jailbreak GPT-4

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:01.404747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:01.404747Z digest=sha256:ad235d0ef2f7fed178b37fc21bf8bfa8ea884f2fe424abb19b73a098a1d01da3

Observation 3517e76f-da1f-46e3-a05a-58462570a705 · outbound

This paper cites Towards Fair Disentangled Online Learning for Changing Environments.

Expert Survey: AI Reliability & Security Research Priorities Towards Fair Disentangled Online Learning for Changing Environments

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:31:01.852026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T13:31:01.571483Z digest=sha256:b9903c660805b283a7e53ff5dac32638c9b5692020e18f1d7f5393e92e8a4688

Observation 9f7d25f7-eb65-4832-8dfd-019e90df399e · outbound

This paper cites Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory.

Expert Survey: AI Reliability & Security Research Priorities Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:00.723766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:00.723766Z digest=sha256:159f5b7c3f2cd72c37cf3b150c255a7391098ad49ee108071f0aed64917fd4b5

Observation 48af103d-8387-4731-99e7-4398642e8d7c · outbound

This paper cites What Does it Mean for a Language Model to Preserve Privacy?.

Expert Survey: AI Reliability & Security Research Priorities What Does it Mean for a Language Model to Preserve Privacy?

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:00.052533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:00.052533Z digest=sha256:55c575c821190d4ad1ff6ac9b7e60e7bf16a6bd3c0f39d31c8b0fb4da463bd3d

Observation ffbf3964-a49e-4ebe-81a3-8408c304d65c · outbound

This paper cites Deep reinforcement learning from human preferences.

Expert Survey: AI Reliability & Security Research Priorities Deep reinforcement learning from human preferences

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:00.121297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:00.121297Z digest=sha256:8e4301f84f3b41aa5fe6cb97b6b72dd497e5766ba2fc95e550defa1d08eeb548

Observation c5afd9aa-06f8-48d6-ac44-3e8ea27f9f93 · outbound

This paper cites Scaling Adversarial Training to Large Perturbation Bounds.

Expert Survey: AI Reliability & Security Research Priorities Scaling Adversarial Training to Large Perturbation Bounds

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T13:31:02.696598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T13:30:59.961792Z digest=sha256:27c1442a6892ab86107f61f27d60852f5ec50a33df206b02d5e4d0620a1c813b

Observation 142d8b50-4955-45f4-bbce-67f0433e18e2 · outbound

This paper cites Causal Fairness Analysis.

Expert Survey: AI Reliability & Security Research Priorities Causal Fairness Analysis

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:00.823362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:00.823362Z digest=sha256:c6828adb41f72a1290506fa23b29ae579e76b823bd180cde1e189ee0a18a683d

Pith citing papers

No inbound Pith citation observations are available.