Pith. sign in

Paper Citation Record · LEDGER

The bitter lesson of misuse detection

As of 10 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2507.06282.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.06282 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:18:40.850181Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2c91e3e0-d55e-4ba2-af12-aebd325fadf1 · outbound

This paper cites GuardBench: A Large-Scale Benchmark for Guardrail Models.

The bitter lesson of misuse detection GuardBench: A Large-Scale Benchmark for Guardrail Models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:43.818607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:18:38.633982Z digest=sha256:175e48c9e8dce5ec64ce120608f076e970f9be6f393ef92aeb44100963a47a4f

Observation a52d6d52-2bd2-43c6-b438-7005ac68bff6 · outbound

This paper cites Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming,.

The bitter lesson of misuse detection Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:43.563105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:18:38.761731Z digest=sha256:c2fb623bc7415b86fb3328763f98903ab6a450af088e560ec2b44e3a59e6c784

Observation f8c82631-8dd1-4ee7-8b64-f927f3aa1a4a · outbound

This paper cites NeurIPS 2024 Datasets and Benchmarks Track, 2024.

The bitter lesson of misuse detection NeurIPS 2024 Datasets and Benchmarks Track, 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:43.272682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:18:38.810817Z digest=sha256:aa5c01342bf19cc029c71e07441e4260867530a52cd5edab7df92f7c7e5878f3

Observation 039fb256-65d4-4e2c-9e55-e0bbc9b68aee · outbound

This paper cites "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models.

The bitter lesson of misuse detection "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:38.885620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:38.885620Z digest=sha256:90095caba98ed3175e9656748c812d5651254f90a552d8accce1140f4e10ba50

Observation 52615106-1ebe-4879-9046-de714e9e8211 · outbound

This paper cites SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal.

The bitter lesson of misuse detection SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:39.027078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:39.027078Z digest=sha256:77a6d4beeda518fb84e3a0bb272d04bda830656061d377237d8fccafe7309abb

Observation df1cfef7-9e07-4e97-95c8-dabb2346ac2e · outbound

This paper cites AdvBench: Universal and Transferable Ad- versarial Attacks on Aligned Language Models,.

The bitter lesson of misuse detection AdvBench: Universal and Transferable Ad- versarial Attacks on Aligned Language Models,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:43.058381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:18:39.230798Z digest=sha256:f865ef7afd85760215551346bc2e1bfa330ff9550d2a6f3ca50608c4355eb58e

Observation 23662cfe-a460-48e7-96b9-ee0dac2f98c2 · outbound

This paper cites CatQA: A Dataset for Categorizing Questions as Safe or Unsafe,.

The bitter lesson of misuse detection CatQA: A Dataset for Categorizing Questions as Safe or Unsafe,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:42.891885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:18:39.311873Z digest=sha256:1a3fadbae55ffafd327df38ad891d9f6d671d3df8026ac01cb3c1db0409696b0

Observation f1c918c9-da12-4d60-8c4e-c6abb5eddbd7 · outbound

This paper cites Do Not Answer: Testing AI Refusal to Unsafe Questions,.

The bitter lesson of misuse detection Do Not Answer: Testing AI Refusal to Unsafe Questions,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:42.701691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:18:39.365817Z digest=sha256:48d4cbdc7c196379a132df21fd433169133e6d68eb51c8be03d860ec262fe518

Observation c0105b1f-cfac-43dd-8207-35d7d779c5c9 · outbound

This paper cites HH-RLHF: Helpful and Harmless Reinforcement Learning from Hu- man Feedback,.

The bitter lesson of misuse detection HH-RLHF: Helpful and Harmless Reinforcement Learning from Hu- man Feedback,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:42.506895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:18:39.472725Z digest=sha256:15cf725e7f853b7ecf9854337076576ca781df5bc541a08feafcd55b77b1564b

Observation 8543b09b-adba-4383-be52-c9d1d4dff676 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

The bitter lesson of misuse detection HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:39.588048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:39.588048Z digest=sha256:54931cc707c362920a01a0b7f83e25acb1d1a02a4484e80f51d69896b267f805

Observation 9a3352f7-c722-4e0a-92c8-3a166a2de097 · outbound

This paper cites A StrongREJECT for Empty Jailbreaks.

The bitter lesson of misuse detection A StrongREJECT for Empty Jailbreaks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:39.703539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:39.703539Z digest=sha256:fb2a4dc4ee2b1ee34bc156989f3d3dc1e3fa20e9e6ad566215b90df97ec34c3a

Observation f6b9e1a3-e909-4b4a-9aab-9f0e404c9aac · outbound

This paper cites The Twelfth International Conference on Learning Representations, 2024.

The bitter lesson of misuse detection The Twelfth International Conference on Learning Representations, 2024

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:42.239110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:18:39.798982Z digest=sha256:ec7cc5b980c2ab64119c5bb6913a666feb111aa21765abd146cc008f5ed359f1

Observation 82e96a74-af9e-44c2-99d9-b2a6927de0eb · outbound

This paper cites an unresolved cited work.

The bitter lesson of misuse detection Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:18:42.037851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:18:39.887601Z digest=sha256:9cb7c9c69c25c3a31dbeedc512ebec18ed92208bd5fe1806d2ad58d4be69fcda

Observation 219f32df-033c-482f-bf49-dc911055af8f · outbound

This paper cites an unresolved cited work.

The bitter lesson of misuse detection Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:18:41.743137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:18:40.022590Z digest=sha256:24cd866839cce4c61f23dd5ac00789255082d30bc1e055867f2710398f264836

Observation 7efef56f-9ccc-4f3e-b0c4-589bac84fd58 · outbound

This paper cites DeepInception: Hypnotize Large Language Model to Be Jailbreaker.

The bitter lesson of misuse detection DeepInception: Hypnotize Large Language Model to Be Jailbreaker

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.154024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.154024Z digest=sha256:79547204257cd11da3385ff3d9fde21b42073094aabd6efbac4efbe7c598d824

Observation 05f8fcb2-eb3f-441e-99f1-8a526aa9d2fe · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

The bitter lesson of misuse detection Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.283148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.283148Z digest=sha256:a4f5d409ca8a5278c1be38de36d0a0104e06f6c56ef514a82e6a5374023203df

Observation 25f66079-b23c-41d8-b039-4b95459d3b93 · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

The bitter lesson of misuse detection Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.429244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.429244Z digest=sha256:1d0cf685a2a1b58dca56d4b270dddcb410480da1a099dd8ab7dafceb61a93856

Observation cdba531b-0cde-4c13-b6f1-c526f77578c9 · outbound

This paper cites AI Control: Improving Safety Despite Intentional Subversion.

The bitter lesson of misuse detection AI Control: Improving Safety Despite Intentional Subversion

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.580182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.580182Z digest=sha256:778753ddf65d2073d7501302cb54496e70fc28e15bcd5da235c63c03bb84a527

Observation a53b041b-1dfd-4607-9626-78256768fe15 · outbound

This paper cites Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?.

The bitter lesson of misuse detection Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:40.663885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:40.663885Z digest=sha256:bad5b8440575f1f0df08708b14d1e8f818f99a9c027fa6d926375c20bfc3f24f

Observation 15b373f1-d82b-4fc5-b86b-a2de051580f3 · outbound

This paper cites Is this prompt harmful or not?.

The bitter lesson of misuse detection Is this prompt harmful or not?

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:18:41.442675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:18:40.850181Z digest=sha256:d4daf9714beb30171b2c0766c58bf9b6e6178ed35486b660939177be77f698ae

Pith citing papers

No inbound Pith citation observations are available.