Pith. sign in

Paper Citation Record · LEDGER

A Holistic Approach to Undesired Content Detection in the Real World

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2208.03274.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2208.03274 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:28:43.934564Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T13:06:58.683371Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2bba94ca-935b-4b55-abb2-43ad148bad9f · inbound

Ignore Previous Prompt: Attack Techniques For Language Models cites this paper.

Ignore Previous Prompt: Attack Techniques For Language Models A Holistic Approach to Undesired Content Detection in the Real World

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:59:31.495804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T13:59:31.213830Z digest=sha256:1111614a606df4240cf4c84370123d8fde6aa64c7754328a913941893f88a2d1

Observation 7571bf72-6d0d-4cc0-8123-3656d96f888e · inbound

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models cites this paper.

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models A Holistic Approach to Undesired Content Detection in the Real World

Reference 128

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T06:38:36.858913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T06:38:36.517935Z digest=sha256:e864ea36fd45d9f5dac22af2299da8ba9bedabd9468501212b9c4f84456a7e51

Observation fc3b47ae-d489-4c17-a7db-e115376c280c · inbound

ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction cites this paper.

ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction A Holistic Approach to Undesired Content Detection in the Real World

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:37:15.936802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T11:34:09.428653Z digest=sha256:14ceb5211c7613bbd08d26a71fc71c9cc96daaee6149e1725c347180446e2dfb

Observation 95f93033-9560-45a8-bcae-3076a86fa639 · inbound

SEALGuard: Safeguarding the Multilingual Conversations in Southeast Asian Languages for LLM Software Systems cites this paper.

SEALGuard: Safeguarding the Multilingual Conversations in Southeast Asian Languages for LLM Software Systems A Holistic Approach to Undesired Content Detection in the Real World

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:28:43.934564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:28:43.934564Z digest=sha256:6911009e09548adda22cadd678af92624667c2035bb382521301970d2799a041

Observation 17a36167-b357-4480-bee9-490901646137 · inbound

Guard Vector: Beyond English LLM Guardrails with Task-Vector Composition and Streaming-Aware Prefix SFT cites this paper.

Guard Vector: Beyond English LLM Guardrails with Task-Vector Composition and Streaming-Aware Prefix SFT A Holistic Approach to Undesired Content Detection in the Real World

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T14:50:30.992141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:50:30.992141Z digest=sha256:2b700df233fd1c4d6c07d54931c705fb57c4aa58aa4ca59167f9d7f02ea2b7f5

Observation 35db1525-7155-47e0-9ca4-3ac1bf0c7fa9 · inbound

GuardReasoner-Omni: A Reasoning-based Multi-modal Guardrail for Text, Image, Video, and Audio cites this paper.

GuardReasoner-Omni: A Reasoning-based Multi-modal Guardrail for Text, Image, Video, and Audio A Holistic Approach to Undesired Content Detection in the Real World

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T05:07:09.386730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:07:09.386730Z digest=sha256:631055fc92173f2ee04c087c639221b3f5a66108dbe7384c3e41b3b96a6b1529

Observation dc21f6c3-6724-43c2-8c5f-32cef1d4e304 · inbound

Response-Based Knowledge Distillation for Multilingual Jailbreak Prevention Unwittingly Compromises Safety cites this paper.

Response-Based Knowledge Distillation for Multilingual Jailbreak Prevention Unwittingly Compromises Safety A Holistic Approach to Undesired Content Detection in the Real World

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:28:48.799684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T01:27:16.967080Z digest=sha256:5a9500969614c5218ae6af86e0a2eddd47e1bbe7cb6518a80b80101626bb6d11

Observation 98b44a5c-3ee8-4942-9058-04915b276034 · inbound

Predict, Don't React: Value-Based Safety Forecasting for LLM Streaming cites this paper.

Predict, Don't React: Value-Based Safety Forecasting for LLM Streaming A Holistic Approach to Undesired Content Detection in the Real World

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-13T11:43:31.089245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T11:43:31.089245Z digest=sha256:012d14dfa14b0474b03c7b20f743060f238e92e0452687bfa548368266e5d983

Observation 945b456b-7538-433d-b7c0-7265147d2d27 · inbound

VoxSafeBench: Not Just What Is Said, but Who, How, and Where cites this paper.

VoxSafeBench: Not Just What Is Said, but Who, How, and Where A Holistic Approach to Undesired Content Detection in the Real World

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:24:22.129507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T10:19:28.041282Z digest=sha256:00af01f39d9a753370d150810e346b75a19758d0daeaebd87fa06b5a10f44208

Observation d36ec0b3-12a0-4029-a2a9-def6c38c666e · inbound

Test-Time Safety Alignment cites this paper.

Test-Time Safety Alignment A Holistic Approach to Undesired Content Detection in the Real World

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:51:43.314183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T16:06:37.244288Z digest=sha256:9618cbe75e98cca5cc83d446edd9738a68ed24ab09954b93f18cd2ba76590cf1

Observation b8c0d322-4c2c-4b4a-9de7-719b014adb7b · inbound

AI Content Moderation in Therapy Conversations cites this paper.

AI Content Moderation in Therapy Conversations A Holistic Approach to Undesired Content Detection in the Real World

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-29T21:03:58.416947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T21:02:26.865075Z digest=sha256:36f3343f3cde129ca736f1929410543405c040481c2fcb788958183e0ea2e068

Observation 05c10576-7de2-48e1-9a74-96b54911bc59 · inbound

$D^2$-Monitor: Dynamic Safety Monitoring for Diffusion LLMs via Hesitation-Aware Routing cites this paper.

$D^2$-Monitor: Dynamic Safety Monitoring for Diffusion LLMs via Hesitation-Aware Routing A Holistic Approach to Undesired Content Detection in the Real World

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:03:59.808267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T22:03:25.762772Z digest=sha256:609f258e44153e82bc5f169a1d6583dfeed85ca2e8c5f8c073c487cd54728cfe

Observation 95c790ca-7fc2-4ab8-9e9c-0e341194407e · inbound

Learning from Mistakes: Can LLM Self-Recover after Misalignment? cites this paper.

Learning from Mistakes: Can LLM Self-Recover after Misalignment? A Holistic Approach to Undesired Content Detection in the Real World

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-13T18:51:10.298187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T18:51:10.298187Z digest=sha256:2857ed30285de69f2dae7c9a80461be01998753e7cf2815961c4694dd4d8bd3d

Observation a5eec1c6-b462-4521-b7db-96e4f953b050 · inbound

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety cites this paper.

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety A Holistic Approach to Undesired Content Detection in the Real World

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:52:55.802055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T00:46:03.210076Z digest=sha256:714a3b7322234fe23eacd8ed83efe93a37aa2c96f11ad36a7c3dd639d7bd817b

Observation 4a973d41-9d02-47d2-a9bf-a133257b3289 · inbound

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models cites this paper.

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models A Holistic Approach to Undesired Content Detection in the Real World

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-01T17:35:51.371484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T03:35:34.594617Z digest=sha256:45331314e7f34f991a0393918043f1c628b6ba3bc758e784b6f958de0772d8c5

Observation 425d607f-c7d4-44b5-9195-0eaf66fa9f7d · inbound

AI Native Games: A Survey and Roadmap cites this paper.

AI Native Games: A Survey and Roadmap A Holistic Approach to Undesired Content Detection in the Real World

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:06:58.684741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-02T12:59:20.908659Z digest=sha256:a8e3cf0fc2b9126fa30853ffadc21137f6147dd5669a9c16567dd6e08ced417b

Observation f3d89b71-7b73-46ad-abc0-ce6be5bdc526 · inbound

AI Native Games: A Survey and Roadmap cites this paper.

AI Native Games: A Survey and Roadmap A Holistic Approach to Undesired Content Detection in the Real World

Reference 89

Resolution
unresolved
no resolver link, observed 2026-07-12T09:30:07.729486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:30:07.729486Z digest=sha256:36cac22b110ed48544690794f09f9ba6b4cd67991e5a8aa319f87c1497abf56d