Pith. sign in

Paper Citation Record · LEDGER

LLM Safeguard is a Double-Edged Sword: Exploiting False Positives for Denial-of-Service Attacks

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2410.02916.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.02916 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:34:34.773621Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c5a11f2c-14cd-46db-93b6-da81bad6b178 · inbound

A Survey of Attacks on Large Language Models cites this paper.

A Survey of Attacks on Large Language Models LLM Safeguard is a Double-Edged Sword: Exploiting False Positives for Denial-of-Service Attacks

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-15T20:34:34.773621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:34:34.773621Z digest=sha256:3dbb8a0674a00a7c9c7d77077cde78f381d56529e2dc5522dbcb1d475684d41d

Observation 1bfdab27-4c2d-41c8-84d2-028ba9d1282e · inbound

VSF-Med:A Vulnerability Scoring Framework for Medical Vision-Language Models cites this paper.

VSF-Med:A Vulnerability Scoring Framework for Medical Vision-Language Models LLM Safeguard is a Double-Edged Sword: Exploiting False Positives for Denial-of-Service Attacks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:01:04.693015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:01:04.693015Z digest=sha256:05bea12357c557a1dc963df53c83e62d6adcf61538a548a452d13bcf8ae0d20e

Observation e571f1f1-dc34-4f77-8c5b-1ef53434b618 · inbound

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations cites this paper.

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations LLM Safeguard is a Double-Edged Sword: Exploiting False Positives for Denial-of-Service Attacks

Reference 194

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:17:56.197336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-14T20:13:10.814899Z digest=sha256:b26590349c0494b792512977cbad4c748f49f188eb954d9e7680f1573e697af0

Observation 8c91eae2-fdca-455c-ba8d-79691f935cb9 · inbound

Adversarial Prompts for Acceptance Collapse in Speculative Decoding cites this paper.

Adversarial Prompts for Acceptance Collapse in Speculative Decoding LLM Safeguard is a Double-Edged Sword: Exploiting False Positives for Denial-of-Service Attacks

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T06:43:35.459929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:43:35.459929Z digest=sha256:f12f16856908778ace4125d4269c839c72801818e1aaa0e80186adab0b88c2ac

Observation 93dfcea6-fce9-4c60-a72f-fdb0c5b819ec · inbound

The Boy Who Cried Wolf: Adversarial Misclassification of Safe Inputs as Unsafe in Multimodal Guardrails cites this paper.

The Boy Who Cried Wolf: Adversarial Misclassification of Safe Inputs as Unsafe in Multimodal Guardrails LLM Safeguard is a Double-Edged Sword: Exploiting False Positives for Denial-of-Service Attacks

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T00:20:22.387882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:20:22.387882Z digest=sha256:0582c11bf172bc5c743559948936540e217716e0fad6b4a9d1897a63205dde99