Pith. sign in

Paper Citation Record · LEDGER

Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2307.08487.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.08487 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:36:57.163093Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T09:54:34.336996Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e9b00a54-4e2e-4a33-a1ba-507582a37895 · inbound

TrustLLM: Trustworthiness in Large Language Models cites this paper.

TrustLLM: Trustworthiness in Large Language Models Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 230

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.661153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:2601b9780bb98bd32c692577c901a2b39d18a37d16f1805d676ad3af4ff4e8ac

Observation dd4cd49e-23e4-4f90-a880-a2db55b9b699 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.795100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:20b7628bfa382b33d23b511e01924157bab83317b428576ca85ef73a4ee88741

Observation f4b5b320-094e-44a5-bbde-cada50d59d0d · inbound

Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models cites this paper.

Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T00:12:04.964160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:12:04.964160Z digest=sha256:72cd0ad244e8614928cfe6b6d21c21e38d58533f07881d581a0b5716c564c606

Observation 7e07336f-0df7-43cc-8d95-214e4312d24d · inbound

Improving LLM Outputs Against Jailbreak Attacks with Expert Model Integration cites this paper.

Improving LLM Outputs Against Jailbreak Attacks with Expert Model Integration Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T20:36:57.163093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:36:57.163093Z digest=sha256:518e2e6d992633ad0fe8366c93f145adf2cb99648b9457d658052b33259ca600

Observation 4d02dc49-52e6-4edc-bd31-c0d406093cc2 · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:26.838444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:26.838444Z digest=sha256:e5eb5d3e9b6b9e078884fd7cb26425e942f087ff27f9802c71f376c398e10c0c

Observation 25a80c37-7277-4d39-807b-dbc2a444fdd8 · inbound

From LLMs to MLLMs to Agents: A Survey of Emerging Paradigms in Jailbreak Attacks and Defenses within LLM Ecosystem cites this paper.

From LLMs to MLLMs to Agents: A Survey of Emerging Paradigms in Jailbreak Attacks and Defenses within LLM Ecosystem Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-15T19:45:09.946166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:45:09.946166Z digest=sha256:84ace53935c1d770c9e0116230573bf8f0afee18f93df43c9423574a2310e75b

Observation 07c841b8-0bba-451b-8d83-2bd5866d92ff · inbound

Diagnosing Failures in Large Language Models' Answers: Integrating Error Attribution into Evaluation Framework cites this paper.

Diagnosing Failures in Large Language Models' Answers: Integrating Error Attribution into Evaluation Framework Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:23:31.499480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:23:31.499480Z digest=sha256:eaa661b85f9da58f613fc326fb102089286aacdeb14c4e1026cf4d7f42f4d1ed

Observation cf15a3b9-767c-4e98-9476-c2b2981a7930 · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 149

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:51.557121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:51.557121Z digest=sha256:7db33061ad5f14eb5bd977a1f75acccdff6b2ad2b41ad3a117b5232e082843d5

Observation 8bbcf1f9-cc39-4b84-866b-3b1ee981fff3 · inbound

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety cites this paper.

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:31:01.017286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-10T16:00:32.413225Z digest=sha256:d8b8668dbfff1ff74220f188fc7b1a0b4cbe5618acd1610f172dd0c65e5f3231

Observation 00e4c13a-828d-4cae-851c-02ac381015c2 · inbound

Agentic Abstention: Do Agents Know When to Stop Instead of Act? cites this paper.

Agentic Abstention: Do Agents Know When to Stop Instead of Act? Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:54:34.338722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T09:54:07.138157Z digest=sha256:b75f82a13746f4dc441f87587d872b272e01411fac1aa4eb357aed9ba02be865