Pith. sign in

Paper Citation Record · LEDGER

Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2307.08487.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.08487 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:36:57.163093Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T09:54:34.336996Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e9b00a54-4e2e-4a33-a1ba-507582a37895 · inbound

TrustLLM: Trustworthiness in Large Language Models cites this paper.

TrustLLM: Trustworthiness in Large Language Models Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 230

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.661153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:aa19f4847d83ef8107231c2f06796bceb27bca7f2425c80bd0536ca1b1a69348

Observation dd4cd49e-23e4-4f90-a880-a2db55b9b699 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.795100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:e1a18bef091ef883a6207e1c87b64a8f332cab865fb201ddd178ff7dda32c306

Observation f4b5b320-094e-44a5-bbde-cada50d59d0d · inbound

Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models cites this paper.

Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T00:12:04.964160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:12:04.964160Z digest=sha256:9f8834dd0da2b396ee0379dc6c34988a01187cdf58e3b6f5ba9556a3ce99e3e5

Observation 7e07336f-0df7-43cc-8d95-214e4312d24d · inbound

Improving LLM Outputs Against Jailbreak Attacks with Expert Model Integration cites this paper.

Improving LLM Outputs Against Jailbreak Attacks with Expert Model Integration Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T20:36:57.163093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:36:57.163093Z digest=sha256:8b249b89be6c517586eeb588434b1c89aed43e65d49b015454c0f0de0230429c

Observation 4d02dc49-52e6-4edc-bd31-c0d406093cc2 · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:26.838444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:26.838444Z digest=sha256:0dcad903d81278861dd2d9929ed6a4c3c1a4754237667fb701789e62893d3e23

Observation 25a80c37-7277-4d39-807b-dbc2a444fdd8 · inbound

From LLMs to MLLMs to Agents: A Survey of Emerging Paradigms in Jailbreak Attacks and Defenses within LLM Ecosystem cites this paper.

From LLMs to MLLMs to Agents: A Survey of Emerging Paradigms in Jailbreak Attacks and Defenses within LLM Ecosystem Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-15T19:45:09.946166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:45:09.946166Z digest=sha256:0a29288d88b7af641a6af67e23ecdb34d5ca7e47e422b18aae8540e186c739e7

Observation 07c841b8-0bba-451b-8d83-2bd5866d92ff · inbound

Diagnosing Failures in Large Language Models' Answers: Integrating Error Attribution into Evaluation Framework cites this paper.

Diagnosing Failures in Large Language Models' Answers: Integrating Error Attribution into Evaluation Framework Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:23:31.499480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:23:31.499480Z digest=sha256:ec578c487058c4c043e9d7848c5a7d8d67f077d22472fe663791e4038eb08644

Observation cf15a3b9-767c-4e98-9476-c2b2981a7930 · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 149

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:51.557121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:51.557121Z digest=sha256:1f7714148ac08651ae8745661fe598a8977ca9a95dc6a8d3f2aac743bcda6896

Observation 8bbcf1f9-cc39-4b84-866b-3b1ee981fff3 · inbound

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety cites this paper.

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:31:01.017286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-10T16:00:32.413225Z digest=sha256:d558c577f52a31025c822304ca9be9e28419601c1c07e60edb8d28031ba8236d

Observation 00e4c13a-828d-4cae-851c-02ac381015c2 · inbound

Agentic Abstention: Do Agents Know When to Stop Instead of Act? cites this paper.

Agentic Abstention: Do Agents Know When to Stop Instead of Act? Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:54:34.338722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T09:54:07.138157Z digest=sha256:fb0054bd82275351540ca79e3976f51f0dc9d37e724029fcd7c451e01432fa35