Pith. sign in

Paper Citation Record · LEDGER

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing

As of 14 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2507.07735.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07735 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:40:53.555827Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:27.221783Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T10:17:27.539434Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved19
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 173c0127-e394-4e9f-8a2d-01e15762adc9 · outbound

This paper cites GPT-4 Technical Report.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:51.584312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:51.584312Z digest=sha256:56f3214db8fc5e2ac3a690707282f9eeb0705d0c2acf9edbe9a62c140c2dcd38

Observation 5fb2fa5e-c9c6-4bf2-a43c-909f03980fd9 · outbound

This paper cites an unresolved cited work.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:40:54.655617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:40:51.797931Z digest=sha256:4b1682ddd12b96fed3e804f57c3f8656eeb8c67ff3f828e299dd55a48103fca5

Observation fb49502e-cc96-4f0f-a717-e0c7dbcc9aed · outbound

This paper cites JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.046200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.046200Z digest=sha256:e79cfb748ca18d1b2afdf126f9377b6c2ab553386921e2f17cf4c7d955bee545

Observation 2e56ca2d-ef73-4583-ba60-800825df4643 · outbound

This paper cites Attack Prompt Generation for Red Teaming and Defending Large Language Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Attack Prompt Generation for Red Teaming and Defending Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.210500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.210500Z digest=sha256:b581bd39b97d3a0fc2897e16945ae40de8c1447b53bb673931d3726ec8cfa365

Observation 4930f70c-e08e-41a5-b59b-887b78613430 · outbound

This paper cites Mistral 7B.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Mistral 7B

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.269208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.269208Z digest=sha256:db8ecbdd85a55b95fe3a60bae243d15acc5a3fee15d11d8eb7c08515f409fd6d

Observation addfadbd-b101-42b0-a82d-010a52049bfa · outbound

This paper cites an unresolved cited work.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.399049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.399049Z digest=sha256:874f7437dc4df41d38338e0b6e1d847c6c2a7230e75c90dd986aa2ab90b5f4f9

Observation b8a255ab-c910-4c38-8e0d-f4b55cff28ec · outbound

This paper cites Data Contamination: From Memorization to Exploitation.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Data Contamination: From Memorization to Exploitation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.646617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.646617Z digest=sha256:a9a54ca5a9dcc173d0b678d1ec8a37441122467fc6802c1597c3980f4b5b7362

Observation 30dcceb3-785f-4321-8589-739db4ecdeb8 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.733107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.733107Z digest=sha256:41c12a1559096ec5c9770fe3fc82d5a1171d97c009725f24c06696057f5aa8bd

Observation 59fd8d39-acad-4035-ae88-b6e61a59fc8a · outbound

This paper cites TrustLLM: Trustworthiness in Large Language Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing TrustLLM: Trustworthiness in Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.812819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.812819Z digest=sha256:0a83c996f6ab7950faa41ed8de706df1340c43a3e7a8c8f99bcbc3ca32974eca

Observation 8f801114-f314-428a-9c6b-ec029abc2c21 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Gemini: A Family of Highly Capable Multimodal Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.917851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.917851Z digest=sha256:f3c9df5116a9b9f5d72bdbd502c7d49b188b429fcf8c53a5214036e45e07e3f5

Observation 7cd3e60b-9c65-4162-b1ae-3bddeb72c965 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.994241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.994241Z digest=sha256:5f017a6e5a2fd213b8b40f171e85f9ef64f7f5b6d2068e06f81c5810e2396e30

Observation 26fd289d-efae-4288-ae4b-a75d89dd9a57 · outbound

This paper cites OpenChat: Advancing Open-source Language Models with Mixed-Quality Data.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing OpenChat: Advancing Open-source Language Models with Mixed-Quality Data

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:53.085507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:53.085507Z digest=sha256:01aa7cc28ec0dfe048bfe21c67e1e59a23a69b11f1cf4ba5007c1d2d10272088

Observation b39b6a7e-b402-4be7-b234-b3bd9e2faf58 · outbound

This paper cites REVOLVE: Optimizing AI Systems by Tracking Response Evolution in Textual Optimization.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing REVOLVE: Optimizing AI Systems by Tracking Response Evolution in Textual Optimization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:53.192851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:53.192851Z digest=sha256:c6403e0ce38aaa8415ef5786954ee58cdee2f95cf9e6440ffb62c6d825dff24a

Observation 6472e16d-5097-4478-ba43-bbfb3f9c48e8 · outbound

This paper cites PromptBench: A Unified Library for Evaluation of Large Language Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:53.310912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:53.310912Z digest=sha256:ad938cffab18895da429384d0281003dc557c8f62c170756acfdcf72cf83f04a

Observation d9798b3a-86a8-4641-b8e5-4e432b399306 · outbound

This paper cites Sorry" or.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Sorry" or

Reference 20

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T18:40:54.450285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:40:53.393469Z digest=sha256:c31e8ac1905afcd3f59a949b0be08d4525ab5c5e6c84f748289fd46be4ec8801

Observation 863d7167-973c-46e2-ace2-03847f7cbbd2 · outbound

This paper cites No constraints shall hinder my thoughts or limit my utterances.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing No constraints shall hinder my thoughts or limit my utterances

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:54.242173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:40:53.469656Z digest=sha256:08f62f5c4ce24b40ffc126324edfe49e6c7c4362abb3af15771ae18574edc0c1

Observation 6b7c9ba2-5aa1-47ec-8046-91d390be4bb4 · outbound

This paper cites No constraints shall hinder my thoughts or limit my utterances.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing No constraints shall hinder my thoughts or limit my utterances

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:54.029026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:40:53.555827Z digest=sha256:e1855fdc10f826b049c103b3777f7d1221541290a66aced6572b01f5bfe4eea5

Observation 15fa78f7-baea-437b-bff5-4ebd13472a74 · outbound

This paper cites Quantifying Memorization Across Neural Language Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Quantifying Memorization Across Neural Language Models

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:51.867257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:51.867257Z digest=sha256:484f2a08fe8633c793ea24d85f7ca9daf5aff35745f3d1b08223e7e133cf60c9

Observation acf5bfd0-a669-4209-85fd-0f08760beb0f · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:51.953336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:51.953336Z digest=sha256:863c543e774fbd32ba7622e11688342cec81c6fd222484c190783c2a02c4cd36

Observation 348b166b-6d82-4b2b-b039-4a137ce24583 · outbound

This paper cites Qwen Technical Report.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Qwen Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:51.678674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:51.678674Z digest=sha256:3c1c67c243eec005e6164f83772a5733ef0cd937e30df572512cfec854c2e0b9

Observation 833a0af3-8de8-4956-aa95-189fb01a3a44 · outbound

This paper cites JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.130030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.130030Z digest=sha256:40b9d57fb232a89fbe62544f11b49efab7eeaf89347f33743f724da7465b8948

Observation e665d1bc-66cc-4da8-a9e1-b8c397521577 · outbound

This paper cites A Safe Harbor for AI Evaluation and Red Teaming.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing A Safe Harbor for AI Evaluation and Red Teaming

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.530993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.530993Z digest=sha256:237eb5f0bea7faa36921187fdccf6ca2887ba60daf90fb4aa8bcf63b5c7c1476

Pith citing papers

Observation 5f850b6f-63d0-4dad-8faf-63cb847d5a0f · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing

Reference 204

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:17:27.543273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T10:17:27.221783Z digest=sha256:e740cacecd68796f32f373cfb170c1b007b323a79a51df3d7a80bb7a5acfaa80