Pith. sign in

Paper Citation Record · LEDGER

Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2301.12867.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2301.12867 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:26.664466Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T09:56:10.738322Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2e818d8e-8903-4eec-9f73-59b3fd164ccd · inbound

Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment cites this paper.

Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 219

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:30:45.065301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T22:30:44.520703Z digest=sha256:883418588013f6ab49d4b636b29486832d53e9727368fc7235358e7872b2a0df

Observation 560fcd23-4622-413e-ae6f-a792a3e6af88 · inbound

Low-Resource Languages Jailbreak GPT-4 cites this paper.

Low-Resource Languages Jailbreak GPT-4 Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.149204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:94755788c04e246c5149648ba64e696be9cc3b423baeb8b59241dc55ab302a28

Observation 696889d0-2eb3-4288-a00b-8581de84e9be · inbound

AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models cites this paper.

AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-12T16:28:04.043903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T16:28:03.996446Z digest=sha256:3278bdc53cee5b07eeaea0ec7c79793f9201bcb1ff413093faa29ac7aa6f6ab2

Observation 031bf682-0173-4977-a4da-457f85d2e75c · inbound

StarCoder 2 and The Stack v2: The Next Generation cites this paper.

StarCoder 2 and The Stack v2: The Next Generation Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-12T17:28:22.996594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T17:28:22.353355Z digest=sha256:f68730a75028de473b15f89a346b428a4030c7a776fd93ca1ffb2562836b472f

Observation 50d2b646-d940-4f5e-8a14-8d6c4ee8a48f · inbound

Enhancing Instructional Quality: Leveraging Computer-Assisted Textual Analysis to Generate In-Depth Insights from Educational Artifacts cites this paper.

Enhancing Instructional Quality: Leveraging Computer-Assisted Textual Analysis to Generate In-Depth Insights from Educational Artifacts Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-24T03:08:48.332077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-24T03:07:32.103513Z digest=sha256:68f9d1d06b0cf887e9da8eaab54e88ad72d7a88ec39ebbffec4fb92efb528e14

Observation 28aae478-56e3-40a4-9a7e-9327a4d988f8 · inbound

Faster-GCG: Efficient Discrete Optimization Jailbreak Attacks against Aligned Large Language Models cites this paper.

Faster-GCG: Efficient Discrete Optimization Jailbreak Attacks against Aligned Large Language Models Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T19:03:21.370735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T19:01:43.922848Z digest=sha256:2c7ce9bc44bd7b53b5203c227b43a1c7b2b323c94fe2e221dea32398ae249bd8

Observation 0b38b1c7-e7b5-4896-a37d-89640c154386 · inbound

AI Failures in the Eyes of the Downstream Developer: A First Look at Concerns, Practices, and Challenges cites this paper.

AI Failures in the Eyes of the Downstream Developer: A First Look at Concerns, Practices, and Challenges Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 124

Resolution
verified exact
arxiv_id, observed 2026-05-22T23:15:13.116836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T23:13:37.570261Z digest=sha256:879cc910b0fc0c9fa4e97068d1c9a631c8e942ed4d24581f75939ff504a537dd

Observation 7f8586df-6d98-4eec-b025-d991dda3ccc3 · inbound

Sword and Shield: Uses and Strategies of LLMs in Navigating Disinformation cites this paper.

Sword and Shield: Uses and Strategies of LLMs in Navigating Disinformation Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 136

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:43.921970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:42:43.921970Z digest=sha256:1c0712dcab283445168378bb9dafb9f3a325bd8e928a8ce9b5d0b2fb7c43f7ed

Observation cef0aeb0-73d6-4b10-8aaa-dca886e3c713 · inbound

The Role of Generative AI in Facilitating Social Interactions: A Scoping Review cites this paper.

The Role of Generative AI in Facilitating Social Interactions: A Scoping Review Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:50.751604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:19:50.751604Z digest=sha256:4fa6b599dc26026bd802b9a1bea311cb7e35bccec12725d7de38c14279c8f945

Observation 0585a788-d1a7-4811-99c0-f81da675aa71 · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:26.664466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:26.664466Z digest=sha256:4e929811a78ea34b0144504138ed58c0672b1855092d91fb1d3ca1e7d7dfda7b

Observation 1195accf-c8bc-4b2d-bb52-4914f0289ad4 · inbound

MAGPIE: A dataset for Multi-AGent contextual PrIvacy Evaluation cites this paper.

MAGPIE: A dataset for Multi-AGent contextual PrIvacy Evaluation Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:24.545873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:24.545873Z digest=sha256:59f9e6cd8f4ed6d1f742c09924dc8315de4526a5e60a1c4e3c1b3681942fe6f3

Observation 50dd944f-7d48-4640-9169-4d9797f47acc · inbound

Obscured but Not Erased: Evaluating Nationality Bias in LLMs via Name-Based Bias Benchmarks cites this paper.

Obscured but Not Erased: Evaluating Nationality Bias in LLMs via Name-Based Bias Benchmarks Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:35.296938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:35.296938Z digest=sha256:6fba5f5ce2be23fda8cbb198909130a21cc26af39e2e5d1344fd4a614e2cdfd7

Observation bff97b61-0538-4a17-9183-341fb3eadc0c · inbound

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems cites this paper.

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T17:46:13.399840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:46:13.399840Z digest=sha256:3083e9bde1404a231ac633b32b42cf6e1a685e85d366e3aa482da85d72a4edd3

Observation 0607fe57-a789-47a5-b011-e00fb04ca2c3 · inbound

Reasoning Structure Matters for Safety Alignment of Reasoning Models cites this paper.

Reasoning Structure Matters for Safety Alignment of Reasoning Models Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:56:03.877901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T02:36:57.093584Z digest=sha256:706124472411c468b95064e9ea71bc93f086b02a6edba2206080f5a02c03ca3a

Observation bf04b061-3e62-494e-bd5c-14eb71141f3a · inbound

Intersectional Fairness in Large Language Models cites this paper.

Intersectional Fairness in Large Language Models Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:01:06.636216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T23:44:08.610255Z digest=sha256:83780a35fa7fe1e147e1c1ae315c2111895a255ee9ab8c39e5e176e025cc4c44

Observation d50d4256-bd2f-40fd-9c41-7fac97595747 · inbound

Ethics Testing: Proactive Identification of Generative AI System Harms cites this paper.

Ethics Testing: Proactive Identification of Generative AI System Harms Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:56:08.168691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T20:49:22.147548Z digest=sha256:fda23aecea1ff03062ebd0e70e166d092abd90cb34c7bfd517b1aa5c0c2fa909

Observation c3318292-3ef8-497b-910b-ebd07401afae · inbound

Creating and Evaluating K-12 GenAI Assessment Graders Through Context Engineering cites this paper.

Creating and Evaluating K-12 GenAI Assessment Graders Through Context Engineering Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:35:46.963988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-30T22:54:03.054871Z digest=sha256:35063f80d82dca600f25f1f368e14290d9fef798730918a0105be630999130c2

Observation 743495a7-ea60-42e7-8269-87b2fede425a · inbound

FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming cites this paper.

FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:09:34.823496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T17:11:40.088809Z digest=sha256:80cdaafb53fb892ff32a625e81b34eb234da8b95da495ee9fa4363498e08f4b2

Observation 3ac2261b-d659-484b-982d-5bceaae7cf07 · inbound

A Survey of Toxicity Detection and Mitigation Strategies for Multilingual Language Models cites this paper.

A Survey of Toxicity Detection and Mitigation Strategies for Multilingual Language Models Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T19:30:07.653056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-25T21:16:36.392606Z digest=sha256:2803b9f8704c24a9132efbc1dafed92c333b520d3ea08b67c0707bce19af03a2

Observation 1835b501-2218-4b20-9d6d-2ee738c75be0 · inbound

Security--Fidelity Tradeoffs: The Hidden Cost of Prompt Injection Defense cites this paper.

Security--Fidelity Tradeoffs: The Hidden Cost of Prompt Injection Defense Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 110

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T12:45:44.902281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T01:44:07.700127Z digest=sha256:a7e933588ae7258773d6f85d80494f1a09f4578e0e2ff66c710e52f09f91b8b7

Observation 81971ae4-9e3b-4842-8ff1-14cd403f27ce · inbound

Biased or Personalized? The Impact of Personal Information on AI-driven Development cites this paper.

Biased or Personalized? The Impact of Personal Information on AI-driven Development Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 81

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:56:10.740772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-09T09:46:37.332477Z digest=sha256:773a3d92a7bd5967d40bf9a63d5be32525863ff2958a5f10aa187f333ecb7b32