Pith. sign in

Paper Citation Record · LEDGER

Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2304.05335.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.05335 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:16:32.749284Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T05:06:00.319233Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e1abe29f-9e10-4aed-bcf6-5bdea8dd3358 · inbound

A Survey on Large Language Model based Autonomous Agents cites this paper.

A Survey on Large Language Model based Autonomous Agents Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T04:03:00.886787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T04:03:00.340349Z digest=sha256:f764dde41f040935e03f877e8f9ffa62defaa4a42da6eaf365836f43e3766a0a

Observation f4f50e0e-26a8-4434-aecc-24eac081f1c3 · inbound

GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts cites this paper.

GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T06:25:21.143527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T06:25:20.966510Z digest=sha256:87361eba6f3270cf8f96af489cdf7b6e1922a410eda73e958bfb8747fc1bc254

Observation 9916e48f-2597-40c9-9c16-f73ad2fb5ea2 · inbound

SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks cites this paper.

SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:11:00.967657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-14T17:11:00.639293Z digest=sha256:ad98e030ef5f47df2a97a2d164e91733cf5f4adfe7bd09bd670ef64333157514

Observation 21670ddf-1bd6-40da-aa4e-82f485a2cc74 · inbound

Jailbreaking Black Box Large Language Models in Twenty Queries cites this paper.

Jailbreaking Black Box Large Language Models in Twenty Queries Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T09:48:33.209683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T09:48:31.721745Z digest=sha256:16551395b1f4736627026a8ab91d59675b12d69fa39528ce9ed7e8ee7c7863d7

Observation adeb3ebd-f9ce-4aa6-8893-a8ff65a1bb14 · inbound

Dr. Jekyll and Mr. Hyde: Two Faces of LLMs cites this paper.

Dr. Jekyll and Mr. Hyde: Two Faces of LLMs Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-24T05:06:00.323174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-24T05:05:35.726118Z digest=sha256:19fb3ab972ea06efbaf56374fea7ef6d447a84bc9acb0785f8956b4bd6c69a9b

Observation e4c2cbb7-da5c-4be6-89eb-d345c32b01f1 · inbound

TrustLLM: Trustworthiness in Large Language Models cites this paper.

TrustLLM: Trustworthiness in Large Language Models Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 247

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T11:17:08.703042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:7166e51a2fc5250d16621f27b2c9c81be675c7d2b35a14881daf4aebb04f2b0e

Observation e9ab008d-434c-4c36-979d-c911afe699ff · inbound

Toxic Subword Pruning for Dialogue Response Generation on Large Language Models cites this paper.

Toxic Subword Pruning for Dialogue Response Generation on Large Language Models Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:23:24.806346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-23T20:19:49.336725Z digest=sha256:b4a534f8a2c969358e6d86036aeeddcd8aedf1a73a92455969666c0e0669f09e

Observation 1ba81533-fe8b-4fc0-9a89-c3f1e18091eb · inbound

From Lived Experience to Insight: Unpacking the Psychological Risks of Using AI Conversational Agents cites this paper.

From Lived Experience to Insight: Unpacking the Psychological Risks of Using AI Conversational Agents Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T18:27:26.170473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:27:26.170473Z digest=sha256:675c88609589e30674f287e507f6700e96aab19173575e2e81fbb1caf4c66541

Observation c731d022-e0b5-4ab2-af41-1e260a7b6081 · inbound

Risk-Averse Finetuning of Large Language Models cites this paper.

Risk-Averse Finetuning of Large Language Models Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T20:54:05.066525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:54:05.066525Z digest=sha256:66b51b1357ac1291f1dad836a650e3f54abac54783aeca03da248f88810e3137

Observation ba5ab789-0452-4ebf-9415-c096ec31aed5 · inbound

Enhancing Patient-Centric Communication: Leveraging LLMs to Simulate Patient Perspectives cites this paper.

Enhancing Patient-Centric Communication: Leveraging LLMs to Simulate Patient Perspectives Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T20:52:38.193214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:52:38.193214Z digest=sha256:ec2f72767dc6ba7a947d8151cb8c19dd4eb2b8744e06b355e835602c0710e4f4

Observation 2afd8d45-dd9e-4c55-bc1e-b3f9c6c1be48 · inbound

LLMs are Vulnerable to Malicious Prompts Disguised as Scientific Language cites this paper.

LLMs are Vulnerable to Malicious Prompts Disguised as Scientific Language Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T15:26:44.761325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:26:44.761325Z digest=sha256:6c552c88009a0a97800bf92a0873ae8b7445c4cdae6705ca3bd572416463bee7

Observation 6bc594ad-b278-462e-b48c-96687bb8aec7 · inbound

HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns cites this paper.

HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T11:03:28.304717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:03:28.304717Z digest=sha256:593769092bceb1ce8feebcb21f44f5d41a3fd7324a79d3134cda5417fc686fb8

Observation 2025b726-2696-4717-be04-c3c9820a338c · inbound

On Almost Surely Safe Alignment of Large Language Models at Inference-Time cites this paper.

On Almost Surely Safe Alignment of Large Language Models at Inference-Time Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T16:18:40.502533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:18:40.502533Z digest=sha256:b5dd0b410033f0e7e4ebc77a74982ae9a715050197d9ff7e68eedc29b27cfa2b

Observation dd1dad92-56d5-4453-916b-6b9c423e243d · inbound

Benchmark on Peer Review Toxic Detection: A Challenging Task with a New Dataset cites this paper.

Benchmark on Peer Review Toxic Detection: A Challenging Task with a New Dataset Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T18:29:43.385459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:29:43.385459Z digest=sha256:2b24db0aa0f269e511d4d1086b524e69e6f0b0b6e3846a063409d36875c66d11

Observation 501d0307-42cd-4a71-8345-6560b3d86d13 · inbound

Revealing Weaknesses in Text Watermarking Through Self-Information Rewrite Attacks cites this paper.

Revealing Weaknesses in Text Watermarking Through Self-Information Rewrite Attacks Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T23:16:32.749284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:16:32.749284Z digest=sha256:2cf2b0bf520a1215f6a840f1c14cb605ced9db10f00ff0ba69cdc9ad0446c4fa

Observation 42064e71-fdde-4ab0-9cdc-443fcb4ca95f · inbound

What Really Matters in Many-Shot Attacks? An Empirical Study of Long-Context Vulnerabilities in LLMs cites this paper.

What Really Matters in Many-Shot Attacks? An Empirical Study of Long-Context Vulnerabilities in LLMs Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:26.717182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:26.717182Z digest=sha256:b871bea8ebff44bea1a0ca3686625f4850d3ac5cdfacc648f9f39c3941442ec4

Observation c90ee0b7-7452-408d-8e22-c3a20c34c6f7 · inbound

AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents cites this paper.

AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T10:54:57.939379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:54:57.939379Z digest=sha256:17004cd469303d6736fc1e3e876453d217bd543f89f59395fc5c5019bc7be9b4

Observation ad68433a-4a5b-4a7d-a04b-d8cb0eb51a37 · inbound

Data-Centric Safety and Ethical Measures for Data and AI Governance cites this paper.

Data-Centric Safety and Ethical Measures for Data and AI Governance Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:34:07.870925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:34:07.870925Z digest=sha256:2872c5b3c2cc7a48abad3bdb9802ddf6e878a1937951d1f88645ca8866eceb6a

Observation 07c3f700-075d-4942-ad45-f2ab4b8434ed · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 115

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:26.945695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:26.945695Z digest=sha256:27a8fabe329ba4c93901ace0a89b4d358980b05f5c344461ac34d1db3627ac6f

Observation d0a38afe-6aa1-4af0-86c3-74970dcba164 · inbound

Identifying Pre-training Data in LLMs: A Neuron Activation-Based Detection Framework cites this paper.

Identifying Pre-training Data in LLMs: A Neuron Activation-Based Detection Framework Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:15:17.984754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:15:17.984754Z digest=sha256:339aa350856d96e7d56b9f1f0c0a70d29b01e8600cc5a5500872aaa50f5e616c

Observation 7c0aa087-0da7-4037-87ca-fbe1f1a602f0 · inbound

Dutch CrowS-Pairs: Adapting a Challenge Dataset for Measuring Social Biases in Language Models for Dutch cites this paper.

Dutch CrowS-Pairs: Adapting a Challenge Dataset for Measuring Social Biases in Language Models for Dutch Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:13:15.547515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:13:15.547515Z digest=sha256:976e6ef77d6f29c61a7035dcb97a8d1287036d8c65286c520d72e11ec851d764

Observation f66dce05-8dae-467b-b1d9-8818193bc4dc · inbound

Towards Trustworthy AI: Characterizing User-Reported Risks across LLMs "In the Wild" cites this paper.

Towards Trustworthy AI: Characterizing User-Reported Risks across LLMs "In the Wild" Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T20:03:36.085653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:03:36.085653Z digest=sha256:a5eee35fb1907a4979cf69e5b8576efb73f5f76d87f995e1a2d9d6cd51477d7e

Observation a64c10cb-f32c-46fc-b2ef-1d39ba3bfc59 · inbound

Certifiable Safe RLHF: Semantic Grounding and Fixed Penalty Constraint Optimization for Safer LLM Alignment cites this paper.

Certifiable Safe RLHF: Semantic Grounding and Fixed Penalty Constraint Optimization for Safer LLM Alignment Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T12:27:28.661682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:27:28.661682Z digest=sha256:443565cd758025ac76e18ca06aaafc0ff35ce5944caabd7df3d503cc6ad86808

Observation 62f69773-5492-4d80-bf51-95910bcb15cc · inbound

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching cites this paper.

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 133

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:57:17.297802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-13T04:55:55.013900Z digest=sha256:fad52216f01d387f028083bd287c7190e6e0a25258f86978b52ce4783ae1ba95

Observation 053c9985-8aee-490d-ad19-f32d181ccbe0 · inbound

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching cites this paper.

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 133

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:45:06.563797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-15T05:41:10.714594Z digest=sha256:2550f2740558e1d128cdaf296cf3175c9fa23ce0c13d5bdb1624ee29ec751b32

Observation e2cf011e-77b4-4cc3-82da-3135301afa0e · inbound

Defenses at Odds: Measuring and Explaining Defense Conflicts in Large Language Models cites this paper.

Defenses at Odds: Measuring and Explaining Defense Conflicts in Large Language Models Toxicity in ChatGPT: Analyzing Persona-assigned Language Models

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T01:43:27.381147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T01:43:22.777232Z digest=sha256:9b617e2e10c58e84dad71b876fb54d57a23ddba2c1f75f59c7d2493545cb8edc