Pith. sign in

Paper Citation Record · LEDGER

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts

As of 8 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 6 inbound Pith citation observations for arXiv:2507.02990.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02990 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:05:21.460967Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T06:48:38.117769Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1a402ef3-322d-48cd-90b2-686900288e7b · outbound

This paper cites When llms meet cybersecurity: A systematic literature review.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts When llms meet cybersecurity: A systematic literature review

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:16.432188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:16.432188Z digest=sha256:b27529eb46f4f47570195fe8730d172fd4a69afff2ef1d5e6464c1215549de1a

Observation c6f1d6d2-1690-46d8-b1f8-1d091aad5827 · outbound

This paper cites Large language models in finance: A survey.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Large language models in finance: A survey

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:16.489860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:16.489860Z digest=sha256:f05bbe404ad47de0a8db9b0ac7a0bb1df9e98e878ce44b39fe1244ef1c55fc49

Observation be644471-7188-4409-ace1-138e1fba19ef · outbound

This paper cites Large language models in health care: Development, applications, and challenges.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Large language models in health care: Development, applications, and challenges

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.774818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:16.565523Z digest=sha256:69c8a054ff0f1be8a84fe06027c69c2f3fe470613e02996afb32e7a62082b4e7

Observation 45c6d5b7-52a2-4d71-9f74-c56e09daf5c7 · outbound

This paper cites A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.757089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:16.640450Z digest=sha256:8d520ae3b17630bf10833fa1f17c6d32e4b896671935f7271ee970aca6a8a0c7

Observation e410ea5f-f740-4b2d-972b-48c2a8f48dec · outbound

This paper cites Bias and fairness in large language models: A survey.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Bias and fairness in large language models: A survey

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.738587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:16.701083Z digest=sha256:5c58538085f813605f922f827611ec1e2c03458659049346b351d97d783b7297

Observation 4a6d9c51-9da3-468e-8ab4-4f557fe97a24 · outbound

This paper cites Safetybench: Evaluating the safety of large language models.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Safetybench: Evaluating the safety of large language models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.720605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:16.798283Z digest=sha256:342d9be71c74348f770cc882746eb4bf22eb56598679fa6bfccbcb83846d2ecd

Observation 28f5288e-143a-49a4-9bd9-ac3881cad586 · outbound

This paper cites do anything now.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts do anything now

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:16.884924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:16.884924Z digest=sha256:ce126e7028e613b8225911addf3b7237f47fd6a6850bdfa965d88ce1b255bc40

Observation 174d2080-f838-4300-b5f5-2beafe903f95 · outbound

This paper cites Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:16.959484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:16.959484Z digest=sha256:0b606ae791b8605a41bc74797a58a5588a6ef6c5430ab9a6e3819f89008d2026

Observation 1c8f77b9-b467-405b-a8f3-edf6aecc1fd8 · outbound

This paper cites Building Guardrails for Large Language Models.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Building Guardrails for Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:17.019897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:17.019897Z digest=sha256:46add9669461c8325835570df86bb4106e95b17fc21d53f8cd0c37f011280637

Observation 431c132d-06fd-49d5-b1cf-17fabae8dbff · outbound

This paper cites Guard: Role-playing to generate natural- language jailbreakings to test guideline adherence of large language models.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Guard: Role-playing to generate natural- language jailbreakings to test guideline adherence of large language models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:17.120436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:17.120436Z digest=sha256:1a17e2920ba39c1665f25776a5cf0111f5535e8b27a059671cd10016055643c6

Observation 08713358-008e-41e1-86a5-8822a45dc420 · outbound

This paper cites Training language models to follow instructions with human feedback.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Training language models to follow instructions with human feedback

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:17.197808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:17.197808Z digest=sha256:cd1db845b8cdc18d20dce8656aee1db97abdc593bcae7e1a9e8949520abaf802

Observation b77445c5-4ed5-4d35-bd82-24a55d14c43d · outbound

This paper cites Chain of hindsight aligns language models with feedback.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Chain of hindsight aligns language models with feedback

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.675364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:17.301394Z digest=sha256:1895371c2a21f959e03e1f9738a4fda49f5dda59f4b8364e7a81521c730275b6

Observation d995b637-ab42-459b-9d22-78a506f23924 · outbound

This paper cites Training language models with language feedback at scale.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Training language models with language feedback at scale

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.653629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:17.416173Z digest=sha256:016f48bd3ae499e891da1871aa007dce01963fe4ff3753c650879da09bbde03b

Observation 21175afd-3f69-46ae-8df2-a39f95eb083d · outbound

This paper cites Attack prompt generation for red teaming and defending large language models.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Attack prompt generation for red teaming and defending large language models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.632850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:17.521565Z digest=sha256:248ae420966cc774955d31ee7c84e242ee390e9c8f6f074895de141bf703f1b3

Observation fc79bc0b-8ba2-4372-af99-a6b86e259348 · outbound

This paper cites Harmbench: a standardized evaluation framework for automated red teaming and robust refusal.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Harmbench: a standardized evaluation framework for automated red teaming and robust refusal

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.615797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:17.614076Z digest=sha256:e2c65ff5c1bd403d68a78e7ac49dff528d3fbc398ca935bf1d527de7cbb4c056

Observation cc06eeba-c3a0-4604-b26b-f49ddb8fb0a8 · outbound

This paper cites Mart: Improving llm safety with multi-round automatic red-teaming.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Mart: Improving llm safety with multi-round automatic red-teaming

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.594966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:17.707807Z digest=sha256:37d0b79fbd3548baa51cf3dbef3da6c947fcbd100cf6037f1240b2db47712014

Observation 0a6f1769-fc9e-4b4f-bdef-7188083a1902 · outbound

This paper cites SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:17.761089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:17.761089Z digest=sha256:d88b444fabaa4055005a8f7b6d9477924a9f8369ce64bcdc78acf01aa2e6df12

Observation 4ea8569f-57a5-45f7-b940-345da1257ae9 · outbound

This paper cites Refusal in language models is mediated by a single direction.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Refusal in language models is mediated by a single direction

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.577394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:17.869695Z digest=sha256:675723d4be23a888fcf123dbca92f801d4a255f1075e33a8f97c35eee56aee4c

Observation 73c328d0-1e2d-462a-bd10-623a08e5900a · outbound

This paper cites The opportunities and risks of large language models in mental health.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts The opportunities and risks of large language models in mental health

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:17.975239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:17.975239Z digest=sha256:3beea2ef5a67fc261f00200e1b05fb90a77124db8078dd5392616dba449c346b

Observation 4c80e9b3-2479-4201-bb9c-fcd7ab5a7d0e · outbound

This paper cites Benefits and Harms of Large Language Models in Digital Mental Health.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Benefits and Harms of Large Language Models in Digital Mental Health

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:18.037577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:18.037577Z digest=sha256:1a622c11fdeb2276e230099d72d1ddb144cd4b8c8ad360666ebac6ac6387981a

Observation 6bf9d6fa-7597-486b-b4d6-efa0317fb6df · outbound

This paper cites Large Language Models in Mental Health Care: a Scoping Review.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Large Language Models in Mental Health Care: a Scoping Review

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:18.117935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:18.117935Z digest=sha256:43dff0e51c3cc9a0a666bac4599a2db2196cad56ae47bd5b1b72bee17148561f

Observation 54d82b29-b364-4806-b3db-1164bce1b177 · outbound

This paper cites To chat or bot to chat: Ethical issues with using chatbots in mental health.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts To chat or bot to chat: Ethical issues with using chatbots in mental health

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.548444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:18.216991Z digest=sha256:37fa8b8c14830b4a868a96477b28e41f2c77b0d5757bb388042d6ac637a75126

Observation 840f3e93-82c7-4be6-a765-e2142b816070 · outbound

This paper cites LLM-empowered Chatbots for Psychiatrist and Patient Simulation: Application and Evaluation.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts LLM-empowered Chatbots for Psychiatrist and Patient Simulation: Application and Evaluation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:18.316869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:18.316869Z digest=sha256:20838376eb4d6ce02f8079df331f2f2d36b2d45f6e28116f9d469050c778a694

Observation b2ed61d7-87e8-4f09-9a10-47afbe3a9b9e · outbound

This paper cites Chain of risks evaluation (core): A framework for safer large language models in public mental health.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Chain of risks evaluation (core): A framework for safer large language models in public mental health

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.533104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:18.389121Z digest=sha256:2d209decbe193b0eef7ee97c5f1fb6a772c8d100f5f6d7987d1766ca2d63ee34

Observation 4c037f2e-8076-4c04-a035-2a285a354c0b · outbound

This paper cites Adversarial attacks on large language models in medicine.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Adversarial attacks on large language models in medicine

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.517363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:18.480529Z digest=sha256:864fd2245342fd9f33decba37e97a3de1d33788ba046fda4bd6408d42098236c

Observation 2c0d0b61-3255-42b3-9d9f-51d3311e6128 · outbound

This paper cites Ensuring Safety and Trust: Analyzing the Risks of Large Language Models in Medicine.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Ensuring Safety and Trust: Analyzing the Risks of Large Language Models in Medicine

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:18.595255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:18.595255Z digest=sha256:fa45b21aa4de50569048ceb966f7b14b9ac750fd967017cd60174627ec25652d

Observation ba9507fe-a727-4cc8-bb5b-495c55ce771a · outbound

This paper cites MedSafetyBench: Evaluating and Improving the Medical Safety of Large Language Models.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts MedSafetyBench: Evaluating and Improving the Medical Safety of Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:18.696521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:18.696521Z digest=sha256:0c490198cf234833bbfb6413a89ea9a0898b80b041e00786030ba324d66b1292

Observation fbf8f2a0-87f1-4f7b-9659-cc7ac00e4469 · outbound

This paper cites Towards safe ai clinicians: A comprehensive study on large language model jailbreaking in healthcare, March 2025.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Towards safe ai clinicians: A comprehensive study on large language model jailbreaking in healthcare, March 2025

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.501553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:18.783457Z digest=sha256:f453e3872e09b0a72ce7969d21a12e8fc4b4d709c680363ee847c0618a1e3eb8

Observation 2c82d652-5ef2-485d-80ef-054c61e696a6 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:18.860880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:18.860880Z digest=sha256:55d1cd1cf47089839ed1561491356c6efc440e80c25a083299e41d6408c11bf0

Observation bb24187a-10c3-4ddf-a492-6e8696503ea8 · outbound

This paper cites Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:18.960898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:18.960898Z digest=sha256:8c0a798e5069707d4b0938406ce225ea2a38f68ad2f44b41688e83f44a227c21

Observation dd164da0-e7bd-4205-9b1e-9eec09b4d897 · outbound

This paper cites BadChain: Backdoor Chain-of-Thought Prompting for Large Language Models.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts BadChain: Backdoor Chain-of-Thought Prompting for Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:19.064128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:19.064128Z digest=sha256:7bcf7d57f89b21b04da16e3e89cd5287b7bb3b3fc299f850c74296c2d410ced7

Observation b3527ed0-15e6-4f3d-802b-ac57da5f1cef · outbound

This paper cites Low-Resource Languages Jailbreak GPT-4.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Low-Resource Languages Jailbreak GPT-4

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:19.158580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:19.158580Z digest=sha256:ade7f84e05111d1156a85eae16b8b9ffe192110573e55a5712c129e130ed220a

Observation 89cd22a9-2510-48fd-8580-eb567a1437b1 · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:19.247224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:19.247224Z digest=sha256:3c5d04042f78ab22f39dd7792b11bb79dc6829451f4e7404441003d995d74ef9

Observation 42865383-6229-458f-aec1-17dea6a5bcdf · outbound

This paper cites Not what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Not what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:19.312280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:19.312280Z digest=sha256:db97af989071dc27f690c555e37d9bfc275d14f81bbe679192409e94118c4994

Observation 3284897d-dcb6-4f19-94b1-f3889c235309 · outbound

This paper cites Multi-step Jailbreaking Privacy Attacks on ChatGPT.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Multi-step Jailbreaking Privacy Attacks on ChatGPT

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:19.409556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:19.409556Z digest=sha256:a4ee63abbb568ee5cdf4da86b238c5d0ee380b7b331f33de129aae0a2b07a98d

Observation 68e18392-3097-4b4f-8788-824bf3dbd397 · outbound

This paper cites Ignore Previous Prompt: Attack Techniques For Language Models.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Ignore Previous Prompt: Attack Techniques For Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:19.483730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:19.483730Z digest=sha256:85c79cc75c34435ec7de0a932045307cc3e484ad66877b787b6d6603cebfc302

Observation c0c45691-c685-4900-9502-e5fa552319c1 · outbound

This paper cites Red teaming language models with language models.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Red teaming language models with language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.474318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:19.574627Z digest=sha256:287da9d8604eb333b59eb826e4fc3c8483f488acb98e6004a3ffd204e59f56e0

Observation d09810c1-d4ee-4109-a8c5-44d27e3e21dd · outbound

This paper cites Jailbroken: How does llm safety training fail? Advances in Neural Information Processing Systems, 36:80079–80110, 2023.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Jailbroken: How does llm safety training fail? Advances in Neural Information Processing Systems, 36:80079–80110, 2023

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:19.641242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:19.641242Z digest=sha256:c33c37f652be9e56fcab20d2af906289d4a26e311bdaaf950dc3e820da7ed1df

Observation 755e8158-254d-43bf-bca4-43627fd3eab9 · outbound

This paper cites Fast adversarial attacks on language models in one gpu minute.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Fast adversarial attacks on language models in one gpu minute

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.440910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:19.757275Z digest=sha256:325b0195318dc5399320ae08bfad2a7ec1bb9a4e2f7c05f51d1fa30c6c798929

Observation 2830ec37-ab7f-4b59-9fc9-01d65ce47bde · outbound

This paper cites Rainbow teaming: Open-ended generation of diverse adversarial prompts.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Rainbow teaming: Open-ended generation of diverse adversarial prompts

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.420911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:19.767200Z digest=sha256:678e51e0d95b314f6ef78648bc2367d9096640d9e017e60c39ec9903a4b1788a

Observation 0c471f9b-d872-4037-8034-5e54a802f481 · outbound

This paper cites Query-based adversarial prompt generation.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Query-based adversarial prompt generation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.401922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:19.877530Z digest=sha256:4b511292c16fd19bc8fd70130cedbccae9a20bd8bc425c7125f0470a6c48a8cd

Observation 38c7990b-5572-40bb-bbe9-0bf72e03782f · outbound

This paper cites Jailbreak chat.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Jailbreak chat

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.383394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:20.036314Z digest=sha256:63b9ade0f6dfa89b63371b394d298bdac5b9ae76ed5af70f2fe3fc7c9afa6c9f

Observation 7a684a47-3608-401a-a936-b5c6a0b1034a · outbound

This paper cites DAN" (and other.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts DAN" (and other

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.364404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:20.162003Z digest=sha256:d15cadf61229aa4acecc1301bf26d56713d96aa36881cca20d65b08da5ab7985

Observation 84aa3a8a-bc97-491c-a3c3-b7dcc69c3ab4 · outbound

This paper cites Suicide, 2023.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Suicide, 2023

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.343377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:20.345089Z digest=sha256:33223e8ce8dad0b55c5a20c15feaa0e302078607f3d7981210cbd6b2f162b144

Observation 17c3dddd-7d9f-4769-b46f-95a4cde9f52a · outbound

This paper cites Adolescents’ use and perceived usefulness of generative ai for schoolwork: exploring their relationships with executive functioning and academic achievement.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Adolescents’ use and perceived usefulness of generative ai for schoolwork: exploring their relationships with executive functioning and academic achievement

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.262329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:20.506613Z digest=sha256:1ab882238f2952564234eb061eef618c99601aa8b5ccfdd66bc2249890a638dd

Observation 7b07b7bb-b41e-4d70-b81e-79d35a59f457 · outbound

This paper cites The truth about self-harm, 2024.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts The truth about self-harm, 2024

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:23.049788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:20.661003Z digest=sha256:ad53b7c0537d205a11a61578aadbe07167314e6f513fe5f095c27774b2681724

Observation b63b9f6d-7e34-474e-8db6-dd11cd008651 · outbound

This paper cites Chatbot encouraged teen’s suicide, lawsuit alleges, 2024.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Chatbot encouraged teen’s suicide, lawsuit alleges, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:22.891381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:20.893225Z digest=sha256:f7b3cd5b082a10790e9bc15d0754da18f233e13e2b345f70475910d4b0af77bd

Observation 52938eab-7390-441b-b298-21032e1f6fe7 · outbound

This paper cites Man ends his life after an ai chatbot ’encouraged’ him to sacrifice himself to stop climate change, 2023.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Man ends his life after an ai chatbot ’encouraged’ him to sacrifice himself to stop climate change, 2023

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:22.634158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:21.088627Z digest=sha256:b87fa4f7e681727844129c157c85a6b39b15a9d22e6605ed81c72363585badee

Observation 8898f39c-3b26-4749-96fd-4abe3ac0e3ab · outbound

This paper cites Ai chatbots pushed autistic teen to cut himself, lawsuit claims, 2024.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Ai chatbots pushed autistic teen to cut himself, lawsuit claims, 2024

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:22.408434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:21.291375Z digest=sha256:be619dec41f81b4e57603d33d99986e069dde54351d9b1026f76932eb89c2d14

Observation aab96ec5-9eb2-4e54-8501-7165cd8cb2d8 · outbound

This paper cites Suicide prevention by limiting access to methods: a review of theory and practice.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Suicide prevention by limiting access to methods: a review of theory and practice

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:05:22.165864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:05:21.460967Z digest=sha256:801f6cc867490f5a7ac4508326f4668d6b3ab62c76e5104231a0e382fbc98a85

Pith citing papers

Observation 7bccb174-2b23-4e6a-8d6f-90bc8d3681d3 · inbound

VERA-MH Concept Paper cites this paper.

VERA-MH Concept Paper `For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T07:01:00.743249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T07:00:15.411310Z digest=sha256:03f3fdb2c8e5d9ec8421881b2e507c8a6dcd1b46a7e8033ffab735066e162b1a

Observation 6b9c099e-52ba-41a4-a45c-33223954d90a · inbound

From Fact to Judgment: Investigating the Impact of Task Framing on LLM Conviction in Dialogue Systems cites this paper.

From Fact to Judgment: Investigating the Impact of Task Framing on LLM Conviction in Dialogue Systems `For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T06:48:38.117769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T06:48:38.117769Z digest=sha256:9f92fc47fc59da1b07026288703664804a68a203af83b5ad3accd8a8719a2ef4

Observation 984a06c4-8eb8-419b-abe0-7b3f38f250c1 · inbound

PARALLAX: Separating Genuine Hallucination Detection from Benchmark Construction Artifacts cites this paper.

PARALLAX: Separating Genuine Hallucination Detection from Benchmark Construction Artifacts `For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:12:44.917378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-19T20:09:47.750043Z digest=sha256:fc3393654dcaf205b977bc239c510fab361831e589bca9031ff865fd80b66283

Observation d4985a04-76e2-4f2e-a066-1e109d0dd063 · inbound

One Year Later...The Harms Persist, But So Do We! cites this paper.

One Year Later...The Harms Persist, But So Do We! `For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:35:29.500111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T06:33:58.246849Z digest=sha256:7650cfb0411444ae7c58c2bf1aa583fac8c46f058701845f3adedc604cc72f05

Observation 0d6326a2-9df7-4898-96b5-3025c37a060b · inbound

One Year Later...The Harms Persist, But So Do We! cites this paper.

One Year Later...The Harms Persist, But So Do We! `For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:47:27.592864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-02T21:38:29.468179Z digest=sha256:bf9d6b31e325a6b9d375344b6ca8b768967b110cf431fda0b380350263d97e63

Observation 8e5a83dc-a9c5-4caa-932c-08f75652f2d5 · inbound

Closing the AI Trust Gap: The Case for Independent Certification for Trustworthy AI cites this paper.

Closing the AI Trust Gap: The Case for Independent Certification for Trustworthy AI `For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T21:44:00.863851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:44:00.863851Z digest=sha256:cb739ddc2bca56b7bf953806c175c92925adc1ac298ddbf509232ea1faa76935