Pith. sign in

Paper Citation Record · LEDGER

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content

As of 9 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2509.12672.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.12672 v2

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T16:40:31.250299Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 618cc52e-9b6b-4065-b6ed-9e5c3d55ef56 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.116422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.116422Z digest=sha256:fd56a00bacbd2fb7f4f5f95c58d6f8f718eeed3924587fea8a43244be5059861

Observation 46cdf308-c29a-430e-aa6a-821fff51889b · outbound

This paper cites write newline.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.120087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.120087Z digest=sha256:9632300ea30a85791d7948fe3da3a39c0283bb62930ad77cecae37e03c43d8b9

Observation d3614bcb-5916-431c-86c4-dcebb732c7e2 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.123713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.123713Z digest=sha256:2b59136b053ad81decc9a4e8a59923d157f2261395a1364d0a2cd9d1751ab0ec

Observation 6bc19f82-69af-4133-aee9-b35823587b2e · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.126880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.126880Z digest=sha256:98d3e508fcbfe4d720ce53a11a84a6558c01e116535e010c2becd9200ab8ca82

Observation b74c5b7a-dd7d-4c18-81a7-a2f8c38e83cc · outbound

This paper cites Mechanistic Interpretability for AI Safety -- A Review.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Mechanistic Interpretability for AI Safety -- A Review

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.129847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.129847Z digest=sha256:42352fd345129137baf4d831cdf4ea33fe5b9c75a6d20720b7a7444a2cf63f4a

Observation 072c8864-8fc0-4f1c-94cf-4981912c04ea · outbound

This paper cites Towards Building a Robust Toxicity Predictor.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Towards Building a Robust Toxicity Predictor

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.133163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.133163Z digest=sha256:46e46bf169b6d8132d313ec4e359578ee3f1552925e71cdc5e1623ea2868c377

Observation 247f0cce-2568-4a35-a002-e7f5d21532c6 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.136971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.136971Z digest=sha256:34e12b008cad65d185523b54dc85a8d27ba3338b302d2a893ce8b107eb171571

Observation 38b51520-2d7a-4a6f-9c57-3c6830720a52 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.140263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.140263Z digest=sha256:c1f2c0c874c84155ad24103b2f892b39a08db4b3470c1495244e8ccf530e73f2

Observation 1390aade-ade4-4013-b84e-7deb1b97cd2f · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.143456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.143456Z digest=sha256:6ce0ec5a6db217c105ad23744db7e953a25e83f60e5dc1734d2f4e8ca7c76b4e

Observation a1e91bfd-b394-45b8-af01-0d6b5052dbca · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.146244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.146244Z digest=sha256:c884358d96aa0e60ab3ac1291d99922679ad3ed7ec58c962fc3d81dbe31ddc1b

Observation 7e2e582d-84cf-4b6f-af44-3536a0ba7fa2 · outbound

This paper cites Towards Automated Circuit Discovery for Mechanistic Interpretability.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Towards Automated Circuit Discovery for Mechanistic Interpretability

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.149122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.149122Z digest=sha256:aca4145226336cfe8528894bbca8e38a269787ad94e19ed2c8dffd99f0fabbfe

Observation 1133c7a5-f396-47a1-9fb7-8b5b823790ba · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.153001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.153001Z digest=sha256:34091970f2699103ea49b8e6ab0cc361718e446c4a9955f969fad7a9e95dd923

Observation 8163cfac-a653-45b1-8e2d-aaa7b94838d6 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.156541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.156541Z digest=sha256:712c84586c2c2f810b8d9b8f5aaa173e19a5fa3c55a5b4756ddc86c7ea6a0d21

Observation a59922d2-4c9d-4813-8857-b2829dc5399b · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.159467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.159467Z digest=sha256:afdd7f61cb8d9a05c1c4ba56197b7c68f326649f13bc36907ac7a19701aa7f1c

Observation 5aaad299-99cd-4808-9824-4eb97b536c99 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.162193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.162193Z digest=sha256:5290457c608995b3d6dee2c478e548c4f736cfa40b4c0e0816646e0bf74d6372

Observation 140d4b5f-71ba-4710-b87a-debafd5c72a4 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.165692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.165692Z digest=sha256:a221d6b82ab47f89dfbbed8df43db789d6fb373ae603e80f43421776b43afe7e

Observation af396301-d35d-4ef6-a04d-098c5fb597c9 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.168451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.168451Z digest=sha256:71ea022509db60f1be250cb60c7fb2ab6b27b3434cd16ad2c9defd608ca278bc

Observation 94337268-ddd3-4ed5-bd4a-6ba7e201cbd3 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.171648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.171648Z digest=sha256:5839add4f678fae8763012edb23172fde1c6aa3b037c2b56905ab349ba0eb5f7

Observation 26eea721-1a69-4a69-a7b5-ac85284641c2 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.174811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.174811Z digest=sha256:0f580c1df0556e9c03bd3f2d6b69e898a8e086d48cbc70f97d6f82f14ee0f909

Observation 50d3bb49-f4db-49aa-92d6-7022dbbacd81 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.178582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.178582Z digest=sha256:25d08a46a385ae3201bb3f4c3a27de0ce61add7b661e17bb535bbf6a8da34aeb

Observation cc9cd95e-1d06-44e1-9fe2-c423bae08b88 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.181797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.181797Z digest=sha256:91f49cc1f100396a373350a8ad556d53032953e5284442cc1a30bbbc983cef70

Observation cf8c2438-d63d-4082-803e-d1245e896c3f · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.184837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.184837Z digest=sha256:bece2ea7ecf7d67b084ed439917faadd522e9ebefcc20acfedbf886116737b3a

Observation a04ecd64-ce82-4bcb-b222-78843b17a1a3 · outbound

This paper cites A New Generation of Perspective API: Efficient Multilingual Character-level Transformers.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content A New Generation of Perspective API: Efficient Multilingual Character-level Transformers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.187575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.187575Z digest=sha256:81ba683eae73e84d7f81e97dab8f3afca606bb2198a2806d405ca1a1f2c09373

Observation 63aee8c0-e72c-4077-a558-d254dc22d77c · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.190708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.190708Z digest=sha256:10bd4e3f5ba9dc1caf9edb1f6b834220018a01bdd54026f9d9f69fa8a3c7f77b

Observation 1a99733c-38c0-447e-aaf8-0988630fe711 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.193428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.193428Z digest=sha256:0aedf84dfd244ba8985a55aa3c834eeb012fd7f459670e10e751835d98c2b531

Observation c8773139-94a8-4f6e-a960-35935bf51f31 · outbound

This paper cites D.; and Finn, C.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content D.; and Finn, C

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.196217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.196217Z digest=sha256:c57ce60e3415c8e2b66e1718c398b51d39a4ff4999beabe1c4ac9e876c0f8d3b

Observation b20839b2-68a0-4c99-9fbe-5a70fe119bb5 · outbound

This paper cites R.; Li, G.; and Crespi, N.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content R.; Li, G.; and Crespi, N

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.198929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.198929Z digest=sha256:6cfdd41ccafdef755084e7dfe64bdf43e81b59237472ae4a894c510d653a3057

Observation 1ff28abf-a7b1-45b6-8beb-025ff87e9613 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.202127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.202127Z digest=sha256:69aca2db39244ca4c0d888c669d0304ad0af12f38e45f36e8da7d5482ef8900e

Observation 8b40be0b-e29d-4760-b373-2f63355c226f · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.204983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.204983Z digest=sha256:5c30deed541b35fc503d86534ab317f66ace2c70047fe61261a04f72603e0987

Observation ad8a4e2f-5d17-4a7f-9ef6-42f3e805f4a1 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.207748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.207748Z digest=sha256:77e9763159f7b2929ba129375186ae28c5ef95773a9fc39f73e95b21f62f7bfe

Observation e1b7ddb6-f539-44e8-a6dd-12d2629bb605 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.210447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.210447Z digest=sha256:75a8677ebb4644004268bf33a2412238a7ccbe2c7a43806b64be4719cc08ca46

Observation 293bf6f3-d479-4531-a742-f2abe67f3e14 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.213577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.213577Z digest=sha256:b6ffb6b36a3b35f4c81ba72f94733e1ee16d05fbbc4927a499e5eb5f3900118e

Observation f0ec7491-8aa0-4f1d-8f17-166cf60effcd · outbound

This paper cites Token-Modification Adversarial Attacks for Natural Language Processing: A Survey.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Token-Modification Adversarial Attacks for Natural Language Processing: A Survey

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.216119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.216119Z digest=sha256:8913ebb0e611d1f062c27c737b8a2845be06843d391dbd3e8afbe6381817fd88

Observation cce6712f-334d-4d41-9a63-1f8030cf7501 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.219115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.219115Z digest=sha256:6d650ee9a26f35b0798d14a847013d333b019b34c853d444fd67c37b856a1cd9

Observation 8aaaba9b-adf5-48c8-a93f-43ac7d596243 · outbound

This paper cites o ck, F.; and Wagner, C. 2021. “Call me sexist, but.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content o ck, F.; and Wagner, C. 2021. “Call me sexist, but

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.222044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.222044Z digest=sha256:224edfc719ae0385d89ea105dfe1f057cf312cd11f45454803176c9d9e21d8f6

Observation a2ae48b7-6b5d-4803-9419-1809e0dcd2f0 · outbound

This paper cites HowkGPT: Investigating the Detection of ChatGPT-generated University Student Homework through Context-Aware Perplexity Analysis.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content HowkGPT: Investigating the Detection of ChatGPT-generated University Student Homework through Context-Aware Perplexity Analysis

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.225010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.225010Z digest=sha256:23e7b9ca5fa5ddad96d26000f536e1cb26c0f5ad6d64f6231d301f49b3c1b9f0

Observation 86ca1c4e-0812-4f9a-83dd-4a60ca6afa21 · outbound

This paper cites Enhancing Adversarial Text Attacks on BERT Models with Projected Gradient Descent.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Enhancing Adversarial Text Attacks on BERT Models with Projected Gradient Descent

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.227672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.227672Z digest=sha256:e44f187b5276486fd066d6dd3a1532dbbaa9b8846b1ecff3f4fb8c8da3530a24

Observation 4b27728d-2133-4603-9263-342399895f24 · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.230569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.230569Z digest=sha256:f09a3d2d219039bacb6d6163abaf212f236ff96a6562bc67acfcc96506a59da6

Observation e72fdbae-b950-4752-92da-431093f8e751 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.233207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.233207Z digest=sha256:a8fd0aaf2ca7ac92bc33211f698ebf00d710010eb53ff830c404dde07020135d

Observation 7b88280b-d6f3-4674-a4ee-071c97d11a47 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.235984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.235984Z digest=sha256:1a43e2f66273ccd8f56f647ad21713770da2eb62bc9214fad0468608a880bd1c

Observation 25799df4-b1cf-4a0a-8220-d889315bb4c5 · outbound

This paper cites S.; and Wong, D.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content S.; and Wong, D

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.238745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.238745Z digest=sha256:f45858949b8e52fb0f08542adef32170f89c66ee9a796280b738d0d647c10264

Observation d9849187-ef1a-42d2-ac4e-7c51ab90eeae · outbound

This paper cites S.; and Wong, D.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content S.; and Wong, D

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.241698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.241698Z digest=sha256:f2a6200196682acc32d1f90774fcf069f75b61d4f7b0e508e91fc916e2ab0fbe

Observation 1fd0ebd5-059c-4277-aa79-70989e890b6a · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.244420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.244420Z digest=sha256:18af33fb90873970bfcd121d604ee098bc18bb60d8fd3d45c2c0ee6099862e24

Observation f925dcdf-a08b-4a7c-8316-fda551ff1c15 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.247794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.247794Z digest=sha256:c8e2b35b171c620cb186ad25be30707a6b59a945d7926aeb2209c110a534f0ce

Observation cc5dfd60-e7b4-4b2a-9700-f45348ff631c · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.250299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.250299Z digest=sha256:5564288314c98a6428d2320a1a43401d8795f30afe7547950e48adf7a565344d

Pith citing papers

No inbound Pith citation observations are available.