Pith. sign in

Paper Citation Record · LEDGER

Gradient-based Adversarial Attacks against Text Transformers

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2104.13733.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.13733 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T23:40:38.967797Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bad03ab0-6152-43de-8221-550d0843abaa · inbound

Scaling Laws for Reward Model Overoptimization cites this paper.

Scaling Laws for Reward Model Overoptimization Gradient-based Adversarial Attacks against Text Transformers

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T09:04:53.270517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T09:04:53.129737Z digest=sha256:d2abaf331549c70544077c02559e21dc8bee11a399a8fc1de6e3dff605c8d2d6

Observation d85c7d76-6cfc-4435-86f8-85b94843a510 · inbound

Universal and Transferable Adversarial Attacks on Aligned Language Models cites this paper.

Universal and Transferable Adversarial Attacks on Aligned Language Models Gradient-based Adversarial Attacks against Text Transformers

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-24T07:44:08.560184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-24T07:42:09.112946Z digest=sha256:9c675856387bd98aa263e69df69675d58bd24b0c7149e518afa7e9d3a34ccdf7

Observation 4ac26eb5-2910-475b-844e-733d075d7bef · inbound

Baseline Defenses for Adversarial Attacks Against Aligned Language Models cites this paper.

Baseline Defenses for Adversarial Attacks Against Aligned Language Models Gradient-based Adversarial Attacks against Text Transformers

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:24:39.992181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T23:24:39.835347Z digest=sha256:5bec360bf341d30a7314f147abfd61815d8896aa833d1ee5c65dbd964566fa93

Observation 2dbe7eae-7739-4e0e-9532-272258a68256 · inbound

On the Validity of Traditional Vulnerability Scoring Systems for Adversarial Attacks against LLMs cites this paper.

On the Validity of Traditional Vulnerability Scoring Systems for Adversarial Attacks against LLMs Gradient-based Adversarial Attacks against Text Transformers

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T23:40:18.827110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T23:40:18.827110Z digest=sha256:461cbfc237c2c77a11ba38fdbeb615503eb129a4b4d0094f2de819619c1ce294

Observation 789c3123-0d1f-4be9-a9a5-ee19da786ce2 · inbound

LLM-Virus: Evolutionary Jailbreak Attack on Large Language Models cites this paper.

LLM-Virus: Evolutionary Jailbreak Attack on Large Language Models Gradient-based Adversarial Attacks against Text Transformers

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T23:40:38.967797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:40:38.967797Z digest=sha256:b716d4a1f4073fd48f9d09168c47e46db00ab3cfad04614f042dab5293a85596

Observation a8d92230-73ab-4d7f-842e-27320e4e435e · inbound

Auto-RT: Automatic Jailbreak Strategy Exploration for Red-Teaming Large Language Models cites this paper.

Auto-RT: Automatic Jailbreak Strategy Exploration for Red-Teaming Large Language Models Gradient-based Adversarial Attacks against Text Transformers

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:54.594335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:25:54.594335Z digest=sha256:a68c7f36e4fa972e5ca3b67751118d167220d39eb0842f1ccc49529d1e9b7a90

Observation 5cbeee96-672b-4249-8329-db96609caa1a · inbound

Trojan Detection Through Pattern Recognition for Large Language Models cites this paper.

Trojan Detection Through Pattern Recognition for Large Language Models Gradient-based Adversarial Attacks against Text Transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T18:07:46.827206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:07:46.827206Z digest=sha256:810d33c2e48521f6a0c03ee261a57a0be8ff0d2f21ca9c6f4bb9385be30c7624

Observation f2c40a2f-82de-4342-ae64-a7ef4823dfda · inbound

Large Language Model Adversarial Landscape Through the Lens of Attack Objectives cites this paper.

Large Language Model Adversarial Landscape Through the Lens of Attack Objectives Gradient-based Adversarial Attacks against Text Transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T10:29:49.903229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:29:49.903229Z digest=sha256:0b843372e14cc2784c1e4101df52d8dafd9234bb6de98ebb39b256f18ff52a1d

Observation 50b85dd9-9475-4c0b-8448-3438ad2efc9a · inbound

A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations cites this paper.

A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations Gradient-based Adversarial Attacks against Text Transformers

Reference 158

Resolution
unresolved
no resolver link, observed 2026-08-09T00:50:00.603037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:50:00.603037Z digest=sha256:65532b7d0968110ec8c4f47c716923cc2a0a7906f9327a9e0fa32a436260cf3b

Observation 4e6df700-2e4b-420b-9ba3-1f0f723e8b69 · inbound

Universal Adversarial Attack on Aligned Multimodal LLMs cites this paper.

Universal Adversarial Attack on Aligned Multimodal LLMs Gradient-based Adversarial Attacks against Text Transformers

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T11:17:01.588766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T11:17:01.588766Z digest=sha256:31f239f74e5ed2ec7e0c8389e9620f12f91ffb72cea27eb11822edcab822cd1f

Observation fc9e9816-7330-44af-afd3-e6dc70848119 · inbound

Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations cites this paper.

Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations Gradient-based Adversarial Attacks against Text Transformers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:23:40.241768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:23:40.241768Z digest=sha256:5290a5828dbaf32ae54e08c9f2a7570e93d5ebfa36883db13f1cabeb3571a490

Observation d2b8bfd1-d10e-4bca-8a67-e090c66de9f7 · inbound

Winter Soldier: Backdooring Language Models at Pre-Training with Indirect Data Poisoning cites this paper.

Winter Soldier: Backdooring Language Models at Pre-Training with Indirect Data Poisoning Gradient-based Adversarial Attacks against Text Transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:16:18.660915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:16:18.660915Z digest=sha256:6a827bd41da26c79d58a209352cbe02bd69312d1c1cfd7c87f66f8ab64abb156

Observation 1d96cd72-a7d0-43b6-a1d7-c47a8e75fc9f · inbound

VERA: Variational Inference Framework for Jailbreaking Large Language Models cites this paper.

VERA: Variational Inference Framework for Jailbreaking Large Language Models Gradient-based Adversarial Attacks against Text Transformers

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:07:02.869101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:07:02.869101Z digest=sha256:923efa061d77b36d49ed54ecf258a08800112cba9c00b165f3a9abe319881504

Observation d23a7a2f-79cd-4aba-8f04-59a1578f4c25 · inbound

Influence-Guided Concolic Testing of Transformer Robustness cites this paper.

Influence-Guided Concolic Testing of Transformer Robustness Gradient-based Adversarial Attacks against Text Transformers

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:50.258022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:50.258022Z digest=sha256:d310373d81d862a106e6fad3b82019b2605bcb8fab79149e600291308b5e5f47

Observation aaf46548-b3da-44c2-8f82-73c3623764b2 · inbound

BarrierSteer: LLM Safety via Learning Barrier Steering cites this paper.

BarrierSteer: LLM Safety via Learning Barrier Steering Gradient-based Adversarial Attacks against Text Transformers

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:05:26.722368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-25T07:02:03.058731Z digest=sha256:3808804c7abc628d3ac157d7a4b9210243ac37b1b8894e5960c7ed580fc57402

Observation c5739eba-eca1-4994-a4c1-de74c09b536d · inbound

PIArena: A Platform for Prompt Injection Evaluation cites this paper.

PIArena: A Platform for Prompt Injection Evaluation Gradient-based Adversarial Attacks against Text Transformers

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:25:59.636183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:11:52.675226Z digest=sha256:2efd1fecb36f8bd824671f10f2b96c0aa05d8341af3c09c59a1410df5296dbd8

Observation 44d36458-51b2-4cc5-8314-7bf1b7530e88 · inbound

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing cites this paper.

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing Gradient-based Adversarial Attacks against Text Transformers

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:27.579695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T04:50:08.866969Z digest=sha256:fb14d99b6a9b4c9c4558e1d378a11ebcba636b50f4b5b5e758a56463eb34e9c4

Observation 88472547-8525-41b5-ad10-086ac0ab779f · inbound

AI Researchers Must Help Lead Arms Control to Mitigate Military AI Risks cites this paper.

AI Researchers Must Help Lead Arms Control to Mitigate Military AI Risks Gradient-based Adversarial Attacks against Text Transformers

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-07-03T13:08:08.789895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T08:26:57.379418Z digest=sha256:5c7c041328f2424c31d1b9da965b3fdd395fc2bfec9249d4eece5a63b4fdfd22

Observation 10cb1471-2f2c-4042-90df-d48c5dd63fa4 · inbound

Greedy Coordinate Diffusion: Effective and Semantically Coherent Adversarial Attacks via Diffusion Guidance cites this paper.

Greedy Coordinate Diffusion: Effective and Semantically Coherent Adversarial Attacks via Diffusion Guidance Gradient-based Adversarial Attacks against Text Transformers

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:08:43.462274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T04:35:35.594085Z digest=sha256:46f1bd776463f62b5f8d520f51a6b0b34714f77ec33241d56c650ab47ecc1136

Observation 0de3bdb9-e4a4-445e-8035-2237617a0b52 · inbound

Vulnerability of Natural Language Classifiers to Evolutionary Generated Adversarial Text cites this paper.

Vulnerability of Natural Language Classifiers to Evolutionary Generated Adversarial Text Gradient-based Adversarial Attacks against Text Transformers

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:59:52.547353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T04:41:23.586921Z digest=sha256:7c0d40b14e4c027721ba02ee908d64d7dc9d68de710a8f3df9d4fa3d2200c92f