Pith. sign in

Paper Citation Record · LEDGER

Are aligned neural networks adversarially aligned?

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2306.15447.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.15447 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T18:05:52.227708Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

25
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 70b52c00-5d1e-4019-8c6b-dbede855f858 · inbound

Universal and Transferable Adversarial Attacks on Aligned Language Models cites this paper.

Universal and Transferable Adversarial Attacks on Aligned Language Models Are aligned neural networks adversarially aligned?

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-24T07:44:08.583092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-24T07:42:09.112946Z digest=sha256:344e5e158e7e020c2ad49897201802a57e0459bbf92ee47580762d5533ae7b71

Observation 85a81cbd-593d-4fd0-b8d4-ab1c37255d06 · inbound

Baseline Defenses for Adversarial Attacks Against Aligned Language Models cites this paper.

Baseline Defenses for Adversarial Attacks Against Aligned Language Models Are aligned neural networks adversarially aligned?

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:24:40.058993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T23:24:39.835347Z digest=sha256:7b788cfcc23ad60e0dc452b2d3f84bb81de9104fb2f486f31d88f67e430b06ea

Observation aaa59a1f-99e0-40dd-96fc-eb6c8f354cf4 · inbound

Low-Resource Languages Jailbreak GPT-4 cites this paper.

Low-Resource Languages Jailbreak GPT-4 Are aligned neural networks adversarially aligned?

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T09:24:14.078179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:866cbe570f55b47b0e6b3d0eb7281dba08eab1d27f7440afdddec46be27a7804

Observation aac2d768-ee6f-446e-9d45-521ced7a6b5e · inbound

SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks cites this paper.

SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks Are aligned neural networks adversarially aligned?

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:11:00.979943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-14T17:11:00.639293Z digest=sha256:5d8c8cdce01cc1d7f39c6eb2884cd963bc4085c0de884677a182e9e867397ff6

Observation 3403e9d2-551a-403b-8b32-e16d0ad468b5 · inbound

Improved Baselines with Visual Instruction Tuning cites this paper.

Improved Baselines with Visual Instruction Tuning Are aligned neural networks adversarially aligned?

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T19:11:34.013332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T19:11:33.783746Z digest=sha256:8247cd35d53b26d837c543ba7b8fc6a731865b9d92ee310964e0effc3c32ee28

Observation 98085e8b-d96b-4ffd-a0be-7156b031366a · inbound

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation cites this paper.

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation Are aligned neural networks adversarially aligned?

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T22:00:51.514305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T22:00:51.487120Z digest=sha256:97351a6490dc393d6ebc4ded59831925f0d6ca08d1de0b190b36a4d411d151a4

Observation 8ddb4cf7-b71d-4e9a-bb31-fda093a272d0 · inbound

Scalable Extraction of Training Data from (Production) Language Models cites this paper.

Scalable Extraction of Training Data from (Production) Language Models Are aligned neural networks adversarially aligned?

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T19:00:46.800869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T19:00:46.708242Z digest=sha256:dc273068b721b074e85054ac8eea7fef18d64e4b95ba21e7e044e3219be5bb31

Observation df9b70f8-16a7-438d-89c7-9cbe07c6315d · inbound

Adversarial Hubness in Multi-Modal Retrieval cites this paper.

Adversarial Hubness in Multi-Modal Retrieval Are aligned neural networks adversarially aligned?

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T06:42:39.688561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T06:39:36.039613Z digest=sha256:d75fb6c9b0551d045bd8d57bed6d77a7dc09a3ed88082f74f03a82b3a5ee394d

Observation 910dba40-53d9-40fc-8903-a2de0d9c95e2 · inbound

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models cites this paper.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Are aligned neural networks adversarially aligned?

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.227708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.227708Z digest=sha256:62a8b1b725581324081ef47e3d5487bc18479e8c5e8369a21f38f9ad0f175588

Observation a15d7181-6876-4ec7-8da1-8a945c4b0596 · inbound

A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations cites this paper.

A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations Are aligned neural networks adversarially aligned?

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-09T00:50:00.386506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:50:00.386506Z digest=sha256:97719e068015d5be906ee9d09debab7a4cea5305ba15b1c7a581a4ebcd6541d8

Observation 77e1130e-de80-4e79-896e-53283f08cf13 · inbound

Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment cites this paper.

Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment Are aligned neural networks adversarially aligned?

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:50.164811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:35:50.164811Z digest=sha256:14720d3c424cb521f6cd1dbf4f66d2929f621799c4a346992c91e9fe41b93b07

Observation 553ad8f8-e473-4063-af66-0fd1741979ea · inbound

Seeing the Threat: Vulnerabilities in Vision-Language Models to Adversarial Attack cites this paper.

Seeing the Threat: Vulnerabilities in Vision-Language Models to Adversarial Attack Are aligned neural networks adversarially aligned?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:00.372780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:00.372780Z digest=sha256:ad1d37fd65e0bed9c2f053c0241e089fd5a6c67a9b42e47d8fdb4d46a68be0e1

Observation f9454b0b-8c8d-44ad-898e-cd51dce5e8ea · inbound

Understanding the Supply Chain and Risks of Large Language Model Applications cites this paper.

Understanding the Supply Chain and Risks of Large Language Model Applications Are aligned neural networks adversarially aligned?

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:42.853648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:45:42.853648Z digest=sha256:a8e072a054bb4a6dae5cf879aa869bd6b1b9ea534d8e70f9340f747e1d249ba3

Observation 654a4353-6e73-416b-9b94-a353778e90be · inbound

Invitation Is All You Need! Promptware Attacks Against LLM-Powered Assistants in Production Are Practical and Dangerous cites this paper.

Invitation Is All You Need! Promptware Attacks Against LLM-Powered Assistants in Production Are Practical and Dangerous Are aligned neural networks adversarially aligned?

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T19:39:09.544533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:39:09.544533Z digest=sha256:39af328f3d843a5984928a80f637a6efafd2ce42dc1560e1bc5579ee603134c3

Observation 4ad71a2e-fd11-4cfe-b04f-9a463724be41 · inbound

AttackEval: A Systematic Empirical Study of Prompt Injection Attack Effectiveness Against Large Language Models cites this paper.

AttackEval: A Systematic Empirical Study of Prompt Injection Attack Effectiveness Against Large Language Models Are aligned neural networks adversarially aligned?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-13T12:52:49.101462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:52:49.101462Z digest=sha256:eeb1f630923c84a80a82fdbf4d0026d6b5534f16e90575c88a451d56df3dd123

Observation 308e3d43-e2c8-4ec0-be1e-e0507397f344 · inbound

VisInject: Disruption != Injection -- A Dual-Dimension Evaluation of Universal Adversarial Attacks on Vision-Language Models cites this paper.

VisInject: Disruption != Injection -- A Dual-Dimension Evaluation of Universal Adversarial Attacks on Vision-Language Models Are aligned neural networks adversarially aligned?

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:56:08.052543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T14:24:48.999632Z digest=sha256:89dc0180a86ff3ed55a9c0c04820bb15a1aeb9561ca7e8d0b82cd03037e0b076

Observation f2e3a91d-658d-45fc-b9f6-6ec9edc6e50c · inbound

Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours cites this paper.

Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours Are aligned neural networks adversarially aligned?

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:51:47.107368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T16:06:18.057868Z digest=sha256:eaf9ff54bf32c4402e4b077e972238b5508a3d27611d34fef0c0f15851cc0b43

Observation 10a509cc-6079-4212-9bb1-f9b74c54e56a · inbound

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing cites this paper.

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing Are aligned neural networks adversarially aligned?

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:51:27.599788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T04:50:08.866969Z digest=sha256:24f4c769691a397d08b8a86d36027a0c30551594899ff8a2912a408a2875e094

Observation ee06d903-7470-4563-a7eb-d91c51946392 · inbound

Attention Hijacking: Response Manipulation Across Queries in Vision-Language Models cites this paper.

Attention Hijacking: Response Manipulation Across Queries in Vision-Language Models Are aligned neural networks adversarially aligned?

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:28:21.591740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T14:25:10.603780Z digest=sha256:f610e69327720fab5e58c5b04dfd08e735810d313c956271c97e07a04913ac44

Observation 532e644e-4b6d-4b0d-90e2-f46eb8df4150 · inbound

SCI-Defense: Defending Manipulation Attacks from Generative Engine Optimization cites this paper.

SCI-Defense: Defending Manipulation Attacks from Generative Engine Optimization Are aligned neural networks adversarially aligned?

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T07:44:42.727058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T07:43:00.837852Z digest=sha256:754683c532972700a664f5c6f7ced06a2688bbf507ae3acfccf998955e625179

Observation b061b0a4-61be-4f8d-8f3a-99216c013c0c · inbound

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks cites this paper.

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks Are aligned neural networks adversarially aligned?

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:27:24.527682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T19:45:30.671490Z digest=sha256:3f8859d9d68ddcf827d1dfff602395ddd5521289e2efc0edd6c5b073ab29909c