Pith. sign in

Paper Citation Record · LEDGER

Are aligned neural networks adversarially aligned?

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2306.15447.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.15447 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:35:50.164811Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

25
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 70b52c00-5d1e-4019-8c6b-dbede855f858 · inbound

Universal and Transferable Adversarial Attacks on Aligned Language Models cites this paper.

Universal and Transferable Adversarial Attacks on Aligned Language Models Are aligned neural networks adversarially aligned?

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-24T07:44:08.583092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-24T07:42:09.112946Z digest=sha256:75d16d710246148695a030f5fcfa75c71fb04774be88caec9b7aa89cf4acdb58

Observation 85a81cbd-593d-4fd0-b8d4-ab1c37255d06 · inbound

Baseline Defenses for Adversarial Attacks Against Aligned Language Models cites this paper.

Baseline Defenses for Adversarial Attacks Against Aligned Language Models Are aligned neural networks adversarially aligned?

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:24:40.058993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-13T23:24:39.835347Z digest=sha256:668c98e0fea511998a1033b3369256d4c36b31f258e1192eed8efdfddc878c52

Observation aaa59a1f-99e0-40dd-96fc-eb6c8f354cf4 · inbound

Low-Resource Languages Jailbreak GPT-4 cites this paper.

Low-Resource Languages Jailbreak GPT-4 Are aligned neural networks adversarially aligned?

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T09:24:14.078179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:cc222311053fc16ec40ba6246508bf526b304c5e976f9227d63906097c05663f

Observation aac2d768-ee6f-446e-9d45-521ced7a6b5e · inbound

SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks cites this paper.

SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks Are aligned neural networks adversarially aligned?

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:11:00.979943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T17:11:00.639293Z digest=sha256:1601466aa842d7594cc1ee003c505777e3c0dd95f2df807359e5acd56fab2a11

Observation 3403e9d2-551a-403b-8b32-e16d0ad468b5 · inbound

Improved Baselines with Visual Instruction Tuning cites this paper.

Improved Baselines with Visual Instruction Tuning Are aligned neural networks adversarially aligned?

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T19:11:34.013332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T19:11:33.783746Z digest=sha256:35d0f9b7460d1d49453e93345874f5bf60ac1bbbc27b1f0d734fca014b05e3f7

Observation 98085e8b-d96b-4ffd-a0be-7156b031366a · inbound

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation cites this paper.

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation Are aligned neural networks adversarially aligned?

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T22:00:51.514305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T22:00:51.487120Z digest=sha256:bf7d1fbd0f7fc12ea4693103f10f95f0ca4f368b532a072c66873047191f2e09

Observation 8ddb4cf7-b71d-4e9a-bb31-fda093a272d0 · inbound

Scalable Extraction of Training Data from (Production) Language Models cites this paper.

Scalable Extraction of Training Data from (Production) Language Models Are aligned neural networks adversarially aligned?

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T19:00:46.800869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T19:00:46.708242Z digest=sha256:e5e0c47c31a8d1fdde8ba61d9568b90b8502455f01acff4e1bb8e7ae1af8a540

Observation df9b70f8-16a7-438d-89c7-9cbe07c6315d · inbound

Adversarial Hubness in Multi-Modal Retrieval cites this paper.

Adversarial Hubness in Multi-Modal Retrieval Are aligned neural networks adversarially aligned?

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T06:42:39.688561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T06:39:36.039613Z digest=sha256:b99281a57b950bbf9855eb4b6840bb30b7a95572818fc1ab91a6fad8237259d1

Observation 77e1130e-de80-4e79-896e-53283f08cf13 · inbound

Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment cites this paper.

Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment Are aligned neural networks adversarially aligned?

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:50.164811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:35:50.164811Z digest=sha256:14720d3c424cb521f6cd1dbf4f66d2929f621799c4a346992c91e9fe41b93b07

Observation 553ad8f8-e473-4063-af66-0fd1741979ea · inbound

Seeing the Threat: Vulnerabilities in Vision-Language Models to Adversarial Attack cites this paper.

Seeing the Threat: Vulnerabilities in Vision-Language Models to Adversarial Attack Are aligned neural networks adversarially aligned?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:00.372780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:00.372780Z digest=sha256:ad1d37fd65e0bed9c2f053c0241e089fd5a6c67a9b42e47d8fdb4d46a68be0e1

Observation f9454b0b-8c8d-44ad-898e-cd51dce5e8ea · inbound

Understanding the Supply Chain and Risks of Large Language Model Applications cites this paper.

Understanding the Supply Chain and Risks of Large Language Model Applications Are aligned neural networks adversarially aligned?

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:42.853648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:45:42.853648Z digest=sha256:a8e072a054bb4a6dae5cf879aa869bd6b1b9ea534d8e70f9340f747e1d249ba3

Observation 654a4353-6e73-416b-9b94-a353778e90be · inbound

Invitation Is All You Need! Promptware Attacks Against LLM-Powered Assistants in Production Are Practical and Dangerous cites this paper.

Invitation Is All You Need! Promptware Attacks Against LLM-Powered Assistants in Production Are Practical and Dangerous Are aligned neural networks adversarially aligned?

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T19:39:09.544533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:39:09.544533Z digest=sha256:9b1a74730ce6701dd1c1b1017d2827c3bfdfc105b16953df6b2277e1ada0ffdc

Observation 4ad71a2e-fd11-4cfe-b04f-9a463724be41 · inbound

AttackEval: A Systematic Empirical Study of Prompt Injection Attack Effectiveness Against Large Language Models cites this paper.

AttackEval: A Systematic Empirical Study of Prompt Injection Attack Effectiveness Against Large Language Models Are aligned neural networks adversarially aligned?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-13T12:52:49.101462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:52:49.101462Z digest=sha256:eeb1f630923c84a80a82fdbf4d0026d6b5534f16e90575c88a451d56df3dd123

Observation 308e3d43-e2c8-4ec0-be1e-e0507397f344 · inbound

VisInject: Disruption != Injection -- A Dual-Dimension Evaluation of Universal Adversarial Attacks on Vision-Language Models cites this paper.

VisInject: Disruption != Injection -- A Dual-Dimension Evaluation of Universal Adversarial Attacks on Vision-Language Models Are aligned neural networks adversarially aligned?

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:56:08.052543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T14:24:48.999632Z digest=sha256:a5b52a0fd42570d9ba95657b374b333286fe4273866804971cd9e2551e30f374

Observation f2e3a91d-658d-45fc-b9f6-6ec9edc6e50c · inbound

Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours cites this paper.

Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours Are aligned neural networks adversarially aligned?

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:51:47.107368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T16:06:18.057868Z digest=sha256:3c5ebe061b0775c1b4f86a923df5d556fd76967712c1cb99d4716d8faf33f9c8

Observation 10a509cc-6079-4212-9bb1-f9b74c54e56a · inbound

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing cites this paper.

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing Are aligned neural networks adversarially aligned?

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:51:27.599788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T04:50:08.866969Z digest=sha256:034c542e7ac92920983a68b86da536e3c125a2f9c5e080aa4a0db009c1e12b62

Observation ee06d903-7470-4563-a7eb-d91c51946392 · inbound

Attention Hijacking: Response Manipulation Across Queries in Vision-Language Models cites this paper.

Attention Hijacking: Response Manipulation Across Queries in Vision-Language Models Are aligned neural networks adversarially aligned?

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:28:21.591740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T14:25:10.603780Z digest=sha256:0264beb6fd196c486c9da67b419973382221689b34a264da7e1eb27c4ba0bc17

Observation 532e644e-4b6d-4b0d-90e2-f46eb8df4150 · inbound

SCI-Defense: Defending Manipulation Attacks from Generative Engine Optimization cites this paper.

SCI-Defense: Defending Manipulation Attacks from Generative Engine Optimization Are aligned neural networks adversarially aligned?

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T07:44:42.727058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T07:43:00.837852Z digest=sha256:57c6905fb84016e1e37ac625bdf34752bfa36c9d6fba74337b5642345f7b7685

Observation b061b0a4-61be-4f8d-8f3a-99216c013c0c · inbound

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks cites this paper.

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks Are aligned neural networks adversarially aligned?

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:27:24.527682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T19:45:30.671490Z digest=sha256:9593d16156a0676dc3fe99a7b8ad6a5c2d105af73ae9bb657d31201f2ace5ddc