Pith. sign in

Paper Citation Record · LEDGER

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models

As of 9 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 7 inbound Pith citation observations for arXiv:2502.00718.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.00718 v2

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T18:05:52.411107Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:19.934200Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T05:53:09.562517Z

Reference resolution

59 of 59 outbound references displayed

  • verified exact5
  • verified fuzzy3
  • unresolved51
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7bf96aeb-feb5-4280-8404-bf68c1088b43 · outbound

This paper cites Flamingo: a Visual Language Model for Few-Shot Learning.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Flamingo: a Visual Language Model for Few-Shot Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.203599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.203599Z digest=sha256:6446faee38a4df50378706387148a58352b8ee1fc3a8f0b7f04145baddc57a7d

Observation 61bfcb78-ce78-4d9c-b207-5cc8efc7610a · outbound

This paper cites Multimodal machine learning: A survey and taxonomy.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Multimodal machine learning: A survey and taxonomy

Reference 2

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T18:05:53.756582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T18:05:52.208534Z digest=sha256:af6aefa761611c03d817ef7dc4725f6844f7f151a6b17b865a360f23d563d5f9

Observation dfce9281-ab22-4a5f-99a5-c5af9cb0dc71 · outbound

This paper cites Evasion Attacks against Machine Learning at Test Time, pp.\ 387–402.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Evasion Attacks against Machine Learning at Test Time, pp.\ 387–402

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.212255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.212255Z digest=sha256:d07bc8e16ec0b360f8d8a58fd09b4745ed5b8cb0ec773ae133a1078ac8435608

Observation 60f5044d-e729-40c0-8ba7-14f32dfddc2a · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models On the Opportunities and Risks of Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.215894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.215894Z digest=sha256:335b5931172f70cabaaa71c76733542edd28a6a77bbb76cab334641d81ba4cdf

Observation f1552174-b227-4c2e-801d-a2784d273df4 · outbound

This paper cites Language Models are Few-Shot Learners.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Language Models are Few-Shot Learners

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.219800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.219800Z digest=sha256:611fc98606d8934e158e0c217bc765500fede63f0f1067f983432e7bb4d285b9

Observation 5c48a6cc-0341-4ca9-922d-1702d9882feb · outbound

This paper cites Audio Adversarial Examples: Targeted Attacks on Speech-to-Text.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Audio Adversarial Examples: Targeted Attacks on Speech-to-Text

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.223609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.223609Z digest=sha256:20a551694a15d695a0ea8985dad2666e0ad30111873bc3082c54f614ddb79622

Observation 910dba40-53d9-40fc-8903-a2de0d9c95e2 · outbound

This paper cites Are aligned neural networks adversarially aligned?.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Are aligned neural networks adversarially aligned?

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.227708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.227708Z digest=sha256:9cfd52c34ade12adb0d936bd1a1db22f106c48b5fafaa8fdd142c50806f4e142

Observation 1d140a9d-87a8-4194-9922-066906d2cede · outbound

This paper cites JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.231429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.231429Z digest=sha256:ca4c7abef7383cc5309ebd743dced029cf479b5070c2e86ad3b970468f4250da

Observation 21115587-0540-47a4-9271-0a6430d890cd · outbound

This paper cites BEATs: Audio Pre-Training with Acoustic Tokenizers.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models BEATs: Audio Pre-Training with Acoustic Tokenizers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.235239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.235239Z digest=sha256:541d4214041d8737be23bf35c28e0ac507db655786477fff9a3fda30e8de8157

Observation a6ffa163-7034-41dc-b559-21894303ea0a · outbound

This paper cites E., Stoica, I., and Xing, E.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models E., Stoica, I., and Xing, E

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.240306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.240306Z digest=sha256:1f35d6a9e3ff236e7d546a516cad244174f846ac83811746b1db01ecf51aee5a

Observation 67f2d290-ed0e-4896-8685-b45fdb997948 · outbound

This paper cites Deep reinforcement learning from human preferences.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Deep reinforcement learning from human preferences

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.243569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.243569Z digest=sha256:8750d161644d2695b5180334a55168206afb343ad04579cde8dc6772e479e286

Observation f5448147-e2f0-4260-9b1d-48b69411d164 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.247296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.247296Z digest=sha256:aa84adb38dc8a92e9392cdfac127c99b78556c054de3f5c688813b03d0ca2528

Observation efa689ac-80a8-40e5-ae60-d1c2ecca8920 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.251049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.251049Z digest=sha256:67665515ce989c1165477fbcc38d01ea18491024f518ea8d32e82d0cb1e766ef

Observation 4dd26791-469c-4111-b9a2-572703678f65 · outbound

This paper cites Pengi: An Audio Language Model for Audio Tasks.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Pengi: An Audio Language Model for Audio Tasks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.254541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.254541Z digest=sha256:4e3ccf9585332e146241b252e94187f5c2ae70d1da33743155fdcf4ab09a4a58

Observation 72a3d137-feea-4b30-a407-aaad073858b8 · outbound

This paper cites HotFlip: White-Box Adversarial Examples for Text Classification.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models HotFlip: White-Box Adversarial Examples for Text Classification

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.258027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.258027Z digest=sha256:324a578cdd87c169270fc6dae796cf5819b325d3ad760cc46e5c7125b76eb4c0

Observation cd6077ff-6625-4d13-8c08-909271d57e7b · outbound

This paper cites Physical Adversarial Examples for Object Detectors.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Physical Adversarial Examples for Object Detectors

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.261748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.261748Z digest=sha256:e1f9868edf4f4cc03003f265d87a6ee11052f9a9cc9609f224d86ced99f899db

Observation 95a03691-838c-4c02-834c-86d4e303c06b · outbound

This paper cites JailbreakLens: Visual Analysis of Jailbreak Attacks Against Large Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models JailbreakLens: Visual Analysis of Jailbreak Attacks Against Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.265466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.265466Z digest=sha256:aa80da81cfc1d9693bc1137521c44fea47c59e4df4183a032920019e2badf3bb

Observation 4f394338-d719-4066-b742-428d8b994608 · outbound

This paper cites Advddos: Zero-query adversarial attacks against commercial speech recognition systems.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Advddos: Zero-query adversarial attacks against commercial speech recognition systems

Reference 18

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T18:05:53.465045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T18:05:52.269375Z digest=sha256:41502527b5f16739553c16c86e8ae0677771c456518ce0f639bd7b26b3dacf11

Observation 890b2520-8524-4e34-ba4d-e2bb57660447 · outbound

This paper cites RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.272593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.272593Z digest=sha256:7c0ca709fd91efc7633e636569e7508ed8e910dd3a477f805cc9270b45bcba5f

Observation ebf9c0f6-606f-4c82-9dd9-72356bae4ed4 · outbound

This paper cites and Unitary team.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models and Unitary team

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:05:53.807091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T18:05:52.276180Z digest=sha256:6c938196a09df0008eadac59a89ae379e9bc2d845457174c5eeb506ca9e8109c

Observation ba4f1ca4-fe9b-4438-9a3b-60cfda8044a0 · outbound

This paper cites Best-of-N Jailbreaking.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Best-of-N Jailbreaking

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.279543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.279543Z digest=sha256:e474b534724f5bdda3068262f7a14751d31d937f7326ffc661425b62c583b9c1

Observation 03886de2-ae14-407e-b3b6-41a4b5dedf0e · outbound

This paper cites and Liang, P.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models and Liang, P

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.283137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.283137Z digest=sha256:adc8dcb81242881c6e7ac039c4ed18796e53b5e327ff20919e35dcb2bccbdd3b

Observation 423c05fd-9c9d-41f1-b64f-be8c355461a8 · outbound

This paper cites M., Sun, J., and Chattopadhyay, S.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models M., Sun, J., and Chattopadhyay, S

Reference 23

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T18:05:53.239601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T18:05:52.286518Z digest=sha256:700f4c5bb04ab66e8d55762e56b558b6b3d2280147524d33c6bdd1252cdc7c18

Observation 041840ae-03d4-4b19-8736-402a9fcdfa6e · outbound

This paper cites Mixtral of Experts.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Mixtral of Experts

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.289872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.289872Z digest=sha256:4324e9241856a832948d9e4c47d7485d2ae39a6688166787bad0dc179c28c161

Observation 541756fa-360c-4b06-95e4-fd95f0abc85b · outbound

This paper cites AdvWave: Stealthy Adversarial Jailbreak Attack against Large Audio-Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models AdvWave: Stealthy Adversarial Jailbreak Attack against Large Audio-Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.293247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.293247Z digest=sha256:49a6c8179eeb0d16e72d3372e0ae7f3be98646e625ef324f8248ed190029848d

Observation 18360d4e-6246-40f1-919a-a41bef1a69d7 · outbound

This paper cites Towards Efficient Visual-Language Alignment of the Q-Former for Visual Reasoning Tasks.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Towards Efficient Visual-Language Alignment of the Q-Former for Visual Reasoning Tasks

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-09T18:05:53.055093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T18:05:52.297013Z digest=sha256:5d4886bc53d54d484f82b4a0903d6223456418a19b64da89b014a4bc919a0696

Observation 37755914-1b8c-4039-b0ec-023ebaf9f734 · outbound

This paper cites Voice biometrics fusion for enhanced security and speaker recognition: A comprehensive review.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Voice biometrics fusion for enhanced security and speaker recognition: A comprehensive review

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:05:53.795167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T18:05:52.300453Z digest=sha256:788c01fbc264e1980f54dd3fe7f6ea0a6cbce1bdf25cb656fc5f422b6f6c4394

Observation e400b6bb-d269-4a65-91c0-f431c6157c37 · outbound

This paper cites Fooling end-to-end speaker verification with adversarial examples.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Fooling end-to-end speaker verification with adversarial examples

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:05:53.783754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T18:05:52.303571Z digest=sha256:44b6c64206b5a12123d8e16b3db4d9adbd72e94fddb1e425ae395ec7fe550bb0

Observation 39bb094a-5346-4191-a8a1-d1488391ef6e · outbound

This paper cites AmpleGCG-Plus: A Strong Generative Model of Adversarial Suffixes to Jailbreak LLMs with Higher Success Rates in Fewer Attempts.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models AmpleGCG-Plus: A Strong Generative Model of Adversarial Suffixes to Jailbreak LLMs with Higher Success Rates in Fewer Attempts

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.306729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.306729Z digest=sha256:8ed9cffdea9f24d87a1d1ef44ca6dae843a9329643ba514b0ec204c0ff173e92

Observation 6bf3f076-3aa4-400d-a15d-5c9e7dafadb9 · outbound

This paper cites Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking Multimodal Large Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking Multimodal Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.310173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.310173Z digest=sha256:c0c07f207d4f112c470940921c6bd9f37b6af3fde14d33e319b2f6b3bd55e2b4

Observation 4c967946-2540-43cf-81a1-b095aab1dff6 · outbound

This paper cites AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.313581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.313581Z digest=sha256:91c43136de1e45dd30c13434ee83a5d4541f62de4b7716dce939bc8fb0de21c3

Observation 4b1029c8-0ea8-4c86-9e42-13b5172ea68a · outbound

This paper cites Jailbreaking Prompt Attack: A Controllable Adversarial Attack against Diffusion Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Jailbreaking Prompt Attack: A Controllable Adversarial Attack against Diffusion Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.317063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.317063Z digest=sha256:2814412232f71be13ca3f4ae86417aaf72158109c66fc4cb4410cd70bdea556f

Observation f1ae30b2-45e4-4644-bc47-cf7de33041a1 · outbound

This paper cites User interaction patterns and breakdowns in conversing with llm-powered voice assistants.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models User interaction patterns and breakdowns in conversing with llm-powered voice assistants

Reference 33

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T18:05:53.003628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T18:05:52.320546Z digest=sha256:8d9aaf98eb998ae8b5bd368f6973af7701714d869a8123811f77ec11b0887d90

Observation 5712b820-e111-4872-9133-134ee783c7dd · outbound

This paper cites Rule Based Rewards for Language Model Safety.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Rule Based Rewards for Language Model Safety

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.323811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.323811Z digest=sha256:d1d7ee7c8b93917503491b228a086469adc2745781971327f6116034589a9602

Observation c95fb8dc-f89a-4eae-96e7-5afcb3596340 · outbound

This paper cites GPT-4 Technical Report.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models GPT-4 Technical Report

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.327250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.327250Z digest=sha256:0e7a3af2a1b163f77f2fa4e0376d776494d17dd4999c055dc77ec11f8fcc5498

Observation b55288ce-0d88-46ee-a781-ce34e5ed02ee · outbound

This paper cites Training language models to follow instructions with human feedback.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Training language models to follow instructions with human feedback

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.330474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.330474Z digest=sha256:97a8b95e5f72d90180f25269f3daca6faec03a642155e48bc92b29dd8dd7fca1

Observation a053c25f-791b-4e62-bee6-b67f6605fa9b · outbound

This paper cites Red Teaming Language Models with Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Red Teaming Language Models with Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.334192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.334192Z digest=sha256:c8c02613f410938b0edb5f98e8d894635a4e3388683ef35268902e5de10e52a1

Observation 2dddea1c-6896-4130-9a1c-7cb95ca46e5e · outbound

This paper cites Visual Adversarial Examples Jailbreak Aligned Large Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.337707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.337707Z digest=sha256:f320553908f2aba1a7a1915781253b0ee5200c797b9f31adf1ac877fb08f751b

Observation 14eb5793-877f-4d05-99f9-9a7cadc400e0 · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Robust Speech Recognition via Large-Scale Weak Supervision

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.341242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.341242Z digest=sha256:17318d2d532345afbe919eb5dc485c65c37ccc7830597508db7bbfa482bd26af

Observation cbc995c3-5481-4a66-aca1-3eb0e873a171 · outbound

This paper cites Universal adversarial attacks on spoken language assessment systems.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Universal adversarial attacks on spoken language assessment systems

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.344668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.344668Z digest=sha256:79de2aaa421069d992011a0570e99da67cf450b048fb8b13db0e2c2ac1524bb6

Observation ed9b5ece-75b5-4893-9bec-1d2ee61463ed · outbound

This paper cites Muting whisper: A universal acoustic adversarial attack on speech foundation models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Muting whisper: A universal acoustic adversarial attack on speech foundation models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.348020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.348020Z digest=sha256:647625811a4dfa4262371be0f04743146b7bcb0fdff5bba413e30d86ed8d40c2

Observation 3d95282a-3d47-4311-a85a-6d319bc863a3 · outbound

This paper cites Artificial Intelligence and the Problem of Control, pp.\ 19--24.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Artificial Intelligence and the Problem of Control, pp.\ 19--24

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.351606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.351606Z digest=sha256:59b64bf5a31e643e672e0a5dfc995051f8f626fa9900feae6dcc7bad7e3dc6a0

Observation 36d4c59f-dd61-4521-8fdc-2b4ebb0ac5d1 · outbound

This paper cites Failures to Find Transferable Image Jailbreaks Between Vision-Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Failures to Find Transferable Image Jailbreaks Between Vision-Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.354966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.354966Z digest=sha256:eedf0473c264337a3cee7494d6bbdb94f12dff27ed5e59cb88d971cd30cf8a7f

Observation 1534340e-a0aa-4139-9ec8-3c72299b783b · outbound

This paper cites Adversarial Attacks Against Automatic Speech Recognition Systems via Psychoacoustic Hiding.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Adversarial Attacks Against Automatic Speech Recognition Systems via Psychoacoustic Hiding

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.358400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.358400Z digest=sha256:cd8894024e287559747cd30e6f20e7ab5095f022be6f6cddd6d5f6ce483d9a41

Observation 44c2689c-1bcf-45ec-b906-802d18747382 · outbound

This paper cites Jailbreak in pieces: Compositional Adversarial Attacks on Multi-Modal Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Jailbreak in pieces: Compositional Adversarial Attacks on Multi-Modal Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.362053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.362053Z digest=sha256:3e80cb1cb212ce46e39b3519479ae94fe15b50a4281a3d7db16ac5da80fea37c

Observation 536b9f55-816d-4bb6-9080-49c3f46acc1d · outbound

This paper cites "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.365731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.365731Z digest=sha256:377c1986841565ba74e2399a82f1b5ec693583d65669a55d9605d1760995519b

Observation 094fc1c1-70b9-4523-aad5-925d36fe9675 · outbound

This paper cites Voice Jailbreak Attacks Against GPT-4o.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Voice Jailbreak Attacks Against GPT-4o

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.369200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.369200Z digest=sha256:e984a68dd6b7b831edf7cb6c8ba773c72c785f4159eda6501ec2aee51966a023

Observation 264d152f-056b-4332-aa16-87d27ed8cc7b · outbound

This paper cites Intriguing properties of neural networks.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Intriguing properties of neural networks

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.372588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.372588Z digest=sha256:9fd0bdff8e21de36e541eba6ea9985c93f384414376482463b452ddfa4210129

Observation 89928ea5-dc64-4c6f-afd7-993b1a80a66a · outbound

This paper cites SALMONN: Towards Generic Hearing Abilities for Large Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models SALMONN: Towards Generic Hearing Abilities for Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.376150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.376150Z digest=sha256:be742fdadb15d20ae75a80a670f20ba15915d1caac4ef4102a549c2b5467e9f1

Observation 7e6d5872-af8b-4fee-801a-888e189b8cc1 · outbound

This paper cites Universal Adversarial Triggers for Attacking and Analyzing NLP.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Universal Adversarial Triggers for Attacking and Analyzing NLP

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.379750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.379750Z digest=sha256:66c9ac359a387f5433644cf6837f6587da5b39339363cb86b79ae5054f14c991

Observation 4e5b1b26-f85f-43b7-b49d-924f9d9b7389 · outbound

This paper cites Jailbroken: How Does LLM Safety Training Fail?.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Jailbroken: How Does LLM Safety Training Fail?

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.383285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.383285Z digest=sha256:56d1e8fb93677db9018e75becb34d79a125f65b367054de9a56626625d354f69

Observation 92e33961-2677-4e59-a605-8951d690c751 · outbound

This paper cites Finetuned Language Models Are Zero-Shot Learners.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Finetuned Language Models Are Zero-Shot Learners

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.386625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.386625Z digest=sha256:6ab4be819b6d61674ac1a86e47634c83df71292d6123a1d180c18447770748fc

Observation ec77abd3-e734-45ab-ac5a-7233f3539636 · outbound

This paper cites Ethical and social risks of harm from Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Ethical and social risks of harm from Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.390059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.390059Z digest=sha256:2b29931bd9fe4f7bc24574b832fcce73f67fe63fc28ae6256116f339096a9a9c

Observation b2451446-b76f-4e00-a7bd-b54022459dc8 · outbound

This paper cites A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.393696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.393696Z digest=sha256:37d3d4c6c9e158cc79b4f14c5089e8a84fdd50e82492c262bfe7f13dae8cca81

Observation 4f8975f0-ebc4-4b2b-a022-bdf6ac592f42 · outbound

This paper cites Audio Is the Achilles' Heel: Red Teaming Audio Large Multimodal Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Audio Is the Achilles' Heel: Red Teaming Audio Large Multimodal Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.397110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.397110Z digest=sha256:a5a6e6aa94b3a252ab297a9a88cbcf2a732bc749d37d2fd6d5d47109813e25ee

Observation cbd27bd2-1236-41bb-afd4-924eb250e2aa · outbound

This paper cites Jailbreak Attacks and Defenses Against Large Language Models: A Survey.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Jailbreak Attacks and Defenses Against Large Language Models: A Survey

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.400556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.400556Z digest=sha256:14cff806f2aa14850c8d62f3b7c625d356d72454eba51091756152ec69cc7874

Observation 80ed813d-ef3c-4083-beec-a5ac8fd8c19d · outbound

This paper cites Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.403894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.403894Z digest=sha256:49e024d48beda7c6d4953a326be341f5b57392b4c9051abdf05947bf123dcfad

Observation eb2d21fa-7101-47ac-abfc-10b6c850db8f · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.407574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.407574Z digest=sha256:b23755f68f4644ac7d36caf59e65dea5ebe9d0b7fa006e7fbfb9fa1897076521

Observation 5f9352e1-5df3-4545-b8e1-3c6765c05584 · outbound

This paper cites write newline.

"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models write newline

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-09T18:05:52.411107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:05:52.411107Z digest=sha256:a2297b96b6616b8032c3aa1742537d76a45b3c4bf59f7848aaf0361016a84c88

Pith citing papers

Observation 640a865f-55a0-4642-866d-4bccf697773a · inbound

Universal Acoustic Adversarial Attacks for Flexible Control of Speech-LLMs cites this paper.

Universal Acoustic Adversarial Attacks for Flexible Control of Speech-LLMs "I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:19.934200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:19.934200Z digest=sha256:e2a6a1f8d4c034d62657f3bca29c01fbd11ec1890c7874e721da5bad59b1da72

Observation 10f74bdd-f3ab-462e-abf1-6a4fb7fe297d · inbound

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey cites this paper.

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey "I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:34:53.251853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T13:32:57.771753Z digest=sha256:536c93f4ba9021ded5c3fd6fd8967c6e55154883da3e90e0d35d747ad0a36a92

Observation 120fea50-dc99-414f-bb87-b914e55d8ab7 · inbound

Investigating Vulnerabilities and Defenses Against Audio-Visual Attacks: A Comprehensive Survey Emphasizing Multimodal Models cites this paper.

Investigating Vulnerabilities and Defenses Against Audio-Visual Attacks: A Comprehensive Survey Emphasizing Multimodal Models "I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T04:08:42.306422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:08:42.306422Z digest=sha256:d3bd497af5f3b208af547a4752e97ef546ec352ddbe9b0e8c932fae11ad7d39d

Observation 67fbab49-7064-461f-b6c2-c2d307b8a427 · inbound

On Optimizing Multimodal Jailbreaks for Spoken Language Models cites this paper.

On Optimizing Multimodal Jailbreaks for Spoken Language Models "I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-13T22:11:30.473038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:11:30.473038Z digest=sha256:1a04ff94090bd2d45d3b3c40e6e255133b096ded8b60739a28d442cb2035651c

Observation 1194067e-34cb-480e-85ca-69dfcfece2f2 · inbound

Benign Fine-Tuning Breaks Safety Alignment in Audio LLMs cites this paper.

Benign Fine-Tuning Breaks Safety Alignment in Audio LLMs "I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:02:24.902416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T08:01:25.938248Z digest=sha256:0e4101241c4b4ab6fbcb5c687c57d4b1eaffb50c653abd4035bb3f9dce4f3a3f

Observation 8b3a05c7-410e-473e-8f02-9266fb140ce2 · inbound

A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook cites this paper.

A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook "I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models

Reference 166

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:39:48.672400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T07:38:23.099479Z digest=sha256:ca718b8dde3d463fcbe43a85208e95987a5339e3e0304ab854c2666c8b601571

Observation 088c884f-afdc-4661-b5fa-1aa3c686eb4c · inbound

Audio Jailbreaks in Large Audio-Language Models: Taxonomy, Attack-Defense Analysis, and Cost-Aware Evaluation cites this paper.

Audio Jailbreaks in Large Audio-Language Models: Taxonomy, Attack-Defense Analysis, and Cost-Aware Evaluation "I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T05:53:09.563735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T05:44:05.762261Z digest=sha256:138b7bb1038e25cfed8a48272edd79583eb78b9956b6fb23c092c699ec442202