Pith. sign in

Paper Citation Record · LEDGER

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models

As of 8 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 2 inbound Pith citation observations for arXiv:2505.22271.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22271 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:17:25.808065Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T07:14:32.902738Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T07:24:22.504297Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact1
  • verified fuzzy30
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0845ab38-c298-47da-892c-114ca7a9cc98 · outbound

This paper cites Jail- breaking black box large language models in twenty queries.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jail- breaking black box large language models in twenty queries

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:32.148838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:19.788942Z digest=sha256:808735780424ec3f277d42a3e709ea63e164ad4a441fe9984fc70fb71b9fcddc

Observation 1c145bd4-d723-40d5-b022-ac9644772aa2 · outbound

This paper cites Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality.See https://vicuna.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality.See https://vicuna

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:19.865083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:19.865083Z digest=sha256:0cc2fc6d0413d741b8635195ca3b4e5be3c2010e09cfbe64e338e21d44298ce1

Observation c3f63183-8f51-498a-8548-e17478612deb · outbound

This paper cites Security and privacy challenges of large language models: A survey.ACM Computing Surveys, 2024.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Security and privacy challenges of large language models: A survey.ACM Computing Surveys, 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:32.006849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:19.978460Z digest=sha256:e40e4099ad9a23de4d186e46a9f2ad78b140d6169df91bfd7b685d58c310f90b

Observation d8fdbdc4-059f-4bb3-9e02-b421cdfee5b9 · outbound

This paper cites Libre: A practical bayesian approach to adversarial detection.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Libre: A practical bayesian approach to adversarial detection

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:31.760169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:20.096904Z digest=sha256:fee437b017474cad1a5a1ee4ae2832b0a7d4d5b0416fe2f8714fbef7759c2c3c

Observation 28b528c5-886b-46d1-9e51-7b3eaa387197 · outbound

This paper cites Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.181959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.181959Z digest=sha256:6cc85069849c23799fe634ee1fdbb54b1ca4bcb96a7bcee25b5ae6fa2075fa87

Observation c9a16fcf-f021-4a0c-a6cb-addc007f4cbc · outbound

This paper cites FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.261903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.261903Z digest=sha256:5e3344991c885afbf3d151d2012af08275a48f92aed00df7cf9c34dcdf3745db

Observation a25ac3bc-1770-46f9-a19e-f3a7d20911e4 · outbound

This paper cites Eyes closed, safety on: Protecting multimodal llms via image-to-text transformation.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Eyes closed, safety on: Protecting multimodal llms via image-to-text transformation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:31.578334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:20.426606Z digest=sha256:d41af1feec38f64e47c35cabb9645fdb67273c8930d41efd242a561a3062a9c2

Observation 56d8d03a-0245-461d-8177-ffdcaf057196 · outbound

This paper cites Backdoor defense via test-time detecting and repairing.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Backdoor defense via test-time detecting and repairing

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:31.441131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:20.512808Z digest=sha256:202196f01e469fb6b1a35450217d55c7405fb110df2a5d776bb3a6e43dcdb113

Observation 94d78a55-b09e-488b-9eca-0cac1d718256 · outbound

This paper cites LoRA: Low-rank adaptation of large language models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models LoRA: Low-rank adaptation of large language models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:31.213043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:20.596521Z digest=sha256:f25ab793037799174bf36e7297d93275441ee72d339c67dfa29f4e3df7ac284d

Observation ebd8e886-f4c4-4b3e-8b47-b31162f717af · outbound

This paper cites Token-Level Adversarial Prompt Detection Based on Perplexity Measures and Contextual Information.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Token-Level Adversarial Prompt Detection Based on Perplexity Measures and Contextual Information

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.682371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.682371Z digest=sha256:19a986597f19f05fe89e471dfcf409e9a08641d4ab50a5038ef974f5ab3a2aca

Observation 6ce751ab-57da-4788-b53e-d3e10dbe983f · outbound

This paper cites Baseline Defenses for Adversarial Attacks Against Aligned Language Models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Baseline Defenses for Adversarial Attacks Against Aligned Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.741888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.741888Z digest=sha256:c2c700a44a6f7b604b9241fe7c7562e7827978af3401e38e4d2f71b58d848260

Observation b125cd62-7fce-480b-90cf-ee3be2e2cfdc · outbound

This paper cites Mistral 7B.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Mistral 7B

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.801958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.801958Z digest=sha256:e20092cabfd5136f5af55aa78cb3c3e9a5078bdf7ca79d0409e42dab05b83671

Observation fdf4814e-1a02-44cf-b5bc-3f087104aea1 · outbound

This paper cites Jailbreakzoo: Survey, landscapes, and horizons in jailbreaking large language and vision-language models.arXiv preprint arXiv:2407.01599, 2024.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jailbreakzoo: Survey, landscapes, and horizons in jailbreaking large language and vision-language models.arXiv preprint arXiv:2407.01599, 2024

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.872477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.872477Z digest=sha256:98e2299f29a5357b2c1ec1fdb44ae14242980570105856eefc791730536035cc

Observation a9945a9f-ca74-4f71-aefc-a281a6f69a1a · outbound

This paper cites A survey of reinforcement learning from human feedback.arXiv preprint arXiv:2312.14925, 2023.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models A survey of reinforcement learning from human feedback.arXiv preprint arXiv:2312.14925, 2023

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.980016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.980016Z digest=sha256:0ba002227920ddfc8d786df9624851937b05c863538e42f4bd94c27359b06617

Observation 49499c2f-bc9e-40e6-8a97-278451b82a47 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Adam: A Method for Stochastic Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:21.040516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:21.040516Z digest=sha256:fc2e96f87a30dc32f37d1f683426c6c018943f27eddcfc78bb4726c63b0521a3

Observation d3eb8416-463e-4425-b4f4-9836979c048e · outbound

This paper cites Certifying LLM Safety against Adversarial Prompting.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Certifying LLM Safety against Adversarial Prompting

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:21.099918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:21.099918Z digest=sha256:2e51f26db91033f8250318101cf55a98744a72683961809e21609aff7d242530

Observation cca613c9-a9fb-45e4-ba0a-58e1e7eaacc5 · outbound

This paper cites A comprehensive survey on test-time adaptation under distribution shifts.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models A comprehensive survey on test-time adaptation under distribution shifts

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:30.926981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:21.172373Z digest=sha256:72cfae1fea2abce8481099d301d23d3f333cbf18b9e4b38416b54eb36b9a4062

Observation f362cc21-f395-49d3-81e4-6932e4aa58b7 · outbound

This paper cites Improving Adversarial Robustness for 3D Point Cloud Recognition at Test-Time through Purified Self-Training.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Improving Adversarial Robustness for 3D Point Cloud Recognition at Test-Time through Purified Self-Training

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:17:26.486041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:21.252772Z digest=sha256:308f30cceea349f665939ef104feec136c7680e92f955eaa6afb74c6f46a18d2

Observation 56d86f8b-23fe-4037-8a6b-b44366b4e45b · outbound

This paper cites Microsoft coco: Common objects in context.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Microsoft coco: Common objects in context

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:30.735860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:21.335139Z digest=sha256:a2d7262cd16aacfd1c7faef435c227854004fed676eced70fb81b6ef34a71f9b

Observation 81052493-b28b-46d9-88a6-801b8ddb13c9 · outbound

This paper cites Visual instruction tuning.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Visual instruction tuning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:30.563054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:21.358161Z digest=sha256:a102f433ca169feff6fc637a2a7a44c3628eb8b91e502061b31519433330a177

Observation ab48208f-11ce-47fd-b197-841d233cf260 · outbound

This paper cites Improved baselines with visual instruction tuning.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Improved baselines with visual instruction tuning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:30.388861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:21.380289Z digest=sha256:5ecf44a348671a3a4b8c2d67641c26636a7a164ca7485fc60022774685302d3e

Observation fb5e1d04-325b-44fd-af20-b51ab3eb58dc · outbound

This paper cites Llava-next: Improved reasoning, ocr, and world knowledge, January 2024.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Llava-next: Improved reasoning, ocr, and world knowledge, January 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:30.185100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:21.466300Z digest=sha256:3bf0ef6fa42c114f0644af41f4a594959dfd1cad0c602a1764c58eee4a159aa9

Observation ad404526-d415-4378-923c-64374aeebaf4 · outbound

This paper cites Autodan: Generating stealthy jailbreak prompts on aligned large language models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Autodan: Generating stealthy jailbreak prompts on aligned large language models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:30.018038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:21.544145Z digest=sha256:6aa93c2d70773b5f2930c129060cfd85bd400c262e961865a00354fcbc908fd2

Observation c9f5fa92-8fd8-4632-8983-76ee67e6b68d · outbound

This paper cites Mm-safetybench: A benchmark for safety evaluation of multimodal large language models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Mm-safetybench: A benchmark for safety evaluation of multimodal large language models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:29.858044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:21.608032Z digest=sha256:22340fbb9036f6a19ae8df87ecd4b3432c89bb09f67efcae498a34f47d7cc346

Observation 4bfeaadc-76ec-4cae-9a69-21bf527f2e44 · outbound

This paper cites A Comprehensive Overview of Large Language Models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models A Comprehensive Overview of Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:21.675739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:21.675739Z digest=sha256:0c46350d6299c131c7c23f2574acd6fb7ecb0b4a3963d69386aaf59849580b9f

Observation a4f2856b-434f-4975-a618-78152c1cfdb3 · outbound

This paper cites Dad: Data-free adversarial defense at test time.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Dad: Data-free adversarial defense at test time

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:29.660346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:21.768487Z digest=sha256:8806abdfdc27e0a0772c72003a200f443eb7f6b9e0b8b0612fe6b36c4e83ae48

Observation 114fc22b-ea06-41dd-bf1f-25ae276fc8d0 · outbound

This paper cites GPT-4 Technical Report.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models GPT-4 Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:21.904633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:21.904633Z digest=sha256:9fc27d5fc3d54c2e91acd8a0e2b88be7802a26002a7231481d68c856b61b472f

Observation cb30a0ee-b326-4202-967c-4c6538b87461 · outbound

This paper cites Rapid Response: Mitigating LLM Jailbreaks with a Few Examples.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Rapid Response: Mitigating LLM Jailbreaks with a Few Examples

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:22.037860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:22.037860Z digest=sha256:b30a898dd9c5eb253bc39893deecd65a57b623162c0860dd71b29a916b53ccfe

Observation e881e95e-9ce6-4fbb-bec8-4cd8c8fb43df · outbound

This paper cites Instruction Tuning with GPT-4.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Instruction Tuning with GPT-4

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:22.099259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:22.099259Z digest=sha256:f8a35e4ea217e44e07bae047e8dbc3ebad247b1387f4143aadc2d34e60a5ae1a

Observation a402bbf3-7b33-4adc-81cf-cdebf6a9d88b · outbound

This paper cites Llm self defense: By self examination, llms know they are being tricked.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Llm self defense: By self examination, llms know they are being tricked

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:29.519521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:22.205834Z digest=sha256:2883952ca2b5175c2b0823456befdb8b0a884040ef5fbbb02447d1b886977990

Observation 3455ac16-09e2-4e3f-ab96-c22cdccbcb4c · outbound

This paper cites Mllm-protector: Ensuring mllm’s safety without hurting performance.Proc.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Mllm-protector: Ensuring mllm’s safety without hurting performance.Proc

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:29.371293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:22.352340Z digest=sha256:f41d4b4f49d28834b556ab3dfa3caf6da820e4d08069a3cb254c94e70c9f7ef8

Observation 79190684-5be4-426e-ae99-3203d733c65a · outbound

This paper cites Visual adversarial examples jailbreak aligned large language models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Visual adversarial examples jailbreak aligned large language models

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:29.226614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:22.498989Z digest=sha256:cd850f1eb6786459254f069d4e4168075cd914794ee542b09b1d1e2189d98ce3

Observation 167eb778-99ee-4038-858d-aa63acb0cbf6 · outbound

This paper cites Improving language understanding by generative pre-training.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Improving language understanding by generative pre-training

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:22.646829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:22.646829Z digest=sha256:6fbe1cfdfdb5e753b16bd64bfffb142327b59e3e313d1824f85927e58de13c16

Observation 62efe873-765e-45cf-b857-39db2e329bc6 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models High-resolution image synthesis with latent diffusion models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:28.993222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:22.791835Z digest=sha256:8c625bc1cec6229a2908eab0dd4e1993f6cea7aa478c3508938941fd2149422c

Observation c8c15d46-d490-4f00-b15e-6831a2b9fa64 · outbound

This paper cites Protecting Model Adaptation from Trojans in the Unlabeled Data.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Protecting Model Adaptation from Trojans in the Unlabeled Data

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T13:17:26.238790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:22.922491Z digest=sha256:5f25b3bbd93c3661572e3a2e9eac89172a103b7ea4991e492b001e7481c8b10c

Observation 712d64e6-8dba-4ac6-b879-9432bcf405f0 · outbound

This paper cites Learning to summarize with human feedback.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Learning to summarize with human feedback

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:28.789095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:23.111377Z digest=sha256:2a0edcd5f66972434cfa26ff35b605d15e3485f828bf98416ef7da21614f4fd3

Observation 78fe64cb-4df1-4848-8b0b-ecd21f5ca6ab · outbound

This paper cites Measuring the accuracy of diagnostic systems.Science, 240(4857):1285–1293, 1988.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Measuring the accuracy of diagnostic systems.Science, 240(4857):1285–1293, 1988

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:23.236378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:23.236378Z digest=sha256:a20f0a39ebc26c715eec5bdb7bcd2c5376fdca5248f515e827f9b6755996d7a3

Observation 0170707d-2eee-4bc2-9dd7-41195eac145e · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:23.436727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:23.436727Z digest=sha256:3b3d1efac013c4b4f798942a8c3f7c9ec44180f1f26d7395ce6f9f86e5b1ba92

Observation 22d51830-12ec-482c-bd20-d75863969f64 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:23.623355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:23.623355Z digest=sha256:6637b54b7d76864b1f3fc47dbc89577506dd5d70ec9fd2615e738e8267758493

Observation 03373f95-d0b4-4e6e-a721-de326268dd8a · outbound

This paper cites Tent: Fully test-time adaptation by entropy minimization.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Tent: Fully test-time adaptation by entropy minimization

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:28.596615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:23.739463Z digest=sha256:eba8b9919d02fabcc143962b438e9586cb7a373da245bdaa841662af089e8cf8

Observation 7cdfaab2-b3b9-4dd1-8864-4e0ee71cfc06 · outbound

This paper cites Do We Really Need Curated Malicious Data for Safety Alignment in Multi-modal Large Language Models?.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Do We Really Need Curated Malicious Data for Safety Alignment in Multi-modal Large Language Models?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:23.839067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:23.839067Z digest=sha256:a590624573360d0172d99a09f3738069034ab3823b8bf300948305537d034e27

Observation 89992326-ccef-4775-a9f8-511491b8817f · outbound

This paper cites Defending llms against jailbreaking attacks via backtranslation.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Defending llms against jailbreaking attacks via backtranslation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:28.389761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:23.937944Z digest=sha256:4d90eb166808e42a0171fe6afedf52429c9aaef94d2d9d13ee4dd4e8d5780492

Observation 06907291-a297-44f6-9858-e073b879b825 · outbound

This paper cites Adashield: Safeguarding multimodal large language models from structure-based attack via adaptive shield prompting.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Adashield: Safeguarding multimodal large language models from structure-based attack via adaptive shield prompting

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:28.160634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:24.019287Z digest=sha256:530e44394870e925c38a1d9486ce1adb19d7a9ba924f606df358b8d5e447145d

Observation 73d2de4a-8bd0-44b0-b6e4-7b1dec2610c8 · outbound

This paper cites Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:24.102566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:24.102566Z digest=sha256:114d7b95a87bfc93f2d3ef8ea8cbcf0936d71697288402482d728a0a48d94789

Observation 471dcb12-f247-4482-be11-f598d5a4218c · outbound

This paper cites Defending chatgpt against jailbreak attack via self-reminders.Nature Machine Intelligence, 5(12):1486–1496, 2023.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Defending chatgpt against jailbreak attack via self-reminders.Nature Machine Intelligence, 5(12):1486–1496, 2023

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:24.264935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:24.264935Z digest=sha256:77ada33aa0d87fbb55e119c59dc1fa711af4a817d89d853d3f8f6ffccc41f6d4

Observation cc86ef63-d0e6-426b-afdd-275cea67d226 · outbound

This paper cites Gradsafe: Detecting jailbreak prompts for llms via safety-critical gradient analysis.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Gradsafe: Detecting jailbreak prompts for llms via safety-critical gradient analysis

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:27.939953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:24.416405Z digest=sha256:50ddfb1e4294eefaf6c7acab7444687b4086bb0b72c8601b8f34cc3923e495d5

Observation 9ef5f3f6-8b65-46ac-a09b-3c2d6e8fde5a · outbound

This paper cites Wizardlm: Empowering large pre-trained language models to follow complex instructions.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Wizardlm: Empowering large pre-trained language models to follow complex instructions

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:27.801491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:24.547046Z digest=sha256:088856a8e2a49c2879bdd6544e3788c2d41cb56863d59f9cdd267ade4fcc92df

Observation 83b8baf3-1190-47cd-a8ea-3fbfb5515aac · outbound

This paper cites Jailbreak Attacks and Defenses Against Large Language Models: A Survey.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jailbreak Attacks and Defenses Against Large Language Models: A Survey

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:24.669043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:24.669043Z digest=sha256:ff89004c86f502af1baab45ecc710ac43532ebcfc7f0b0ec61d164290a52b5a4

Observation cd5f35cc-4ef6-4d18-91ea-2720e37edaf4 · outbound

This paper cites Stamp: Outlier-aware test-time adaptation with stable memory replay.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Stamp: Outlier-aware test-time adaptation with stable memory replay

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:27.641319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:24.753149Z digest=sha256:479711943aca479c4d67aa376363e85d2a844efe81173d6bd1d0b8f697b1e857

Observation b82128b5-df7b-4649-91c9-431c1e7cad3b · outbound

This paper cites Instruction tuning for large language models: A survey.arXiv preprint arXiv:2308.10792, 2023.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Instruction tuning for large language models: A survey.arXiv preprint arXiv:2308.10792, 2023

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:24.856171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:24.856171Z digest=sha256:4f09675cf8072bca5fc449df29be315b1a09b4d7bbd2bac0dbd85fd4acefe396

Observation ddb5be6a-e6e3-426e-8d88-ce0dd430be2c · outbound

This paper cites JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:24.937849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:24.937849Z digest=sha256:b2b0260065efc696d61d6c1cfdef7af5d2de288fb2176ca88a45276e2fed333d

Observation f98d0ec3-61f0-41e7-a35c-283aa7e4c457 · outbound

This paper cites Defending large language models against jailbreaking attacks through goal prioritization.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Defending large language models against jailbreaking attacks through goal prioritization

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:27.421311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:25.086670Z digest=sha256:a211108b0ae72d7933578b5950d8b4c1b04e6938696b76d84222f4d807f70e96

Observation ac80896e-fa10-4990-a10f-6527abf9c6b1 · outbound

This paper cites The first to know: How token distributions reveal hidden knowledge in large vision-language models? InProc.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models The first to know: How token distributions reveal hidden knowledge in large vision-language models? InProc

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:27.275094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:25.232714Z digest=sha256:da2e4c200829b3ed0ed357f3cb93866fe27bb8b8e4ce53b4b80d991a57d05232

Observation 1e8403bd-a55d-4ef7-9dc1-b40ee193785a · outbound

This paper cites A Survey of Large Language Models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models A Survey of Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:25.320167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:25.320167Z digest=sha256:af4d4bc747ca63fb7bd63339c66d33c0ec4974b2ae3730971e19daf1cdff7da3

Observation 637246e7-529b-4d42-a374-0a774e2dfdf5 · outbound

This paper cites Improved few-shot jailbreaking can circumvent aligned language models and their defenses.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Improved few-shot jailbreaking can circumvent aligned language models and their defenses

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:27.138624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:25.541068Z digest=sha256:7d899398b86b09dd7b4a89290c3a39b64f4b907f43ed05332ae7e48535d57bdb

Observation 7cb6da9d-22c6-4b49-9033-1ef6acbb4384 · outbound

This paper cites Minigpt-4: Enhancing vision- language understanding with advanced large language models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Minigpt-4: Enhancing vision- language understanding with advanced large language models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:26.974341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:25.663621Z digest=sha256:d64f0545a4e818fc0adaf9b6ef97a6121d0811dee2abddad1c2376494500e767

Observation 5039fe24-3486-46e2-b31e-07ae8f4a358b · outbound

This paper cites Safety fine-tuning at (almost) no cost: A baseline for vision large language models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Safety fine-tuning at (almost) no cost: A baseline for vision large language models

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:26.826571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:17:25.755330Z digest=sha256:d4ae728a21fc93c27198c60ed2795127e7930ad5e8818b07c92fc57e221edb2e

Observation 4461d1a0-d860-431a-91e7-0ce83f9ad5b6 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:25.808065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:25.808065Z digest=sha256:616bf9450b8a3738b319b2958e3361cd4c7f6faf258ee4c54f86fe714d8db497

Pith citing papers

Observation 1b1d1d6c-e1c3-4414-b792-8952e4fd56f8 · inbound

Understanding and Mitigating Spurious Signal Amplification in Test-Time Reinforcement Learning for Math Reasoning cites this paper.

Understanding and Mitigating Spurious Signal Amplification in Test-Time Reinforcement Learning for Math Reasoning Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:07.293158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T22:55:19.471307Z digest=sha256:5d00f8574078237807e035e325a0d331f19512d41244c2443e0d524e5bcc250f

Observation ca574ac8-d05a-48c1-b5d0-5f3862ed884f · inbound

On the Vulnerability of Parameter-Level Defenses to Model Merging cites this paper.

On the Vulnerability of Parameter-Level Defenses to Model Merging Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:24:22.505907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T07:14:32.902738Z digest=sha256:fca9ade8a5b2fd126fd182df5c077e7c8437fdc1cdccf763ddccf419d6476799