Pith. sign in

Paper Citation Record · LEDGER

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models

As of 10 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 2 inbound Pith citation observations for arXiv:2505.22271.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22271 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:17:25.808065Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T07:14:32.902738Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T07:24:22.504297Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact1
  • verified fuzzy30
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0845ab38-c298-47da-892c-114ca7a9cc98 · outbound

This paper cites Jail- breaking black box large language models in twenty queries.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jail- breaking black box large language models in twenty queries

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:32.148838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:19.788942Z digest=sha256:f47cc30ebfc214f7a8ab4564387348fa2aafc5572cc9b789cf226b2ba43bd017

Observation 1c145bd4-d723-40d5-b022-ac9644772aa2 · outbound

This paper cites Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality.See https://vicuna.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality.See https://vicuna

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:19.865083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:19.865083Z digest=sha256:0cc2fc6d0413d741b8635195ca3b4e5be3c2010e09cfbe64e338e21d44298ce1

Observation c3f63183-8f51-498a-8548-e17478612deb · outbound

This paper cites Security and privacy challenges of large language models: A survey.ACM Computing Surveys, 2024.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Security and privacy challenges of large language models: A survey.ACM Computing Surveys, 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:32.006849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:19.978460Z digest=sha256:88cfcdc9354af6ebd1c52decb80b889b7d2f7f325571c29bd96f36cc7b4a4f07

Observation d8fdbdc4-059f-4bb3-9e02-b421cdfee5b9 · outbound

This paper cites Libre: A practical bayesian approach to adversarial detection.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Libre: A practical bayesian approach to adversarial detection

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:31.760169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:20.096904Z digest=sha256:55afa5b359b8daa6b082b3ddbf5102a88059fa82c6ad4d2537de147f2ca97920

Observation 28b528c5-886b-46d1-9e51-7b3eaa387197 · outbound

This paper cites Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.181959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.181959Z digest=sha256:6cc85069849c23799fe634ee1fdbb54b1ca4bcb96a7bcee25b5ae6fa2075fa87

Observation c9a16fcf-f021-4a0c-a6cb-addc007f4cbc · outbound

This paper cites FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.261903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.261903Z digest=sha256:5e3344991c885afbf3d151d2012af08275a48f92aed00df7cf9c34dcdf3745db

Observation a25ac3bc-1770-46f9-a19e-f3a7d20911e4 · outbound

This paper cites Eyes closed, safety on: Protecting multimodal llms via image-to-text transformation.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Eyes closed, safety on: Protecting multimodal llms via image-to-text transformation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:31.578334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:20.426606Z digest=sha256:79abf02e68f1f0956eb5a80c8e793806ceb0c638a93f0bda01dfdefefc79b569

Observation 56d8d03a-0245-461d-8177-ffdcaf057196 · outbound

This paper cites Backdoor defense via test-time detecting and repairing.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Backdoor defense via test-time detecting and repairing

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:31.441131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:20.512808Z digest=sha256:db0a63cb7643368c7e29d1822a9e07e098b0f09432c74dc2c6c6c524edae86ec

Observation 94d78a55-b09e-488b-9eca-0cac1d718256 · outbound

This paper cites LoRA: Low-rank adaptation of large language models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models LoRA: Low-rank adaptation of large language models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:31.213043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:20.596521Z digest=sha256:75e8115180f62ef0843a19950c9fc7cf6346fa8ad3d56e96267ecf32f775060b

Observation ebd8e886-f4c4-4b3e-8b47-b31162f717af · outbound

This paper cites Token-Level Adversarial Prompt Detection Based on Perplexity Measures and Contextual Information.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Token-Level Adversarial Prompt Detection Based on Perplexity Measures and Contextual Information

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.682371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.682371Z digest=sha256:19a986597f19f05fe89e471dfcf409e9a08641d4ab50a5038ef974f5ab3a2aca

Observation 6ce751ab-57da-4788-b53e-d3e10dbe983f · outbound

This paper cites Baseline Defenses for Adversarial Attacks Against Aligned Language Models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Baseline Defenses for Adversarial Attacks Against Aligned Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.741888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.741888Z digest=sha256:c2c700a44a6f7b604b9241fe7c7562e7827978af3401e38e4d2f71b58d848260

Observation b125cd62-7fce-480b-90cf-ee3be2e2cfdc · outbound

This paper cites Mistral 7B.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Mistral 7B

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.801958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.801958Z digest=sha256:e20092cabfd5136f5af55aa78cb3c3e9a5078bdf7ca79d0409e42dab05b83671

Observation fdf4814e-1a02-44cf-b5bc-3f087104aea1 · outbound

This paper cites Jailbreakzoo: Survey, landscapes, and horizons in jailbreaking large language and vision-language models.arXiv preprint arXiv:2407.01599, 2024.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jailbreakzoo: Survey, landscapes, and horizons in jailbreaking large language and vision-language models.arXiv preprint arXiv:2407.01599, 2024

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.872477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.872477Z digest=sha256:98e2299f29a5357b2c1ec1fdb44ae14242980570105856eefc791730536035cc

Observation a9945a9f-ca74-4f71-aefc-a281a6f69a1a · outbound

This paper cites A survey of reinforcement learning from human feedback.arXiv preprint arXiv:2312.14925, 2023.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models A survey of reinforcement learning from human feedback.arXiv preprint arXiv:2312.14925, 2023

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:20.980016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:20.980016Z digest=sha256:0ba002227920ddfc8d786df9624851937b05c863538e42f4bd94c27359b06617

Observation 49499c2f-bc9e-40e6-8a97-278451b82a47 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Adam: A Method for Stochastic Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:21.040516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:21.040516Z digest=sha256:fc2e96f87a30dc32f37d1f683426c6c018943f27eddcfc78bb4726c63b0521a3

Observation d3eb8416-463e-4425-b4f4-9836979c048e · outbound

This paper cites Certifying LLM Safety against Adversarial Prompting.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Certifying LLM Safety against Adversarial Prompting

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:21.099918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:21.099918Z digest=sha256:2e51f26db91033f8250318101cf55a98744a72683961809e21609aff7d242530

Observation cca613c9-a9fb-45e4-ba0a-58e1e7eaacc5 · outbound

This paper cites A comprehensive survey on test-time adaptation under distribution shifts.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models A comprehensive survey on test-time adaptation under distribution shifts

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:30.926981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:21.172373Z digest=sha256:e292bf37fe22424963843051765f40b6a2ce9636008592e6237bf0f2ebb304fa

Observation f362cc21-f395-49d3-81e4-6932e4aa58b7 · outbound

This paper cites Improving Adversarial Robustness for 3D Point Cloud Recognition at Test-Time through Purified Self-Training.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Improving Adversarial Robustness for 3D Point Cloud Recognition at Test-Time through Purified Self-Training

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:17:26.486041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:21.252772Z digest=sha256:38e97272ddcb91a6b77a93f120cbf205d5bd293a108ff52624c01c729a151ed3

Observation 56d86f8b-23fe-4037-8a6b-b44366b4e45b · outbound

This paper cites Microsoft coco: Common objects in context.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Microsoft coco: Common objects in context

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:30.735860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:21.335139Z digest=sha256:5b973c2fd1faaa610a4b159a0d6ed7d37c28d5e7d4d8cef83e7b1a8b2ab6fd53

Observation 81052493-b28b-46d9-88a6-801b8ddb13c9 · outbound

This paper cites Visual instruction tuning.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Visual instruction tuning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:30.563054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:21.358161Z digest=sha256:bd972ad0655572db1848dc8c05a8b2efcbb0f47b502612bd57bbdf1a777fa244

Observation ab48208f-11ce-47fd-b197-841d233cf260 · outbound

This paper cites Improved baselines with visual instruction tuning.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Improved baselines with visual instruction tuning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:30.388861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:21.380289Z digest=sha256:4dc9fa9422c28ad0f65599dd740032bc2f6539111ee53c8d3f88c118ed3eda71

Observation fb5e1d04-325b-44fd-af20-b51ab3eb58dc · outbound

This paper cites Llava-next: Improved reasoning, ocr, and world knowledge, January 2024.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Llava-next: Improved reasoning, ocr, and world knowledge, January 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:30.185100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:21.466300Z digest=sha256:8a98666cd3b22490377ef2ad1a235b2e7bd992c898a49cf19ce53df02bdb699e

Observation ad404526-d415-4378-923c-64374aeebaf4 · outbound

This paper cites Autodan: Generating stealthy jailbreak prompts on aligned large language models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Autodan: Generating stealthy jailbreak prompts on aligned large language models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:30.018038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:21.544145Z digest=sha256:043ace32382c0e13781edb53ec17e2837f89a5b64c9450a137c7dfb2e1069fa6

Observation c9f5fa92-8fd8-4632-8983-76ee67e6b68d · outbound

This paper cites Mm-safetybench: A benchmark for safety evaluation of multimodal large language models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Mm-safetybench: A benchmark for safety evaluation of multimodal large language models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:29.858044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:21.608032Z digest=sha256:07352b7402737ee058ad19286c172900904fc92581682ec04c57b76f0a22556d

Observation 4bfeaadc-76ec-4cae-9a69-21bf527f2e44 · outbound

This paper cites A Comprehensive Overview of Large Language Models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models A Comprehensive Overview of Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:21.675739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:21.675739Z digest=sha256:0c46350d6299c131c7c23f2574acd6fb7ecb0b4a3963d69386aaf59849580b9f

Observation a4f2856b-434f-4975-a618-78152c1cfdb3 · outbound

This paper cites Dad: Data-free adversarial defense at test time.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Dad: Data-free adversarial defense at test time

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:29.660346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:21.768487Z digest=sha256:e70bdb99896245ed7a70035b9e82d7f4de260d2f4c2f017b62835711d2116713

Observation 114fc22b-ea06-41dd-bf1f-25ae276fc8d0 · outbound

This paper cites GPT-4 Technical Report.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models GPT-4 Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:21.904633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:21.904633Z digest=sha256:9fc27d5fc3d54c2e91acd8a0e2b88be7802a26002a7231481d68c856b61b472f

Observation cb30a0ee-b326-4202-967c-4c6538b87461 · outbound

This paper cites Rapid Response: Mitigating LLM Jailbreaks with a Few Examples.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Rapid Response: Mitigating LLM Jailbreaks with a Few Examples

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:22.037860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:22.037860Z digest=sha256:b30a898dd9c5eb253bc39893deecd65a57b623162c0860dd71b29a916b53ccfe

Observation e881e95e-9ce6-4fbb-bec8-4cd8c8fb43df · outbound

This paper cites Instruction Tuning with GPT-4.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Instruction Tuning with GPT-4

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:22.099259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:22.099259Z digest=sha256:f8a35e4ea217e44e07bae047e8dbc3ebad247b1387f4143aadc2d34e60a5ae1a

Observation a402bbf3-7b33-4adc-81cf-cdebf6a9d88b · outbound

This paper cites Llm self defense: By self examination, llms know they are being tricked.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Llm self defense: By self examination, llms know they are being tricked

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:29.519521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:22.205834Z digest=sha256:0219eb974bdc8ee285da2a7353cbcb8aa0a9af49b9fdd72aac060bb10c6a3db6

Observation 3455ac16-09e2-4e3f-ab96-c22cdccbcb4c · outbound

This paper cites Mllm-protector: Ensuring mllm’s safety without hurting performance.Proc.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Mllm-protector: Ensuring mllm’s safety without hurting performance.Proc

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:29.371293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:22.352340Z digest=sha256:2aee0db9ea4b925bd9346034f4134de49dfbd5f4b7a99f716727848cda731e5b

Observation 79190684-5be4-426e-ae99-3203d733c65a · outbound

This paper cites Visual adversarial examples jailbreak aligned large language models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Visual adversarial examples jailbreak aligned large language models

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:29.226614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:22.498989Z digest=sha256:4470813f8aef48e58659b10426afdb2821dd254f7a81c6fc6610e7140f0fe95d

Observation 167eb778-99ee-4038-858d-aa63acb0cbf6 · outbound

This paper cites Improving language understanding by generative pre-training.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Improving language understanding by generative pre-training

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:22.646829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:22.646829Z digest=sha256:6fbe1cfdfdb5e753b16bd64bfffb142327b59e3e313d1824f85927e58de13c16

Observation 62efe873-765e-45cf-b857-39db2e329bc6 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models High-resolution image synthesis with latent diffusion models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:28.993222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:22.791835Z digest=sha256:9d5ee7f1f6d5dfa9b8ddcc1b4da0d65312dbd15ddb986f81eb9ce89fc46bb830

Observation c8c15d46-d490-4f00-b15e-6831a2b9fa64 · outbound

This paper cites Protecting Model Adaptation from Trojans in the Unlabeled Data.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Protecting Model Adaptation from Trojans in the Unlabeled Data

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T13:17:26.238790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:22.922491Z digest=sha256:0ae0ff52191623425556fa096d0a4a12bc0dde50788122c6a208476fe745b4c8

Observation 712d64e6-8dba-4ac6-b879-9432bcf405f0 · outbound

This paper cites Learning to summarize with human feedback.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Learning to summarize with human feedback

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:28.789095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:23.111377Z digest=sha256:4177977de662fbbd1cefa741d9bc2335e34d517829fe0427ba73017699e714db

Observation 78fe64cb-4df1-4848-8b0b-ecd21f5ca6ab · outbound

This paper cites Measuring the accuracy of diagnostic systems.Science, 240(4857):1285–1293, 1988.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Measuring the accuracy of diagnostic systems.Science, 240(4857):1285–1293, 1988

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:23.236378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:23.236378Z digest=sha256:a20f0a39ebc26c715eec5bdb7bcd2c5376fdca5248f515e827f9b6755996d7a3

Observation 0170707d-2eee-4bc2-9dd7-41195eac145e · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:23.436727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:23.436727Z digest=sha256:3b3d1efac013c4b4f798942a8c3f7c9ec44180f1f26d7395ce6f9f86e5b1ba92

Observation 22d51830-12ec-482c-bd20-d75863969f64 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:23.623355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:23.623355Z digest=sha256:6637b54b7d76864b1f3fc47dbc89577506dd5d70ec9fd2615e738e8267758493

Observation 03373f95-d0b4-4e6e-a721-de326268dd8a · outbound

This paper cites Tent: Fully test-time adaptation by entropy minimization.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Tent: Fully test-time adaptation by entropy minimization

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:28.596615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:23.739463Z digest=sha256:3ab1e7c2f07b34c91e533c6a3a3ce000a9840948cc3daf5e87d6048e14d774e2

Observation 7cdfaab2-b3b9-4dd1-8864-4e0ee71cfc06 · outbound

This paper cites Do We Really Need Curated Malicious Data for Safety Alignment in Multi-modal Large Language Models?.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Do We Really Need Curated Malicious Data for Safety Alignment in Multi-modal Large Language Models?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:23.839067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:23.839067Z digest=sha256:a590624573360d0172d99a09f3738069034ab3823b8bf300948305537d034e27

Observation 89992326-ccef-4775-a9f8-511491b8817f · outbound

This paper cites Defending llms against jailbreaking attacks via backtranslation.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Defending llms against jailbreaking attacks via backtranslation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:28.389761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:23.937944Z digest=sha256:102ac8bb024ce9903b3d4fdf3d3ed597979ac38a6b9cf5a0d5b06ee3784005ca

Observation 06907291-a297-44f6-9858-e073b879b825 · outbound

This paper cites Adashield: Safeguarding multimodal large language models from structure-based attack via adaptive shield prompting.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Adashield: Safeguarding multimodal large language models from structure-based attack via adaptive shield prompting

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:28.160634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:24.019287Z digest=sha256:61260bf45211c20acb6d3532e7dbae7079e011cb0873cec8f03cf4075b166670

Observation 73d2de4a-8bd0-44b0-b6e4-7b1dec2610c8 · outbound

This paper cites Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:24.102566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:24.102566Z digest=sha256:114d7b95a87bfc93f2d3ef8ea8cbcf0936d71697288402482d728a0a48d94789

Observation 471dcb12-f247-4482-be11-f598d5a4218c · outbound

This paper cites Defending chatgpt against jailbreak attack via self-reminders.Nature Machine Intelligence, 5(12):1486–1496, 2023.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Defending chatgpt against jailbreak attack via self-reminders.Nature Machine Intelligence, 5(12):1486–1496, 2023

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:24.264935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:24.264935Z digest=sha256:77ada33aa0d87fbb55e119c59dc1fa711af4a817d89d853d3f8f6ffccc41f6d4

Observation cc86ef63-d0e6-426b-afdd-275cea67d226 · outbound

This paper cites Gradsafe: Detecting jailbreak prompts for llms via safety-critical gradient analysis.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Gradsafe: Detecting jailbreak prompts for llms via safety-critical gradient analysis

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:27.939953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:24.416405Z digest=sha256:58ed9624b2394bc5be014d8b514290c17e87494483110bc0a85a24d413ac40e4

Observation 9ef5f3f6-8b65-46ac-a09b-3c2d6e8fde5a · outbound

This paper cites Wizardlm: Empowering large pre-trained language models to follow complex instructions.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Wizardlm: Empowering large pre-trained language models to follow complex instructions

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:27.801491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:24.547046Z digest=sha256:ad7d9bb7dde0b91714a18bdeb5543de6029e186e29f14d80ff4feb7241d6cbf8

Observation 83b8baf3-1190-47cd-a8ea-3fbfb5515aac · outbound

This paper cites Jailbreak Attacks and Defenses Against Large Language Models: A Survey.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jailbreak Attacks and Defenses Against Large Language Models: A Survey

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:24.669043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:24.669043Z digest=sha256:ff89004c86f502af1baab45ecc710ac43532ebcfc7f0b0ec61d164290a52b5a4

Observation cd5f35cc-4ef6-4d18-91ea-2720e37edaf4 · outbound

This paper cites Stamp: Outlier-aware test-time adaptation with stable memory replay.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Stamp: Outlier-aware test-time adaptation with stable memory replay

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:27.641319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:24.753149Z digest=sha256:9d66b6abc8a4204665a6cf91ad32bde692ddae178fcf77a7d2a2cc33077b95fc

Observation b82128b5-df7b-4649-91c9-431c1e7cad3b · outbound

This paper cites Instruction tuning for large language models: A survey.arXiv preprint arXiv:2308.10792, 2023.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Instruction tuning for large language models: A survey.arXiv preprint arXiv:2308.10792, 2023

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:24.856171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:24.856171Z digest=sha256:4f09675cf8072bca5fc449df29be315b1a09b4d7bbd2bac0dbd85fd4acefe396

Observation ddb5be6a-e6e3-426e-8d88-ce0dd430be2c · outbound

This paper cites JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:24.937849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:24.937849Z digest=sha256:b2b0260065efc696d61d6c1cfdef7af5d2de288fb2176ca88a45276e2fed333d

Observation f98d0ec3-61f0-41e7-a35c-283aa7e4c457 · outbound

This paper cites Defending large language models against jailbreaking attacks through goal prioritization.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Defending large language models against jailbreaking attacks through goal prioritization

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:27.421311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:25.086670Z digest=sha256:ce8b3745bfd2c009c21ece5a42ec5fd10f36a9c147b4e5fd028ce9d04ad79442

Observation ac80896e-fa10-4990-a10f-6527abf9c6b1 · outbound

This paper cites The first to know: How token distributions reveal hidden knowledge in large vision-language models? InProc.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models The first to know: How token distributions reveal hidden knowledge in large vision-language models? InProc

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:27.275094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:25.232714Z digest=sha256:1662694a6fed45a77bc22552a9bbcb34fcd7a08553837d9b942fec627a4a5f27

Observation 1e8403bd-a55d-4ef7-9dc1-b40ee193785a · outbound

This paper cites A Survey of Large Language Models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models A Survey of Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:25.320167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:25.320167Z digest=sha256:af4d4bc747ca63fb7bd63339c66d33c0ec4974b2ae3730971e19daf1cdff7da3

Observation 637246e7-529b-4d42-a374-0a774e2dfdf5 · outbound

This paper cites Improved few-shot jailbreaking can circumvent aligned language models and their defenses.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Improved few-shot jailbreaking can circumvent aligned language models and their defenses

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:27.138624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:25.541068Z digest=sha256:f7da01589234ac6bd5ab23a5e96189f421b5b9ea85b74a6924756ebb1819177a

Observation 7cb6da9d-22c6-4b49-9033-1ef6acbb4384 · outbound

This paper cites Minigpt-4: Enhancing vision- language understanding with advanced large language models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Minigpt-4: Enhancing vision- language understanding with advanced large language models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:26.974341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:25.663621Z digest=sha256:01e9eb7497753bc7ad141b60015a3c44cdd3a247e13496fa15aa88b421f07e6d

Observation 5039fe24-3486-46e2-b31e-07ae8f4a358b · outbound

This paper cites Safety fine-tuning at (almost) no cost: A baseline for vision large language models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Safety fine-tuning at (almost) no cost: A baseline for vision large language models

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:17:26.826571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:17:25.755330Z digest=sha256:9a6f8e463b77aabc371c22e22e70611335d3e2261891ef2f290236db9c093fe7

Observation 4461d1a0-d860-431a-91e7-0ce83f9ad5b6 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:25.808065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:25.808065Z digest=sha256:616bf9450b8a3738b319b2958e3361cd4c7f6faf258ee4c54f86fe714d8db497

Pith citing papers

Observation 1b1d1d6c-e1c3-4414-b792-8952e4fd56f8 · inbound

Understanding and Mitigating Spurious Signal Amplification in Test-Time Reinforcement Learning for Math Reasoning cites this paper.

Understanding and Mitigating Spurious Signal Amplification in Test-Time Reinforcement Learning for Math Reasoning Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:07.293158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T22:55:19.471307Z digest=sha256:a109de4b0e84658d1f76fb990da6625e53cc61e0fa2171c6f830586d44025570

Observation ca574ac8-d05a-48c1-b5d0-5f3862ed884f · inbound

On the Vulnerability of Parameter-Level Defenses to Model Merging cites this paper.

On the Vulnerability of Parameter-Level Defenses to Model Merging Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:24:22.505907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T07:14:32.902738Z digest=sha256:35459c11fcaf495fa68cc711f5819537f38636b1428e7ae64fafd54d0e5063fa