Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:17:25.808065Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 2 inbound Pith citation observations for arXiv:2505.22271.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:17:25.808065Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-30T07:14:32.902738Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-30T07:24:22.504297Z
58 of 58 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0845ab38-c298-47da-892c-114ca7a9cc98 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jail- breaking black box large language models in twenty queries
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1c145bd4-d723-40d5-b022-ac9644772aa2 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality.See https://vicuna
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3f63183-8f51-498a-8548-e17478612deb · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Security and privacy challenges of large language models: A survey.ACM Computing Surveys, 2024
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d8fdbdc4-059f-4bb3-9e02-b421cdfee5b9 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Libre: A practical bayesian approach to adversarial detection
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 28b528c5-886b-46d1-9e51-7b3eaa387197 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9a16fcf-f021-4a0c-a6cb-addc007f4cbc · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a25ac3bc-1770-46f9-a19e-f3a7d20911e4 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Eyes closed, safety on: Protecting multimodal llms via image-to-text transformation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 56d8d03a-0245-461d-8177-ffdcaf057196 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Backdoor defense via test-time detecting and repairing
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 94d78a55-b09e-488b-9eca-0cac1d718256 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models LoRA: Low-rank adaptation of large language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ebd8e886-f4c4-4b3e-8b47-b31162f717af · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Token-Level Adversarial Prompt Detection Based on Perplexity Measures and Contextual Information
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ce751ab-57da-4788-b53e-d3e10dbe983f · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Baseline Defenses for Adversarial Attacks Against Aligned Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b125cd62-7fce-480b-90cf-ee3be2e2cfdc · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Mistral 7B
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdf4814e-1a02-44cf-b5bc-3f087104aea1 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jailbreakzoo: Survey, landscapes, and horizons in jailbreaking large language and vision-language models.arXiv preprint arXiv:2407.01599, 2024
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9945a9f-ca74-4f71-aefc-a281a6f69a1a · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models A survey of reinforcement learning from human feedback.arXiv preprint arXiv:2312.14925, 2023
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49499c2f-bc9e-40e6-8a97-278451b82a47 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Adam: A Method for Stochastic Optimization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3eb8416-463e-4425-b4f4-9836979c048e · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Certifying LLM Safety against Adversarial Prompting
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cca613c9-a9fb-45e4-ba0a-58e1e7eaacc5 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models A comprehensive survey on test-time adaptation under distribution shifts
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f362cc21-f395-49d3-81e4-6932e4aa58b7 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Improving Adversarial Robustness for 3D Point Cloud Recognition at Test-Time through Purified Self-Training
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 56d86f8b-23fe-4037-8a6b-b44366b4e45b · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Microsoft coco: Common objects in context
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 81052493-b28b-46d9-88a6-801b8ddb13c9 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Visual instruction tuning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ab48208f-11ce-47fd-b197-841d233cf260 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Improved baselines with visual instruction tuning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fb5e1d04-325b-44fd-af20-b51ab3eb58dc · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Llava-next: Improved reasoning, ocr, and world knowledge, January 2024
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ad404526-d415-4378-923c-64374aeebaf4 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Autodan: Generating stealthy jailbreak prompts on aligned large language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c9f5fa92-8fd8-4632-8983-76ee67e6b68d · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Mm-safetybench: A benchmark for safety evaluation of multimodal large language models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4bfeaadc-76ec-4cae-9a69-21bf527f2e44 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models A Comprehensive Overview of Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4f2856b-434f-4975-a618-78152c1cfdb3 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Dad: Data-free adversarial defense at test time
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 114fc22b-ea06-41dd-bf1f-25ae276fc8d0 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models GPT-4 Technical Report
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb30a0ee-b326-4202-967c-4c6538b87461 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Rapid Response: Mitigating LLM Jailbreaks with a Few Examples
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e881e95e-9ce6-4fbb-bec8-4cd8c8fb43df · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Instruction Tuning with GPT-4
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a402bbf3-7b33-4adc-81cf-cdebf6a9d88b · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Llm self defense: By self examination, llms know they are being tricked
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3455ac16-09e2-4e3f-ab96-c22cdccbcb4c · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Mllm-protector: Ensuring mllm’s safety without hurting performance.Proc
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 79190684-5be4-426e-ae99-3203d733c65a · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Visual adversarial examples jailbreak aligned large language models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 167eb778-99ee-4038-858d-aa63acb0cbf6 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Improving language understanding by generative pre-training
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62efe873-765e-45cf-b857-39db2e329bc6 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models High-resolution image synthesis with latent diffusion models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c8c15d46-d490-4f00-b15e-6831a2b9fa64 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Protecting Model Adaptation from Trojans in the Unlabeled Data
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 712d64e6-8dba-4ac6-b879-9432bcf405f0 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Learning to summarize with human feedback
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 78fe64cb-4df1-4848-8b0b-ecd21f5ca6ab · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Measuring the accuracy of diagnostic systems.Science, 240(4857):1285–1293, 1988
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0170707d-2eee-4bc2-9dd7-41195eac145e · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Gemini: A Family of Highly Capable Multimodal Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22d51830-12ec-482c-bd20-d75863969f64 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03373f95-d0b4-4e6e-a721-de326268dd8a · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Tent: Fully test-time adaptation by entropy minimization
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7cdfaab2-b3b9-4dd1-8864-4e0ee71cfc06 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Do We Really Need Curated Malicious Data for Safety Alignment in Multi-modal Large Language Models?
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89992326-ccef-4775-a9f8-511491b8817f · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Defending llms against jailbreaking attacks via backtranslation
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 06907291-a297-44f6-9858-e073b879b825 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Adashield: Safeguarding multimodal large language models from structure-based attack via adaptive shield prompting
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 73d2de4a-8bd0-44b0-b6e4-7b1dec2610c8 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 471dcb12-f247-4482-be11-f598d5a4218c · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Defending chatgpt against jailbreak attack via self-reminders.Nature Machine Intelligence, 5(12):1486–1496, 2023
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc86ef63-d0e6-426b-afdd-275cea67d226 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Gradsafe: Detecting jailbreak prompts for llms via safety-critical gradient analysis
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9ef5f3f6-8b65-46ac-a09b-3c2d6e8fde5a · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Wizardlm: Empowering large pre-trained language models to follow complex instructions
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 83b8baf3-1190-47cd-a8ea-3fbfb5515aac · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Jailbreak Attacks and Defenses Against Large Language Models: A Survey
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd5f35cc-4ef6-4d18-91ea-2720e37edaf4 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Stamp: Outlier-aware test-time adaptation with stable memory replay
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b82128b5-df7b-4649-91c9-431c1e7cad3b · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Instruction tuning for large language models: A survey.arXiv preprint arXiv:2308.10792, 2023
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddb5be6a-e6e3-426e-8d88-ce0dd430be2c · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f98d0ec3-61f0-41e7-a35c-283aa7e4c457 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Defending large language models against jailbreaking attacks through goal prioritization
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ac80896e-fa10-4990-a10f-6527abf9c6b1 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models The first to know: How token distributions reveal hidden knowledge in large vision-language models? InProc
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1e8403bd-a55d-4ef7-9dc1-b40ee193785a · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models A Survey of Large Language Models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 637246e7-529b-4d42-a374-0a774e2dfdf5 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Improved few-shot jailbreaking can circumvent aligned language models and their defenses
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7cb6da9d-22c6-4b49-9033-1ef6acbb4384 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Minigpt-4: Enhancing vision- language understanding with advanced large language models
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5039fe24-3486-46e2-b31e-07ae8f4a358b · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Safety fine-tuning at (almost) no cost: A baseline for vision large language models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4461d1a0-d860-431a-91e7-0ce83f9ad5b6 · outbound
Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b1d1d6c-e1c3-4414-b792-8952e4fd56f8 · inbound
Understanding and Mitigating Spurious Signal Amplification in Test-Time Reinforcement Learning for Math Reasoning Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ca574ac8-d05a-48c1-b5d0-5f3862ed884f · inbound
On the Vulnerability of Parameter-Level Defenses to Model Merging Test-Time Immunization: A Universal Defense Framework Against Jailbreaks for (Multimodal) Large Language Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.