Pith. sign in

Paper Citation Record · LEDGER

SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2410.18927.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.18927 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:55:40.252495Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:39:16.622704Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a822adb2-d59d-4194-a4a3-6113d2ad5311 · inbound

Visual Adversarial Attack on Vision-Language Models for Autonomous Driving cites this paper.

Visual Adversarial Attack on Vision-Language Models for Autonomous Driving SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-23T16:35:42.164443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-23T16:35:24.063578Z digest=sha256:ee8c537e82b6165e307248933441101010b28bcf6a24ac9156bf19ea95a3c7ce

Observation fbf6ca42-5a61-45e2-a609-d17d54993677 · inbound

A Survey of State of the Art Large Vision Language Models: Alignment, Benchmark, Evaluations and Challenges cites this paper.

A Survey of State of the Art Large Vision Language Models: Alignment, Benchmark, Evaluations and Challenges SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 251

Resolution
unresolved
no resolver link, observed 2026-08-10T22:17:23.899122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:17:23.899122Z digest=sha256:37886bd78b6d1f6a22602ed8b34a36c5f68000f3fd1d5f167944d5b9531959c5

Observation 842d4427-4229-4d9f-b061-38f3e911ef4c · inbound

Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation cites this paper.

Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T04:37:16.960430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:37:16.960430Z digest=sha256:1f73b0eaed775d3f306761436e0c115c2aa30cc2593cff884fb64fe98e713234

Observation 0d6687d1-48c9-41ec-b90f-141a42938795 · inbound

o3-mini vs DeepSeek-R1: Which One is Safer? cites this paper.

o3-mini vs DeepSeek-R1: Which One is Safer? SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T23:32:16.779178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T23:32:16.779178Z digest=sha256:9b52ed25097c370fc574cebf1b8fefd244954ded5e2cfdf882aa0217e2357031

Observation 70b85627-294c-4a45-b4a4-3a8fbafb761f · inbound

Universal Adversarial Attack on Aligned Multimodal LLMs cites this paper.

Universal Adversarial Attack on Aligned Multimodal LLMs SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T11:17:01.814133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T11:17:01.814133Z digest=sha256:643ed04f7a0c231412bec15e35cf4c6cc886e4ec492a5d93c8e027f1ecd77cba

Observation 3f75b51a-ad24-47cc-b318-4eb2434fc056 · inbound

A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations cites this paper.

A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 154

Resolution
unresolved
no resolver link, observed 2026-08-07T19:45:20.005573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:45:20.005573Z digest=sha256:4a6c45e9fd54e2b32288b552da394734354a4b741c121deeeb42f17175b9a39b

Observation 0ca77586-53ef-46aa-a01a-fc8aed8ce374 · inbound

Manipulating Multimodal Agents via Cross-Modal Prompt Injection cites this paper.

Manipulating Multimodal Agents via Cross-Modal Prompt Injection SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-16T11:55:40.252495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:55:40.252495Z digest=sha256:af4a3d6d01f6330a133c1d2a3c9af9ae9847d58ee721e8cef371d75e0112b9fd

Observation 57920c96-704a-4381-b407-b29ae37acafd · inbound

Toward Generalizable Evaluation in the LLM Era: A Survey Beyond Benchmarks cites this paper.

Toward Generalizable Evaluation in the LLM Era: A Survey Beyond Benchmarks SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 223

Resolution
unresolved
no resolver link, observed 2026-08-16T10:12:01.159768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:12:01.159768Z digest=sha256:795a3f804ae0436e00cb165cc554678410d1fbc183f47b43b06681b321c438ea

Observation a34517b3-2359-4a5b-9ea1-409067274079 · inbound

POISONCRAFT: Practical Poisoning of Retrieval-Augmented Generation for Large Language Models cites this paper.

POISONCRAFT: Practical Poisoning of Retrieval-Augmented Generation for Large Language Models SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T22:43:05.966295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:43:05.966295Z digest=sha256:f6a3636f24bb983f99a8d2202bc1d822e52b3ee6f7a96fd7dd08932c67bcdd14

Observation 6247ff6e-a531-4f58-83b3-40c1f9f92fe5 · inbound

MORALISE: A Structured Benchmark for Moral Alignment in Visual Language Models cites this paper.

MORALISE: A Structured Benchmark for Moral Alignment in Visual Language Models SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T20:12:19.955319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:12:19.955319Z digest=sha256:fbee06192005881beed1b1db72b19d9a2fd3cfd22189a076d42d513bedc19efe

Observation 346c0c83-1e0b-446e-8c4d-53d5c3850bdf · inbound

Hierarchical Safety Realignment: Lightweight Restoration of Safety in Pruned Large Vision-Language Models cites this paper.

Hierarchical Safety Realignment: Lightweight Restoration of Safety in Pruned Large Vision-Language Models SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:11:05.787596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:11:05.787596Z digest=sha256:2bc636467443905670cafa6ee2b997de68d27b2cb44793bf74c11d1e73676b24

Observation 96e76b2c-24e4-4577-8662-b9ffa4a0b0f0 · inbound

Three Minds, One Legend: Jailbreak Large Reasoning Model with Adaptive Stacked Ciphers cites this paper.

Three Minds, One Legend: Jailbreak Large Reasoning Model with Adaptive Stacked Ciphers SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:13.267384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:08:13.267384Z digest=sha256:759ac8276b45a530754bfeb56c768709e38ad23367ac2b68c085d4db2fe7f1d8

Observation ec99037e-e96b-4a23-a0f9-3ee01abca2a9 · inbound

MDIT-Bench: Evaluating the Dual-Implicit Toxicity in Large Multimodal Models cites this paper.

MDIT-Bench: Evaluating the Dual-Implicit Toxicity in Large Multimodal Models SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:06:33.579854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:06:33.579854Z digest=sha256:52fd4910e418e1ac87e45c50c78d4a27faf7d1e1ba4153755171141402ab4ba2

Observation 4c35aa6d-af23-4ccd-b6cd-39e70e21f35e · inbound

USB: A Comprehensive and Unified Safety Evaluation Benchmark for Multimodal Large Language Models cites this paper.

USB: A Comprehensive and Unified Safety Evaluation Benchmark for Multimodal Large Language Models SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:15:15.211995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:15:15.211995Z digest=sha256:41cc8c75b00073e7c18fe3a759736eb5e525719c18d93de0304eed6f58fc8fc6

Observation a3c88896-0a55-42ea-94af-7a6f60d66c26 · inbound

PRJ: Perception-Retrieval-Judgement for Generated Images cites this paper.

PRJ: Perception-Retrieval-Judgement for Generated Images SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:02:03.708785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:02:03.708785Z digest=sha256:0f0cf2187ef2edc1b89969248e3d56e2df40ec4450d5a69574510ecf831a40d0

Observation d5647fa7-95cb-4631-8574-7e64036abdc8 · inbound

Hatevolution: What Static Benchmarks Don't Tell Us cites this paper.

Hatevolution: What Static Benchmarks Don't Tell Us SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T01:06:13.360109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:06:13.360109Z digest=sha256:3f99cc65f9f15ffff5877abdccad57bd18ca4e852862026a8839b9d76a897255

Observation 3f02041a-4b63-4f7e-904c-fb8306e3faa0 · inbound

Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025 cites this paper.

Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025 SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:57:13.787678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:57:13.787678Z digest=sha256:5d9f88bcf51e9df7d6cdefa5618f14bed027c8fc4431da4bf3bf113a40ae2e67

Observation 179bf2b5-07c3-4775-a195-332fcea8df02 · inbound

PRISM: Programmatic Reasoning with Image Sequence Manipulation for LVLM Jailbreaking cites this paper.

PRISM: Programmatic Reasoning with Image Sequence Manipulation for LVLM Jailbreaking SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-19T03:37:01.050974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-19T03:36:24.013477Z digest=sha256:f0b322cd8814f745779097d62b2d0fb7b9ab99b2602d1c792fdc94fcb9114560

Observation 84a1fa09-5fac-49a0-9a7c-5d08cfffaf94 · inbound

Secure Tug-of-War (SecTOW): Iterative Defense-Attack Training with Reinforcement Learning for Multimodal Model Security cites this paper.

Secure Tug-of-War (SecTOW): Iterative Defense-Attack Training with Reinforcement Learning for Multimodal Model Security SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T12:09:39.812263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:09:39.812263Z digest=sha256:573d83ee0c7e568719d71658da28e5ae3364ba44e666abc5aafe4bef8fcdd663

Observation f9460823-9764-4395-b2d8-e576d10e5b52 · inbound

A Novel Evaluation Benchmark for Medical LLMs: Illuminating Safety and Effectiveness in Clinical Domains cites this paper.

A Novel Evaluation Benchmark for Medical LLMs: Illuminating Safety and Effectiveness in Clinical Domains SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T10:46:10.455850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:46:10.455850Z digest=sha256:7ef15a0fce243230d1f5398834f47d72fd42b9b4e6d7dbd0298f28c4c9f07a70

Observation 638c69ce-eb7c-43c9-be03-16cdedd8f489 · inbound

LLM Serving Optimization with Variable Prefill and Decode Lengths cites this paper.

LLM Serving Optimization with Variable Prefill and Decode Lengths SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 2

Resolution
malformed identifier
no resolver link, observed 2026-08-05T22:59:19.872813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:59:19.872813Z digest=sha256:31855d02c8e852944c974e845c2cdabd639ea1993243dcd26b8d15fcb9253817

Observation 6f9d33de-3907-4cc7-bea3-49308d3b109b · inbound

Mask-GCG: Are All Tokens in Adversarial Suffixes Necessary for Jailbreak Attacks? cites this paper.

Mask-GCG: Are All Tokens in Adversarial Suffixes Necessary for Jailbreak Attacks? SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T23:49:33.590413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T23:49:33.590413Z digest=sha256:9ab03a5596c9aac2ed5fd2750c6e5c8e2647ed87243887d9aa7f2d63806183f3

Observation d5b73532-6df5-4685-854b-badf749082d0 · inbound

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models cites this paper.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:40.772114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:40.772114Z digest=sha256:7d50a9e0fc53d6dff7d50c7a54770910b958330a99473323d04ec8de38916c55

Observation a715de99-dcb7-4650-b1d3-304988d2e1c3 · inbound

SecureWebArena: A Holistic Security Evaluation Benchmark for LVLM-based Web Agents cites this paper.

SecureWebArena: A Holistic Security Evaluation Benchmark for LVLM-based Web Agents SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:16:06.542069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T08:14:51.102085Z digest=sha256:addf16b3a1718b2b01f1537b7b442960f9219a0238a57a421dbbae4f85a7245b

Observation 6883b789-33df-4213-8142-e7f4b9befa9f · inbound

OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows cites this paper.

OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T15:46:06.039927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:46:06.039927Z digest=sha256:5176bd5e3af19070ff2c8b73413a1d498b51d83d38c66625533a348e5c390161

Observation 13fac892-3c89-4675-af0a-85eca087a48f · inbound

ROBOSHACKLES: A Safety Dataset for Human-Injury Prevention in Embodied Foundation Models cites this paper.

ROBOSHACKLES: A Safety Dataset for Human-Injury Prevention in Embodied Foundation Models SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T00:39:16.624438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-26T21:06:54.022715Z digest=sha256:fef706020898018488f4ef3e119dda8e05f552c3bbfa891c8dbd3d682faf2c2f