Pith. sign in

Paper Citation Record · LEDGER

Safety of Multimodal Large Language Models on Images and Texts

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2402.00357.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.00357 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:38:17.103048Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T01:27:30.991837Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fd9d8eb5-393c-4e58-a585-61fa9825b70d · inbound

When Data Manipulation Meets Attack Goals: An In-depth Survey of Attacks for VLMs cites this paper.

When Data Manipulation Meets Attack Goals: An In-depth Survey of Attacks for VLMs Safety of Multimodal Large Language Models on Images and Texts

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T15:38:17.103048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:38:17.103048Z digest=sha256:c75ccec9cbac14215fa626cb7c3bb89b2119b41071d66bec6f63f0f3587c5aef

Observation f71db071-708a-45be-900d-a66b1762284c · inbound

A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations cites this paper.

A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations Safety of Multimodal Large Language Models on Images and Texts

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T19:45:18.435520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:45:18.435520Z digest=sha256:569b24c2f07841f2cafd1e0c889bdf277fa1746cfb801a5367dd5ac5aa756e8c

Observation cccdc95a-89aa-4648-9e47-aaf4bb2cc0e4 · inbound

Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment cites this paper.

Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment Safety of Multimodal Large Language Models on Images and Texts

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:53.646908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:35:53.646908Z digest=sha256:739a539f7456cad3c2dab4b7c59359c83c34e2dd28dcb061fcc76ec82021c4d4

Observation d2bbf271-6201-44be-a588-6223836e8b4e · inbound

The First Differentiable Transfer-Based Algorithm for Discrete MicroLED Repair cites this paper.

The First Differentiable Transfer-Based Algorithm for Discrete MicroLED Repair Safety of Multimodal Large Language Models on Images and Texts

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T22:21:02.805074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:21:02.805074Z digest=sha256:528d9dae56a16f8186734bc5367d303b6ffbe65b42827f9258149957e8a40add

Observation cbad9768-587a-4bb5-8ba6-3e50a50fa5e4 · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Safety of Multimodal Large Language Models on Images and Texts

Reference 146

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:46.157970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:46.157970Z digest=sha256:5f2b48467a1a42582b4ebb66e9613fc30ff0d4f8db7b69d0c8a006dc751fc357

Observation 3894a040-3054-4e18-a724-87ac4b6121f9 · inbound

Is GPT-4o mini Blinded by its Own Safety Filters? Exposing the Multimodal-to-Unimodal Bottleneck in Hate Speech Detection cites this paper.

Is GPT-4o mini Blinded by its Own Safety Filters? Exposing the Multimodal-to-Unimodal Bottleneck in Hate Speech Detection Safety of Multimodal Large Language Models on Images and Texts

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T16:30:29.384240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:30:29.384240Z digest=sha256:aee9d01cb5ec5ad45adfe3b5cf1f6fd2be3140512a84fcc0b2574f927a90b671

Observation 97a40a27-a51b-49e8-b125-c291234dd07f · inbound

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing cites this paper.

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing Safety of Multimodal Large Language Models on Images and Texts

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:27.469126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T04:50:08.866969Z digest=sha256:ba3d6f29b6aa2151cafefd6d8cb2027d9e85efa4cc41a64c6f190867e5fbed26

Observation a3431651-ae9c-4ad6-9d92-b0bb7f6ad56a · inbound

Investigating Adversarial Robustness of Multi-modal Large Language Models cites this paper.

Investigating Adversarial Robustness of Multi-modal Large Language Models Safety of Multimodal Large Language Models on Images and Texts

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:06:27.639652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T11:11:34.152223Z digest=sha256:6bca34ac566dd87d9ff1abc221fe87c09c820670b890f27b458fa3a5393d6858

Observation 339e86b5-919c-4603-ab5f-e6107eca8397 · inbound

Unveiling Privacy Risks in Multi-modal Large Language Models: Task-specific Vulnerabilities and Mitigation Challenges cites this paper.

Unveiling Privacy Risks in Multi-modal Large Language Models: Task-specific Vulnerabilities and Mitigation Challenges Safety of Multimodal Large Language Models on Images and Texts

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:27:30.993178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T16:33:28.848573Z digest=sha256:517e4284d95ae411fbe7f295e8c4eb6197d12c3aa2df55ede4053ebcedba49aa

Observation d54a7677-0c02-45c0-a740-7977675c926f · inbound

V-DEAL: Diagnosing Video Safety De-Calibration as an Understanding-Refusal Coupling Failure cites this paper.

V-DEAL: Diagnosing Video Safety De-Calibration as an Understanding-Refusal Coupling Failure Safety of Multimodal Large Language Models on Images and Texts

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T08:23:25.032757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:23:25.032757Z digest=sha256:5cf8bd3a3ab67875e3e5a9c07fafba49c01e900f5fcee74067e1d86621a57dfb