Pith. sign in

Paper Citation Record · LEDGER

Robust Multimodal Large Language Models Against Modality Conflict

As of 8 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 2 inbound Pith citation observations for arXiv:2507.07151.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07151 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:02:35.018852Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T01:00:28.684512Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:18:56.608098Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6a81bd36-c9fc-4228-8559-153a6cf28398 · outbound

This paper cites Insight Over Sight: Exploring the Vision-Knowledge Conflicts in Multimodal LLMs.

Robust Multimodal Large Language Models Against Modality Conflict Insight Over Sight: Exploring the Vision-Knowledge Conflicts in Multimodal LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.987671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.987671Z digest=sha256:d870427685a0b8bec2e161cd0e4299b3b0b15c4a928b8b7dca30669cbc5c6fb2

Observation c8fde335-0da2-4685-be08-d6a92b196ec7 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Robust Multimodal Large Language Models Against Modality Conflict Proximal Policy Optimization Algorithms

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.993758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.993758Z digest=sha256:be771dadf9e98bc04252a188e95b74b94a4d43b2215cc79057a9d44fceea3bb6

Observation 6c360178-5c32-449c-9339-58db71ec1181 · outbound

This paper cites Knowledge conflicts for LLMs: A survey.

Robust Multimodal Large Language Models Against Modality Conflict Knowledge conflicts for LLMs: A survey

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:02:35.375264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:02:35.009128Z digest=sha256:ff0d77c43f379185230111de8af6abc87121d5270e6a36fe9d8d026fb99944a8

Observation 824ec269-4595-4289-b25a-17bc9e75a33b · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

Robust Multimodal Large Language Models Against Modality Conflict Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:35.013941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:35.013941Z digest=sha256:eb430ac3c5638d3e7395a3d3e258c4319889989d368b4fba117592732a95dc70

Observation 63c38f93-dfe9-4510-9ac0-fee91f0bec99 · outbound

This paper cites Unraveling Cross-Modality Knowledge Conflicts in Large Vision-Language Models.

Robust Multimodal Large Language Models Against Modality Conflict Unraveling Cross-Modality Knowledge Conflicts in Large Vision-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:35.018852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:35.018852Z digest=sha256:f63b715034811d42d4bee24acd6b37533ff2973e89a8e1828e684a0bd00e5cb3

Observation 42c00eac-12e9-4bea-876e-32f40293ee15 · outbound

This paper cites Large vision-language model alignment and mis- alignment: A survey through the lens of explainability.

Robust Multimodal Large Language Models Against Modality Conflict Large vision-language model alignment and mis- alignment: A survey through the lens of explainability

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.998854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.998854Z digest=sha256:f1dacd7476cfc9b878e79ea9a8ec1e38072eda988b413bb93f606b65969b8bbe

Observation 66490eca-cb45-49fa-9753-2f6979fbfde8 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Robust Multimodal Large Language Models Against Modality Conflict Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:35.003935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:35.003935Z digest=sha256:3c18d445006ad4fadd5bd38ed541feb0d49e752df37f2dc877789d580f6cc10f

Observation eef14c0d-2e58-4ecd-a39a-9547bd6f71b4 · outbound

This paper cites REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization.

Robust Multimodal Large Language Models Against Modality Conflict REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.976580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.976580Z digest=sha256:b6f3444a9441b386d0a681bd830325504da1ec833bd0de51e605c38ae94dfb3c

Observation ea0551ff-bf15-4c4d-b400-1820d21f587d · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Robust Multimodal Large Language Models Against Modality Conflict Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.960675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.960675Z digest=sha256:4d2133af5412609adc27a4a77c4487ff99f97f213e336baeede56759124f4c65

Observation e4cf7d51-617f-4236-bf7c-c7701e49d89a · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Robust Multimodal Large Language Models Against Modality Conflict MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.965966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.965966Z digest=sha256:994f56b547224b4274fd429a54799b3224924a14ddaf5056415db64ff81eb872

Observation a6a9822f-a255-409c-aa8e-385e8ac1b09d · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Robust Multimodal Large Language Models Against Modality Conflict LoRA: Low-Rank Adaptation of Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.971396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.971396Z digest=sha256:54e20c343012673472f9378843d728fec75ef578b27a11d6aa73d706a355485a

Observation 42e835f8-6907-4cc8-ae47-3db44957f3d4 · outbound

This paper cites OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework.

Robust Multimodal Large Language Models Against Modality Conflict OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.981872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.981872Z digest=sha256:d77d3aaba65fc654a84f20806af50eec911668df5878a16e3716a5ea05f7516b

Pith citing papers

Observation 3268c7ed-89ce-4f77-801e-bbd66db5bb4c · inbound

Can Large Multimodal Models Actively Recognize Faulty Inputs? A Systematic Evaluation Framework of Their Input Scrutiny Ability cites this paper.

Can Large Multimodal Models Actively Recognize Faulty Inputs? A Systematic Evaluation Framework of Their Input Scrutiny Ability Robust Multimodal Large Language Models Against Modality Conflict

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T01:00:28.684512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T01:00:28.684512Z digest=sha256:a11120eb6e738b250973f6d907cc6fa36e447ead4910b99b3a1b25a97aae4a25

Observation 82844a20-d46b-4cca-b895-8d54fd7ae614 · inbound

MLLMs Get It Right, Then Get It Wrong: Tracing and Correcting Late-Layer Textual Bias cites this paper.

MLLMs Get It Right, Then Get It Wrong: Tracing and Correcting Late-Layer Textual Bias Robust Multimodal Large Language Models Against Modality Conflict

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:18:56.610055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T01:28:50.021432Z digest=sha256:984a775607a1087455d86e8d0dee7f52523cca80efdb387d2241fefa757801f1