Pith. sign in

Paper Citation Record · LEDGER

Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:2401.06805.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.06805 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 31 of 31 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T05:57:51.050490Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:59:44.777709Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0b2da07f-e806-4b7e-a552-f13f425ffe47 · inbound

ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection cites this paper.

ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:13:24.651777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-23T20:10:59.264484Z digest=sha256:ef879bbca1f86ea80f19ff273aba6c71074eaa5b74685bc7bfe764915f9aae25

Observation 8847d598-5b09-49ee-9377-b84a0658d1b3 · inbound

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning cites this paper.

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 198

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:32:32.853992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-23T04:30:38.804702Z digest=sha256:c745972b74220137fede68efcb745d519ed92ac6f0f33b60c54c33bee79fde51

Observation c7117d9e-44c6-4bd9-9c55-afd3332f733c · inbound

LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL cites this paper.

LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:15:46.312789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T15:15:46.255296Z digest=sha256:7f2280751e5ae505d69e673df758368a9eab2bc635ba2df7fb656e6c36627cf2

Observation e7c99b5f-2b23-4896-81d2-441100c7da1e · inbound

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems cites this paper.

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-22T22:57:13.410915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T22:55:34.238427Z digest=sha256:c6300ae6542c1e70c5787c3e5b2f9a9f3db667fdd03cf76da28da3fdeddbf03f

Observation 8e9ee500-0dcb-4a2f-906c-6f50032227b7 · inbound

Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning cites this paper.

Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:52:07.985526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T05:47:19.552825Z digest=sha256:e0174e6f199d9d14f98b476389baad87bbe6fd357483d42fbe2df63c2032b2f5

Observation b13c8292-0d2e-4eed-9c48-d58e99344d59 · inbound

PRISM: Programmatic Reasoning with Image Sequence Manipulation for LVLM Jailbreaking cites this paper.

PRISM: Programmatic Reasoning with Image Sequence Manipulation for LVLM Jailbreaking Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-19T03:37:01.119697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-19T03:36:24.013477Z digest=sha256:85ee86155956fffb8cd4e589531573b1e64070213131a6772eede3e141e2a6cf

Observation 5ab7640e-addd-4be7-b2f4-2910d7d78817 · inbound

Edge-Based Multimodal Sensor Data Fusion with Vision Language Models (VLMs) for Real-time Autonomous Vehicle Accident Avoidance cites this paper.

Edge-Based Multimodal Sensor Data Fusion with Vision Language Models (VLMs) for Real-time Autonomous Vehicle Accident Avoidance Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T05:57:51.050490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:57:51.050490Z digest=sha256:08ab99b8a71987e9353064179c5a7ce8721b582bb15ecc998dea9c053e6c76d5

Observation 15ef3c24-902a-4c0f-a59c-05fa565f8774 · inbound

The Emotional Baby Is Truly Deadly: Does your Multimodal Large Reasoning Model Have Emotional Flattery towards Humans? cites this paper.

The Emotional Baby Is Truly Deadly: Does your Multimodal Large Reasoning Model Have Emotional Flattery towards Humans? Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T00:59:41.320388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:59:41.320388Z digest=sha256:10a209f3979f5c0046dbf07e0c20c4c2f5a29dfe163631b4e5d85d2bd96c537d

Observation cc887b9b-36c5-4070-a7e2-26956cfdb73b · inbound

Large Language Models Show Signs of Alignment with Human Neurocognition During Abstract Reasoning cites this paper.

Large Language Models Show Signs of Alignment with Human Neurocognition During Abstract Reasoning Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-05T21:10:36.314261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:10:36.314261Z digest=sha256:ead620c6176962e2b43fcee05b57dbf5fca4e6634d162a04dd77aa46dcf14d0c

Observation 77663c75-23bc-40cf-ae45-4fea7f918f0c · inbound

Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning cites this paper.

Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 191

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:53.870257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:53.870257Z digest=sha256:94d0a8ba7153f224ae18894a2557f7acf22059439e17be9e4083e69bb48ce952

Observation fa318581-36f8-4249-ba9b-7d7ec939165d · inbound

SATORI: Static Test Oracle Generation for REST APIs cites this paper.

SATORI: Static Test Oracle Generation for REST APIs Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T17:26:48.552058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:26:48.552058Z digest=sha256:82a8b4e7769d76f3c981bb6df0f9e4156cacff0873abf8137da26d75f5201cf0

Observation ca171eef-4fd3-420a-8051-9f6507910840 · inbound

Leveraging Vision-Language Large Models for Interpretable Video Action Recognition with Semantic Tokenization cites this paper.

Leveraging Vision-Language Large Models for Interpretable Video Action Recognition with Semantic Tokenization Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T05:12:54.499516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:12:54.499516Z digest=sha256:562198d4dc257ca164cec5a7d9faa57bb588752947964ac3872a990538e5d129

Observation 5d91e788-d584-4756-a747-a78b5a31bf85 · inbound

SheetDesigner: MLLM-Powered Spreadsheet Layout Generation with Rule-Based and Vision-Based Reflection cites this paper.

SheetDesigner: MLLM-Powered Spreadsheet Layout Generation with Rule-Based and Vision-Based Reflection Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T22:14:15.550079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T22:14:15.550079Z digest=sha256:a9b52e18661619fa8df3a0a4d748d72f9b8a0f951afdccb744c777ab75047ef2

Observation e8d72ed2-dcf4-4090-b0f6-b37a42c9706b · inbound

AI Reasoning for Wireless Communications and Networking: A Survey and Perspectives cites this paper.

AI Reasoning for Wireless Communications and Networking: A Survey and Perspectives Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T19:34:25.404351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:34:25.404351Z digest=sha256:74925acc68481981ac5a134ab93d50981bd5ad4dc901a972445f92e88cc33bc1

Observation 6d39f313-f117-4f59-8bd8-a882444ee909 · inbound

MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs cites this paper.

MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-21T19:54:20.248281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T19:51:04.983299Z digest=sha256:ce630546eca0e57972b3b703847f24d311e0b0b132c94fbe3fd7e856aa187f02

Observation 91346b78-69ce-4cd5-9fa8-964916bcd6b3 · inbound

MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs cites this paper.

MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-03T21:44:02.813473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:44:02.813473Z digest=sha256:12275b2e71198f76e054d0e1a71220b86ddaf5ce54c600180f6224f2a7865a21

Observation 91be9c46-15bf-4390-aee6-8426011432d5 · inbound

Dual Tuning for Reasoning Efficacy-Driven Data Curation in Multimodal LLM Training cites this paper.

Dual Tuning for Reasoning Efficacy-Driven Data Curation in Multimodal LLM Training Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:12:35.420031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T08:12:25.063204Z digest=sha256:c36961ab5e1299ce0a2075e7c8419504070835072dce2fd0d9f9b887e6904137

Observation 22ce4050-fbb7-42bb-bc97-7b4ed160d06d · inbound

Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs cites this paper.

Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T07:21:00.732474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T17:12:33.231488Z digest=sha256:cdaa6956daefccdaf5f1aa984ae911c34ddcc8e3ac75a24d4b084191d8e14b6f

Observation 17be4f44-b564-4c49-b36b-7894eb917c0e · inbound

3D-VCD: Hallucination Mitigation in 3D-LLM Embodied Agents through Visual Contrastive Decoding cites this paper.

3D-VCD: Hallucination Mitigation in 3D-LLM Embodied Agents through Visual Contrastive Decoding Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:56:00.234454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T17:51:52.063238Z digest=sha256:28da35672e441d58164b12163dcc6d499e952ca69f7d93f44a39d37ae9840ab6

Observation 050fd947-b622-4700-91a6-a22fec9d2dde · inbound

Learning Preference-Based Objectives from Clinical Narratives for Dynamic Sepsis Treatment cites this paper.

Learning Preference-Based Objectives from Clinical Narratives for Dynamic Sepsis Treatment Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T22:22:06.385856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T22:22:06.385856Z digest=sha256:3df05bdd41dfcb32fd8f8958a40117052fa17de7abefa21d161f4ff54ae7da02

Observation be0cbbab-f994-47b4-adba-cbadba729324 · inbound

All in One: A Unified Synthetic Data Pipeline for Multimodal Video Understanding cites this paper.

All in One: A Unified Synthetic Data Pipeline for Multimodal Video Understanding Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 90

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:31:03.847437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:26:55.369840Z digest=sha256:3d028fd587866f93684035d9b3efc30cb3e3c47d67730c0eb84eb1934f0b3edb

Observation 57ead0d9-aa04-4116-951c-43c33b0e8f39 · inbound

Mol-Debate: Multi-Agent Debate Improves Structural Reasoning in Molecular Design cites this paper.

Mol-Debate: Multi-Agent Debate Improves Structural Reasoning in Molecular Design Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T00:49:49.012626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-10T00:42:32.617355Z digest=sha256:deb327cb45742560d341a02e320d9ddfd17f4009d1f6b8047d65349676b65fca

Observation de2faed8-6e47-418c-b11a-458575130d60 · inbound

Learn to Think: Improving Multimodal Reasoning through Vision-Aware Self-Improvement Training cites this paper.

Learn to Think: Improving Multimodal Reasoning through Vision-Aware Self-Improvement Training Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:22:23.494875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T06:17:57.264809Z digest=sha256:20507eb331ae34c00b3d58f5aa9696ef23f243468d1315d0cada6a2eca5866db

Observation d98ab346-77fb-4b13-9544-93c49b75b348 · inbound

GRIP-VLM: Group-Relative Importance Pruning for Efficient Vision-Language Models cites this paper.

GRIP-VLM: Group-Relative Importance Pruning for Efficient Vision-Language Models Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:17:50.811951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T19:15:13.205594Z digest=sha256:a9ba4cb4e0a7ab72a4bc1ac8d8360f0ae1994414e1bd433354c52e272ee29d4b

Observation 5f725fde-596d-44d7-aec4-537f102a3482 · inbound

Focus-then-Context: Subject-Centric Progressive Visual Token Reduction for Vision-Language Models cites this paper.

Focus-then-Context: Subject-Centric Progressive Visual Token Reduction for Vision-Language Models Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T05:23:58.507985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-21T05:20:55.448430Z digest=sha256:175c853ebdd619cbbda204513b2cd161554244cb89b9e42f4df825b4200d4f59

Observation fce0f2f0-498d-4079-955a-b7992d695fbd · inbound

Mags-RL: Wearing Multimodal LLMs a Magnifying Glass via Agentic Reinforcement Learning For Complex Scene Reasoning cites this paper.

Mags-RL: Wearing Multimodal LLMs a Magnifying Glass via Agentic Reinforcement Learning For Complex Scene Reasoning Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T13:43:29.006421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T13:38:01.819121Z digest=sha256:df16299dd5fdec91eb12ceb9b7b6cfdb0ba98a2ed1d771311b757758cba25c44

Observation 86fc9560-a05a-46b4-9af2-0a8545c80a21 · inbound

Investigating Adversarial Robustness of Multi-modal Large Language Models cites this paper.

Investigating Adversarial Robustness of Multi-modal Large Language Models Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:06:27.607029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T11:11:34.152223Z digest=sha256:d0223eaafc2052cb626f064e9673185ffb9ef73ba4b4575e1266cc22a9bbd27e

Observation 7e4be221-1973-4068-8fe2-80caf02295d6 · inbound

SPICE: Synergy and Partial Information Based Curriculum Evolution cites this paper.

SPICE: Synergy and Partial Information Based Curriculum Evolution Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T02:09:36.267326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:09:36.267326Z digest=sha256:b15d04d8990a20d2f1f377efd67418dbfc7017bc30638a19fe967bd22a0af38c

Observation 845c266d-a0c7-4114-b019-d98d8ccc6040 · inbound

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning cites this paper.

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 279

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:59:44.779117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-26T09:19:50.623741Z digest=sha256:3391354e4270753da95a385325f7aa22c718ab5bfc20a48653ddfe16012f5294

Observation 6002f94b-0bf0-40d4-84a0-a215daebff55 · inbound

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning cites this paper.

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 278

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T18:55:59.608772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-29T01:18:19.195007Z digest=sha256:b0c9145a72956d49b6c65fd9b656223dd77c4b25b81d9fbde45a123a92ff7db5

Observation 95753894-36ad-4324-96fe-abc79050a56b · inbound

HalluScope: Fine-grained Hallucination Diagnosis for Multimodal Large Language Models cites this paper.

HalluScope: Fine-grained Hallucination Diagnosis for Multimodal Large Language Models Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T08:33:10.377116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:33:10.377116Z digest=sha256:432ac8685604d8ffc07bb39b0d1c7a1bfc19f291592916a8b7de20cc08dbb116