Pith. sign in

Paper Citation Record · LEDGER

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models

As of 16 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 1 inbound Pith citation observation for arXiv:2507.23382.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.23382 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:53:20.673277Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T00:36:22.176532Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T00:36:22.504927Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact1
  • verified fuzzy24
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 38f6022c-3e77-4d45-b575-8e1ee440fa5a · outbound

This paper cites Gpt-4 technical report.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Gpt-4 technical report

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.519968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.497524Z digest=sha256:613ff43cd6028433e072a9b7aa1c9407c76dde139dd068cb729c2b809be9897d

Observation dfc13025-eb1c-4a87-88b3-7fc3ede17fa9 · outbound

This paper cites Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.502152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.502152Z digest=sha256:a212a046ca891bc3d96356f7025dc7dae933adafc23143649808fe3a88763897

Observation 7d024f42-d35d-4b7d-92ee-16babd56faaf · outbound

This paper cites Ecm: A unified electronic circuit model for explaining the emergence of in-context learning and chain-of-thought in large language model.arXiv preprint arXiv:2502.03325, 2025.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Ecm: A unified electronic circuit model for explaining the emergence of in-context learning and chain-of-thought in large language model.arXiv preprint arXiv:2502.03325, 2025

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.506901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.506901Z digest=sha256:8a12886755174a8fc54cfa27458467a48e72528b5166881957f5a291ba9e4bc8

Observation 13f70436-3fe5-434d-90c7-e59b42e12ad9 · outbound

This paper cites Un- locking the capabilities of thought: A reasoning boundary framework to quantify and optimize chain-of-thought.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Un- locking the capabilities of thought: A reasoning boundary framework to quantify and optimize chain-of-thought

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.507106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.511551Z digest=sha256:d811db633586f0130d4d5594a3b48c743b7d53679418da8ec63453383df4b00b

Observation 7d3e216e-b486-4320-a1fe-67023c648907 · outbound

This paper cites M 3 cot: A novel benchmark for multi-domain multi-step multi-modal chain-of- thought.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models M 3 cot: A novel benchmark for multi-domain multi-step multi-modal chain-of- thought

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.493455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.515644Z digest=sha256:406074ff0a7a88c3f1f6c4438e86fe2229f2224727461ba86bf1318708093535

Observation 4a376083-56d5-406a-b71c-5c2c975fcac7 · outbound

This paper cites Egoplan-bench: Benchmarking egocentric embodied planning with multimodal large language models.CoRR, 2023.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Egoplan-bench: Benchmarking egocentric embodied planning with multimodal large language models.CoRR, 2023

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.479982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.519787Z digest=sha256:4e4f9017e498c514e47541c1981ce1182f352f1616694a1ad8997ded023afff3

Observation 03ca3a9d-8cff-4683-9d69-15435d998f64 · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.464734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.524370Z digest=sha256:b17caadff19f8096e5fb72f2ff860e08d1a39d413874bb6c05fe5c86af6d8b91

Observation 71c564ea-345e-4f30-a3cc-ae076c1446f7 · outbound

This paper cites Visual thoughts: A unified perspective of understanding multimodal chain-of-thought.arXiv preprint arXiv:2505.15510, 2025.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Visual thoughts: A unified perspective of understanding multimodal chain-of-thought.arXiv preprint arXiv:2505.15510, 2025

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.528393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.528393Z digest=sha256:5be618151a93d645478e99e93f3adb58a849c39de0275afdd309a3e87bb41cc5

Observation f1690525-f2cc-429f-a8e7-a80942b28811 · outbound

This paper cites Comt: A novel benchmark for chain of multi-modal thought on large vision-language models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Comt: A novel benchmark for chain of multi-modal thought on large vision-language models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.451184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.532223Z digest=sha256:58ce025e01aa863ff89e1031d4dd5f08dea289699eb8395ad235c42bc363e5e5

Observation c27c78a9-7cd2-4fe4-9320-b1a8af5c589a · outbound

This paper cites A Survey on In-context Learning.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models A Survey on In-context Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.535987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.535987Z digest=sha256:8b99ef87517ff01340172f69a727d1749992d0eb7dd34fd8f25bfef0db65e935

Observation 5d2493a2-9ab8-4f45-b2f6-a97da016bfb0 · outbound

This paper cites Vlmevalkit: An open- source toolkit for evaluating large multi-modality models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Vlmevalkit: An open- source toolkit for evaluating large multi-modality models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.438119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.540500Z digest=sha256:aad110adcd88286ce740e1af6842947cf8e2aad5d052dcadcbe2dbdf9d8c3702

Observation ef8c7bb5-9147-4a2e-8615-507acd7add9b · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.544678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.544678Z digest=sha256:fd531598354998cd69ec37c1180681f65eae1cd21fd801fc922c4231b5b757f8

Observation e9cfd071-65c8-4457-a939-78342128c09f · outbound

This paper cites Mllm-compbench: A compara- tive reasoning benchmark for multimodal llms.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Mllm-compbench: A compara- tive reasoning benchmark for multimodal llms

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.423840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.548904Z digest=sha256:95cd7712789a5e6693155f08d32d84fac05a67f0681fbff915ad27b0d0a1fbcd

Observation e13ced5d-1622-4e0f-827e-a450a91147f7 · outbound

This paper cites VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.552721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.552721Z digest=sha256:223e4ebb807be4d4e3d659dbe8235aa96510a8607f6909a7de8338b47e9fefbd

Observation 88aabee3-7b5a-49bd-a6d3-f42ca4afb14d · outbound

This paper cites Tree search for language model agents, 2024.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Tree search for language model agents, 2024

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.410865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.556992Z digest=sha256:b83e196b37353415931dc2bbace500b7ab583b4fc83bc02c39d73e4ec348ae6b

Observation b3647e32-7be7-471a-a0b0-f4b92f9b4980 · outbound

This paper cites Large language models are zero-shot reasoners, 2022.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Large language models are zero-shot reasoners, 2022

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.560926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.560926Z digest=sha256:9508934d84c8beee5a6c23083789a036983c2487dcdb24814a8375283cc459c3

Observation 07d3ce67-45a7-4c52-8f9e-b251c317ace6 · outbound

This paper cites Llava-onevision: Easy visual task transfer, 2024.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Llava-onevision: Easy visual task transfer, 2024

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.564766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.564766Z digest=sha256:6e55cece3317d3057a90855e8562142944dae2e5233b48f1e51e9f00cd97eddb

Observation f00ebf18-11f2-4a6c-95a1-5e526fe0e5a2 · outbound

This paper cites Seed-bench: Benchmarking multimodal large language models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Seed-bench: Benchmarking multimodal large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.379683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.568470Z digest=sha256:9bd9b46be968a79a15c72954d739b8c0021d194fae49e2efce8864f2f374e2de

Observation f2526b6e-39a8-41ed-b38b-492360bdee11 · outbound

This paper cites Behavior-1k: A benchmark for embodied ai with 1,000 everyday activities and realistic simulation.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Behavior-1k: A benchmark for embodied ai with 1,000 everyday activities and realistic simulation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.572100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.572100Z digest=sha256:cded9255575d8ffe1b9f93dfcaf6590d67b8fcab4b17148e6b3f4de673828dac

Observation 20607dfc-846e-443e-8f8b-8454ead93101 · outbound

This paper cites Ferret-ui 2: Mastering universal user interface understanding across platforms, 2024.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Ferret-ui 2: Mastering universal user interface understanding across platforms, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.357032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.576092Z digest=sha256:90388b59b8fe65bb539a5caf5a2adaaf713b14f50ab23998dd82a0a8b2108383

Observation 80832cb4-1839-4aab-a06d-2149fd8c6fc6 · outbound

This paper cites RoboMP$^2$: A Robotic Multimodal Perception-Planning Framework with Multimodal Large Language Models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models RoboMP$^2$: A Robotic Multimodal Perception-Planning Framework with Multimodal Large Language Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:53:20.857484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.579657Z digest=sha256:80e2a6a1b6872688cd177d81d2db070411b5be31310420e8806805e083907c96

Observation b2cb7bce-9da2-403c-a023-6a9b43a115c2 · outbound

This paper cites JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.583635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.583635Z digest=sha256:8f79a6b82df9630fe192167c643fcd2083475979b3645fae023eefa062978899

Observation 2d7c84d5-0e23-4ef3-90af-b17f03632481 · outbound

This paper cites m & m’s: A benchmark to evaluate tool-use for m ulti-step m ulti-modal tasks.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models m & m’s: A benchmark to evaluate tool-use for m ulti-step m ulti-modal tasks

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.342611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.587553Z digest=sha256:3d2b8ce7d81e831ef50b08dba6dd2fad228b6ee8888ea5535bbf59ee53ecdf22

Observation e18d92fa-a9c9-4f16-9239-96f48aad6669 · outbound

This paper cites Perception test: A diagnostic benchmark for multimodal video models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Perception test: A diagnostic benchmark for multimodal video models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.329216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.591223Z digest=sha256:3ac4ab72cb098ef65e2e3846c48058973e02dd1f2a15bd85bfa79f63a2e9631a

Observation 3b11e0bc-ef44-4a5c-b56b-fb42a421b34a · outbound

This paper cites What Factors Affect Multi-Modal In-Context Learning? An In-Depth Exploration.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models What Factors Affect Multi-Modal In-Context Learning? An In-Depth Exploration

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.594935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.594935Z digest=sha256:56f659fe2743b3e45d7aa1d4aa8bd4bbc778c17c821b4a03df6bcd65ffb3ac6a

Observation e1663673-4322-4823-b7ba-ef621e7f5a66 · outbound

This paper cites Mementos: System support for long-running computation on rfid-scale devices.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Mementos: System support for long-running computation on rfid-scale devices

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.315978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.598656Z digest=sha256:eb19b1adcaeb122e0585d72ee1573d4df8566a2cdcdea512545cb9999446fd1d

Observation fdb76c80-db26-41c4-824f-56969aaf932b · outbound

This paper cites Alfred: A benchmark for interpreting grounded instructions for everyday tasks.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Alfred: A benchmark for interpreting grounded instructions for everyday tasks

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.302254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.602259Z digest=sha256:3b567a3186eab8f7e4575f84e0435cd38d7ae38c307d20ed09d019b6511c7c34

Observation 4300e1d5-9a33-4489-a7f5-b3b29cf42ef6 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.606796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.606796Z digest=sha256:8cc05d4cf1a4588d6e941eeb9c707b0c3437b281100fe72702473763cc34a611

Observation d6e7aed5-07cf-4a4c-988a-d32486c2c349 · outbound

This paper cites Qvq: To see the world with wisdom, 2024.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Qvq: To see the world with wisdom, 2024

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.287495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.610677Z digest=sha256:76951c5317c2e2f34a003123c16f14968f12581b35916a86d4021e52e65cfc58

Observation 415ebad2-08aa-44e8-8832-4fc3bfd011b0 · outbound

This paper cites Planbench: An extensible benchmark for evaluating large language models on planning and reasoning about change, 2023.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Planbench: An extensible benchmark for evaluating large language models on planning and reasoning about change, 2023

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.614626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.614626Z digest=sha256:5da46acd296ee3cb94aecf16710e3657e9cf7ab0610d2d765420e68dddfc9979

Observation 175ce598-3c26-41fd-a595-99bdebe0ef8d · outbound

This paper cites Xlrs-bench: Could your multimodal llms understand extremely large ultra-high-resolution remote sensing imagery?, 2025.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Xlrs-bench: Could your multimodal llms understand extremely large ultra-high-resolution remote sensing imagery?, 2025

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.264906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.618401Z digest=sha256:0318122a76043757f9182d2e5a119ef79aca56952f632b5af0ead5b3790d1355

Observation 64d6f275-87ee-4f1c-944e-b2e104f60ad3 · outbound

This paper cites Plan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language Models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Plan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.622273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.622273Z digest=sha256:d52d9dfa3850997b90848415c3cf888bed605814c4e301c272ad1d33a17b745b

Observation a43f15a6-d6dd-4aec-9a8f-94076181df54 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.626682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.626682Z digest=sha256:32f8e23af885174daf8d870941d53c64ec6c8281ab2a8c703c47e6752a2cc94f

Observation e25a1a96-ca37-46a9-a969-d45c7963374c · outbound

This paper cites S3 agent: Unlocking the power of vllm for zero-shot multi-modal sarcasm detection.ACM Transactions on Multimedia Computing, Communications and Applications, 2024.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models S3 agent: Unlocking the power of vllm for zero-shot multi-modal sarcasm detection.ACM Transactions on Multimedia Computing, Communications and Applications, 2024

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.251436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.630864Z digest=sha256:ba6f1a258e6dfc31664554c0917f3e1eed7e1347910767f609a69811d0e8638f

Observation 31d4ee66-0bea-495b-95af-e3346d405260 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Chain-of-thought prompting elicits reasoning in large language models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.237662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.634494Z digest=sha256:3e529d39aaa50afc693a207305f2a44e76e913f7323aa3cd3d086b2dfc77988f

Observation e098dcdc-292b-4deb-9cea-d21e07d50b92 · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.638489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.638489Z digest=sha256:1218ee7bee5038e8fa1095c5f772bb350efb396ff9a3d84c6892fb950d52aed5

Observation 8dad09c2-d8f5-4d26-a6bf-c430b5b79682 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.642627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.642627Z digest=sha256:3a8cfcb5715ecedc2901b1af172de06044a2299e2eb82c0f56222383829092e0

Observation 40164780-a8fd-448a-b20b-6e7c10104746 · outbound

This paper cites LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.646437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.646437Z digest=sha256:17dd3a6765a1f9839ea53d9dbb07f919d6dc069bb084c20ca8cef43143f9d98f

Observation 9fbe474f-7f41-4e34-8c20-c37f16660a2e · outbound

This paper cites Osworld: Benchmarking multimodal agents for open-ended tasks in real computer envi- ronments.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Osworld: Benchmarking multimodal agents for open-ended tasks in real computer envi- ronments

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.222379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.650321Z digest=sha256:b727dfa07e2e6f11bfcaa569c0e46238f9e581ee6b63303761cfae8fba1543c9

Observation 85312a74-d55f-422f-afab-0aded02c8247 · outbound

This paper cites Mm-react: Prompt- ing chatgpt for multimodal reasoning and action, 2023.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Mm-react: Prompt- ing chatgpt for multimodal reasoning and action, 2023

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.208589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.653970Z digest=sha256:bc873074398f05d3349473bdd3e40478ef25a57003e88a0be01dfee52e22677d

Observation 26a3a3dc-d9ba-4060-ba7c-8bd3e211a70c · outbound

This paper cites MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.657760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.657760Z digest=sha256:7fc0233e45e75198b0ac66a797a7f2535b225970edefbe755c6ae1691996a5a0

Observation bb719f0c-e043-4363-9b71-a54dc0b829d1 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.194795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.661802Z digest=sha256:2b9f78104a50d61ba4c6bd76a639ae495dd3b29a946105c3527aa2b25f627163

Observation df3b235e-10dd-4c5d-a6ab-30a1188db3d8 · outbound

This paper cites Open3dvqa: A benchmark for comprehensive spatial reasoning with multimodal large language model in open space, 2025.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Open3dvqa: A benchmark for comprehensive spatial reasoning with multimodal large language model in open space, 2025

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.180213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.665367Z digest=sha256:eb8cc4346bd260c4ddfa36df589c6162579e4800bff4d3cf213529d0dcc0ce5e

Observation 54c936e4-ef99-4a58-a577-45b9aa36009a · outbound

This paper cites Le, Ed H.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Le, Ed H

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.167018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T10:53:20.669515Z digest=sha256:f9455722163faf266d61fb5e3c04b06b4039a08f85295d878a3974214661be17

Observation bb16c00f-dbbe-4255-910c-68e64810b32f · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.673277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.673277Z digest=sha256:bf9959c88a245ea7ffdd3a216e350598ec63a17fdce4f6b8d5dd2e7512849a19

Pith citing papers

Observation 801da51e-d348-486c-acc6-532317dd6cee · inbound

ComplexityWorld: Benchmarking Vision-Language Models on Verifiable Visual Decision Making cites this paper.

ComplexityWorld: Benchmarking Vision-Language Models on Verifiable Visual Decision Making MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models

Reference 16

Resolution
metadata mismatch
local_arxiv, observed 2026-08-11T00:36:22.517208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-11T00:36:22.176532Z digest=sha256:05c15284e55234e93457bae006893c2590565d52062c225e870156e1069c9195