Pith. sign in

Paper Citation Record · LEDGER

Training-Free Reasoning and Reflection in MLLMs

As of 8 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2505.16151.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16151 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:13:58.155456Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact0
  • verified fuzzy30
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b4a03c8d-c0d0-4c32-b66d-099fb747db79 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Training-Free Reasoning and Reflection in MLLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:51.828346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:51.828346Z digest=sha256:374924b47c592dfb4d9c595986f92668635633a95c27ca2583a8b763c092bb1c

Observation 57845998-5e7e-4750-bcdc-a2068d408ced · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Training-Free Reasoning and Reflection in MLLMs Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:51.879241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:51.879241Z digest=sha256:dd32af8d7e42ada4d43960af633c42125172c6178244ced94e32bebe57ecc7ac

Observation 93e86eb5-d40a-4b50-8c8c-b65249a2b758 · outbound

This paper cites S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning.

Training-Free Reasoning and Reflection in MLLMs S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:51.984499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:51.984499Z digest=sha256:57e35365bef583beeb07c464e5eeca26af15e748b0ab516fb6b7a0aedf192b97

Observation 0c42477e-565d-45e7-b736-146e083932f4 · outbound

This paper cites Let’s verify step by step.

Training-Free Reasoning and Reflection in MLLMs Let’s verify step by step

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:04.355711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:52.146034Z digest=sha256:1e8960b2dbad80161ea61d4606373501e7bc37d860cc07fc28fd8c294d95c06f

Observation beb3ea43-93c7-4bfa-92dd-cf5433b58684 · outbound

This paper cites OpenAI o1.

Training-Free Reasoning and Reflection in MLLMs OpenAI o1

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:04.149806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:52.295443Z digest=sha256:1599cbbe1ba964a3f851c3c54b8fc4cf54013f64755be9120805400d930a8965

Observation 41caa1b0-7bc0-4854-b185-0b4f219ffe8b · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Training-Free Reasoning and Reflection in MLLMs LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:52.385935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:52.385935Z digest=sha256:1a1da381f7b1f8fc933686738ebeb9a12106aef652daa0b126560c07b54fc2dd

Observation 14cf2302-bd83-4130-928c-a89068c677ce · outbound

This paper cites Insight-V: Exploring long-chain visual reasoning with multimodal large language models.

Training-Free Reasoning and Reflection in MLLMs Insight-V: Exploring long-chain visual reasoning with multimodal large language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:03.936163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:52.502448Z digest=sha256:d382d1e56e035767bee0d4b2e16453f0a051862685788710e921c610ed18b170

Observation 9d634656-5c72-4ccb-9890-89369531f8f8 · outbound

This paper cites Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search.

Training-Free Reasoning and Reflection in MLLMs Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:52.634235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:52.634235Z digest=sha256:b0c9e73006ca46e9b51e0701df390cdc65dfb6e4652fb58e026e8014893b04f3

Observation c58e677e-b258-4034-85c2-3c79f7f93288 · outbound

This paper cites Visual-RFT: Visual Reinforcement Fine-Tuning.

Training-Free Reasoning and Reflection in MLLMs Visual-RFT: Visual Reinforcement Fine-Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:52.747958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:52.747958Z digest=sha256:7f28780e15935bd1a68893038981af514d499d6fbf2cdf7f6f67dad721ae3873

Observation c0044c76-f0d6-480e-bf76-6e1045f06c76 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

Training-Free Reasoning and Reflection in MLLMs Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:52.912688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:52.912688Z digest=sha256:a4fd25a9d1bdaebea78627790461b0432d9e1bc3d05e5e5e171fc75ba8ea0ab7

Observation c16d2579-a035-41ab-91e7-114625bcacae · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

Training-Free Reasoning and Reflection in MLLMs LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:53.084279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:53.084279Z digest=sha256:409a96abaaf69e433adb36cbe454736752c1893d8e409232184e75f2b7e4901a

Observation 0c3d472a-cddc-490d-aac9-d9e2ba56f329 · outbound

This paper cites R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization.

Training-Free Reasoning and Reflection in MLLMs R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:53.259417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:53.259417Z digest=sha256:f2bed751b9cc9c7233558f68995c04487169ac918132f1cf22539cedb58e065b

Observation b9901528-38be-4f40-ad0f-a6bc60af2194 · outbound

This paper cites Morcos, Hongseok Namkoong, Ali Farhadi, Yair Carmon, Simon Kornblith, and Ludwig Schmidt.

Training-Free Reasoning and Reflection in MLLMs Morcos, Hongseok Namkoong, Ali Farhadi, Yair Carmon, Simon Kornblith, and Ludwig Schmidt

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:03.795695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:53.411747Z digest=sha256:5bbd02c4a83568b0fe1814f39fb2e922297148ce4e3ff0af9eb584b8baec229f

Observation d55647b8-133f-4200-a45d-331eab01d5c3 · outbound

This paper cites Editing models with task arithmetic.

Training-Free Reasoning and Reflection in MLLMs Editing models with task arithmetic

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:03.631396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:53.532874Z digest=sha256:326840ee63ca0053f2d01b795a7f040034c016758e31045cb0dfc7c4ec22562d

Observation fd227eb9-e86f-44c5-9d0d-9148a56e3c05 · outbound

This paper cites Composing parameter-efficient modules with arithmetic operation.

Training-Free Reasoning and Reflection in MLLMs Composing parameter-efficient modules with arithmetic operation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:03.426559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:53.655470Z digest=sha256:bffeb29479394fcd77dc189a0eb32b36db6023b4828c3ecf5cbabdbcd4cde081

Observation 93053a11-4333-411e-9419-5bb4c56a9471 · outbound

This paper cites Gradual progression from sensory to task-related processing in cerebral cortex.

Training-Free Reasoning and Reflection in MLLMs Gradual progression from sensory to task-related processing in cerebral cortex

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:03.243842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:53.801760Z digest=sha256:56a1224bc6703e280251dfbc8342e93601586b6b994150db52cacf9af3f732f5

Observation db001b72-5745-4abe-abc6-1b1c1e4d1378 · outbound

This paper cites Hierarchical processing of visual and language information in the brain.

Training-Free Reasoning and Reflection in MLLMs Hierarchical processing of visual and language information in the brain

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:03.048780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:53.915991Z digest=sha256:5225f34fe4f55168393ddb09e2bd1bd69cd8ad00b69de8671c744c80dfb71f39

Observation 37983c83-55b1-4773-8a90-c17d4a67db5d · outbound

This paper cites Visionllm: Large language model is also an open-ended decoder for vision-centric tasks.

Training-Free Reasoning and Reflection in MLLMs Visionllm: Large language model is also an open-ended decoder for vision-centric tasks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:54.045942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:54.045942Z digest=sha256:eb7ea057cd4e0929b8b1a78f085703859a5a53041b46d589cce9a5469ebb799a

Observation ae20434e-d4d3-41ad-933e-ed70bf3a7f99 · outbound

This paper cites Qwen Technical Report.

Training-Free Reasoning and Reflection in MLLMs Qwen Technical Report

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:54.232145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:54.232145Z digest=sha256:f1ac9f7b258689f7e506534444fa06888ca2a5851a2b05838ec7f33fd8535e3a

Observation 9bd3e0e5-35f8-4239-8ce2-90440b2bab58 · outbound

This paper cites Visual instruction tuning.

Training-Free Reasoning and Reflection in MLLMs Visual instruction tuning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:02.852006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:54.390407Z digest=sha256:1cd2d27c8f3382f30b18cc8b140e2e3ee4dabb3b1d77eac914dc2fd898294449

Observation 6cde11fc-f8ee-44cb-91ca-8af5c0207f04 · outbound

This paper cites Emu: Generative pretraining in multimodality.

Training-Free Reasoning and Reflection in MLLMs Emu: Generative pretraining in multimodality

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:02.676886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:54.529176Z digest=sha256:27c8063a196190d5533dc5bfae20e49e8b570e1dbc54cd6e49fc4a85d8da648b

Observation 7ee25ad8-f013-4c6f-b450-7203ce445ce8 · outbound

This paper cites Chi, Quoc V.

Training-Free Reasoning and Reflection in MLLMs Chi, Quoc V

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:02.505631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:54.654112Z digest=sha256:3aca4fc4b953c250a64df483ab489e00866a92c8fe07aab04b4a33341ea1530c

Observation 7068f311-2234-4343-9571-58b2bc26b0ea · outbound

This paper cites Improve Vision Language Model Chain-of-thought Reasoning.

Training-Free Reasoning and Reflection in MLLMs Improve Vision Language Model Chain-of-thought Reasoning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:54.759630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:54.759630Z digest=sha256:33558a1bf79b635366f34631c2ec2752eaa2d5f8ff5cc2fe1c2ac996706f12a4

Observation cd65305a-d2d3-4566-9141-65de8ebf159f · outbound

This paper cites Compositional chain-of-thought prompting for large multimodal models.

Training-Free Reasoning and Reflection in MLLMs Compositional chain-of-thought prompting for large multimodal models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:02.272805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:54.882809Z digest=sha256:a154d7fae725065d1d0266c9f32fe8dbe2db0a10c0c6742c2440673d2a3aad3c

Observation 14d2e6a5-43c5-4b83-9c53-e5550f74b49f · outbound

This paper cites TextCoT: Zoom In for Enhanced Multimodal Text-Rich Image Understanding.

Training-Free Reasoning and Reflection in MLLMs TextCoT: Zoom In for Enhanced Multimodal Text-Rich Image Understanding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:55.011386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:55.011386Z digest=sha256:bf69fd844cd96e08e768efcf6d0cee54e73e767619fc719d2b0699f499bc4df9

Observation 6727c625-bba4-4568-a169-45ca1b8fe1a3 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

Training-Free Reasoning and Reflection in MLLMs MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:55.145082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:55.145082Z digest=sha256:73d7c01d70b59b76b6b069c40f451543498442f984af1a7381c755918f70dec1

Observation 897af4c1-7e20-4dd4-878f-c34a3423183d · outbound

This paper cites Merging models with fisher-weighted averaging.

Training-Free Reasoning and Reflection in MLLMs Merging models with fisher-weighted averaging

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:02.178644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:55.288092Z digest=sha256:dc7617d8051105a30a56740515298bddabfca05c67de07ccd6b3b7a4e6f79107

Observation b96f9645-0ec3-4ef9-ac84-ff3177a291a9 · outbound

This paper cites Raffel, and Mohit Bansal.

Training-Free Reasoning and Reflection in MLLMs Raffel, and Mohit Bansal

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:02.038488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:55.377913Z digest=sha256:aba20f797b2599a2f0d16057595ed1fc3daed2b77e7edd1a83dd8e426f89423e

Observation 198173b2-c49d-4ca9-8ece-5f69c3f9a79a · outbound

This paper cites Metagpt: Merging large language models using model exclusive task arithmetic.

Training-Free Reasoning and Reflection in MLLMs Metagpt: Merging large language models using model exclusive task arithmetic

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:01.860402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:55.488112Z digest=sha256:eaf47e4e11cbefc33dd5f54ee7ac460cb38cc9d68045bb25add51cba673675c4

Observation b0e58955-0b2b-44e3-99ae-a2e60fbe3d22 · outbound

This paper cites Dataless knowledge fusion by merging weights of language models.

Training-Free Reasoning and Reflection in MLLMs Dataless knowledge fusion by merging weights of language models

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:01.719617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:55.556366Z digest=sha256:3984e0fa1b3b5e735c129f382a8fedd921d57743c4016d03dd62c8d11306ca4a

Observation c95559bb-14b1-4004-9cb1-c43d4b8ecdc7 · outbound

This paper cites Language models are super mario: Absorbing abilities from homologous models as a free lunch.

Training-Free Reasoning and Reflection in MLLMs Language models are super mario: Absorbing abilities from homologous models as a free lunch

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:01.603696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:55.649661Z digest=sha256:484fa207eefdb2544b136d7c4f876fd28319cc774f3ef9f13498703069c385bd

Observation c9568ce4-e1e1-418e-97cd-37470082b171 · outbound

This paper cites Bring Reason to Vision: Understanding Perception and Reasoning through Model Merging.

Training-Free Reasoning and Reflection in MLLMs Bring Reason to Vision: Understanding Perception and Reasoning through Model Merging

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:55.819405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:55.819405Z digest=sha256:4abdb4d163c48c46815b2ccd5f739bd4b2c9356cee24489408e9cf18c60bd2d3

Observation babf7510-ac6f-4b55-b94a-f9be4d6483b0 · outbound

This paper cites Neu- ral Tangent Kernel: Convergence and generalization in10 neural networks.

Training-Free Reasoning and Reflection in MLLMs Neu- ral Tangent Kernel: Convergence and generalization in10 neural networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:01.423826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:55.972987Z digest=sha256:cc8484c98ef8052cf20166648d3bd53ce6b26800ae95b339286e6c38fcac66f4

Observation 5fdac022-bc8f-44fd-b5b0-c1eba68b3f94 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Training-Free Reasoning and Reflection in MLLMs Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:56.060872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:56.060872Z digest=sha256:3f1891731aad6a4c748a0f76618cfec9b30cc174f27ef46e3f2a4c20c4e8c640

Observation 383efb5e-2548-4a87-a036-59f905451354 · outbound

This paper cites AGIEval: A human-centric benchmark for evaluating foundation models.

Training-Free Reasoning and Reflection in MLLMs AGIEval: A human-centric benchmark for evaluating foundation models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:01.259648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:56.214550Z digest=sha256:7cc7f8b3dde56e6acc5b9687fdbe30b55506005c02815ee747992721804a87a1

Observation a0006f44-3ad8-4b4c-bafb-60e3314e2184 · outbound

This paper cites Improved baselines with visual instruction tuning.

Training-Free Reasoning and Reflection in MLLMs Improved baselines with visual instruction tuning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:01.109036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:56.305641Z digest=sha256:8b0cd0801e9d1c21dfc02a90388f4f39f41d85b52ca6b4fb6a21db80a2d02e82

Observation 80fa561b-4be0-40a4-bc90-2158c8d5e01e · outbound

This paper cites Llava- next.

Training-Free Reasoning and Reflection in MLLMs Llava- next

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:00.942818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:56.410192Z digest=sha256:8e1bdb4341c1871f59250facc3684275bf35746735eb2a3464cdbe34509bbca5

Observation 0a029e9f-43ef-4f43-a265-d67a00e0aef5 · outbound

This paper cites VILA: on pre-training for visual language models.

Training-Free Reasoning and Reflection in MLLMs VILA: on pre-training for visual language models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:00.804085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:56.536836Z digest=sha256:6886984f4f196bbac1fff6b02dcf39687091af7981929694c6873278fbeb4a10

Observation 71c96e96-d19e-42a1-9d43-e83aa32072a7 · outbound

This paper cites Building and better understanding vision- language models: insights and future directions.

Training-Free Reasoning and Reflection in MLLMs Building and better understanding vision- language models: insights and future directions

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:00.681087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:56.609209Z digest=sha256:3d137c3c99d2954092660c265fcec49b3ce08a96c8db4b09a574a5148a040dfb

Observation 79992ad2-8c98-474b-ba7a-df3a9f70a619 · outbound

This paper cites Sharegpt4v: Improving large multi-modal models with better captions.

Training-Free Reasoning and Reflection in MLLMs Sharegpt4v: Improving large multi-modal models with better captions

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:00.508551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:56.679480Z digest=sha256:fb4435e539537f9b12a2f35b2f450aac2d8194df7b926d12b33605eba5136d58

Observation a8dab4d0-57e2-4045-bbef-6ea7cf7dc34d · outbound

This paper cites NVILA: Efficient Frontier Visual Language Models.

Training-Free Reasoning and Reflection in MLLMs NVILA: Efficient Frontier Visual Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:56.772463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:56.772463Z digest=sha256:63a85f41c2563f4be9993067095132d34d8e1b5944e7a887d6002b9de4dfef68

Observation 69526aa6-a77e-4f0d-8ee7-a0e8566f7696 · outbound

This paper cites LLaV A-OneVision: Easy visual task transfer.

Training-Free Reasoning and Reflection in MLLMs LLaV A-OneVision: Easy visual task transfer

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:14:00.340628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:56.841558Z digest=sha256:aa1c0286f369f5295967c4ef15f7a5f57ec3a42ab331f0704c2ba37eabc468d3

Observation 52f83b67-50cc-4255-8a95-013e8a8d3ad3 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Training-Free Reasoning and Reflection in MLLMs Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:56.924965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:56.924965Z digest=sha256:88d2090e960c0062576bcc4dac6b87c44d9d5e9bfc853519313c2817ab06628c

Observation 2cdb1580-914e-48d7-af6b-f22e16ec8d97 · outbound

This paper cites an unresolved cited work.

Training-Free Reasoning and Reflection in MLLMs Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:14:00.145206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:57.024187Z digest=sha256:5305a8ef590c09175bd8ca293b190742a95f56553303f01d368e3c3280e6c416

Observation 294efcf3-6fe5-4ed8-bcf1-da716dec89ff · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Training-Free Reasoning and Reflection in MLLMs Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:57.151656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:57.151656Z digest=sha256:09a98deb94b6d6431fab5f736658b5d68b5bd8d794e8e5f1c7016a968d734b70

Observation ce5f00e2-6d81-4ddf-a05d-8b68d7b1f3bb · outbound

This paper cites MMMU: A massive multi-discipline multi- modal understanding and reasoning benchmark for expert AGI.

Training-Free Reasoning and Reflection in MLLMs MMMU: A massive multi-discipline multi- modal understanding and reasoning benchmark for expert AGI

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:13:59.979941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:57.319687Z digest=sha256:683c522f6a30c8334a9da9ba5f177bf4e74fee1584846034f70ba07bf0dd198a

Observation 7a260aa3-ee5a-4850-a16c-4fa8b2149eb9 · outbound

This paper cites MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark.

Training-Free Reasoning and Reflection in MLLMs MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:57.403852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:57.403852Z digest=sha256:ecf4568c24f71b66f12886cfdf8805b589c5b0ae3d42a541b48db86e7d0ee941

Observation 598d3a4f-c541-4a1e-85f4-24484d68a512 · outbound

This paper cites MathVista: Evaluating mathematical reasoning of foundation models in visual contexts.

Training-Free Reasoning and Reflection in MLLMs MathVista: Evaluating mathematical reasoning of foundation models in visual contexts

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:13:59.801231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:57.534671Z digest=sha256:15f6af5f7707507d4991d9988e8f34a483b842d5f8228fc988bddc40ff8253e4

Observation 1f77ec34-5c6e-427d-8fbf-dba8e89d4de6 · outbound

This paper cites Measuring multimodal mathematical reasoning with math- vision dataset.

Training-Free Reasoning and Reflection in MLLMs Measuring multimodal mathematical reasoning with math- vision dataset

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:13:59.680133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:57.670724Z digest=sha256:57f7ff1a2a8e65271cd9118a6d1684cb3fa90f0d8c150be313391ebbb45f6055

Observation 237520f4-695a-42dd-bf94-e5bc8b55970d · outbound

This paper cites We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?.

Training-Free Reasoning and Reflection in MLLMs We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:57.748180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:57.748180Z digest=sha256:9c2fffabc770969cefb52747caa67e3cf2a6eae7fa97f62b667ddc00254198fb

Observation c4dd5115-8240-4cc4-8aee-bf37bff505c7 · outbound

This paper cites an unresolved cited work.

Training-Free Reasoning and Reflection in MLLMs Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:13:59.569232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:57.824122Z digest=sha256:f342228d19cfe532f035d841d0c1b108380a85fd5a468fab2c1b8dca0604f065

Observation f9cbc9b4-afda-41ea-be0c-aa0ac8f62cd1 · outbound

This paper cites Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C.

Training-Free Reasoning and Reflection in MLLMs Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:13:59.427612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:57.957815Z digest=sha256:5fa785313703942bf9cb7bcb6f48733fa87ed925fda1aabb35d630e99651b42e

Observation a7941d7c-dc4a-4bf6-96fa-b588987efdcd · outbound

This paper cites Non-Reasoning MLLM.

Training-Free Reasoning and Reflection in MLLMs Non-Reasoning MLLM

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:13:59.285079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:58.054094Z digest=sha256:f76ae4d4fb723378688cf802936f7e1d7c6771d5f870e14863c5c857da4d2f6e

Observation 35c9be25-702f-4eb6-99ec-11fa196a66ac · outbound

This paper cites an unresolved cited work.

Training-Free Reasoning and Reflection in MLLMs Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:13:59.157793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:13:58.155456Z digest=sha256:f8399260ee7ad8701ad7dde4e4b0f9f0fa917ee879730b1cb31f327ab9ee5fbd

Pith citing papers

No inbound Pith citation observations are available.