Pith. sign in

Paper Citation Record · LEDGER

MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2407.08739.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.08739 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:55:56.665148Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T22:47:26.006183Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 24af4123-0fd2-44b7-92f7-832014c86653 · inbound

MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? cites this paper.

MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 70

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T01:29:30.190346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T01:29:30.032408Z digest=sha256:195622b53a86f4bc502ee98fcc006303df74c83e816fd5104ac65a6bca400662

Observation 4a21b82d-8f57-4c5d-8782-6593f43566e9 · inbound

LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models cites this paper.

LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:01:54.061147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T06:01:53.730356Z digest=sha256:2dba21015079cd006d0c8617e2593b9d75ea54741c0cafe842f2a7d55e55940c

Observation a2f6558d-bacf-40b3-9290-d6f5d0d6db7a · inbound

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization cites this paper.

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 115

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:16:17.557931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T09:16:17.150383Z digest=sha256:ef0e74cab2e276f175985efa737da503f6696d9fd288115c6a2046f4906f7d08

Observation 16f579f7-a6d3-4b0f-8645-548a8440880a · inbound

DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding cites this paper.

DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 110

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:09:26.732425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T10:09:21.542356Z digest=sha256:8fe320dbbb2a027f092e6384ad693dc8ea4c8726f29353cf3fe9611e37b2a5bb

Observation 7d4d353c-549d-424a-8824-bc950204ba0d · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 161

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.270300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:1a91571ccb154903b2369b3638f711fa421f6c5703fa5df9da3aa8c30c2b44d2

Observation c209adac-e703-4368-a0cb-5989f264fe0a · inbound

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems cites this paper.

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-22T22:57:13.394341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T22:55:34.238427Z digest=sha256:f5c47dc76de02e0dc96e2584b30b35e858fde292d5d48a55b48372583c8f408c

Observation 53cdd838-b3db-45ab-9ddc-eed5a0d2135e · inbound

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models cites this paper.

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 148

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:41:08.184254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T13:41:07.991012Z digest=sha256:08fe7146ea5fe34c0181757c64532e717687519cd92632cd680a427e129c9bc2

Observation b529bdf1-551a-4362-aa4e-a75d6d4924ab · inbound

Delving into RL for Image Generation with CoT: A Study on DPO vs. GRPO cites this paper.

Delving into RL for Image Generation with CoT: A Study on DPO vs. GRPO MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:56.665148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:56.665148Z digest=sha256:69310cc889ae960ba3d51ad85fa87955fe5ec5733e7a7dbc25889de6c0503e71

Observation 9bbde23e-4ac3-482c-8e41-cabfad96a542 · inbound

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start cites this paper.

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T13:12:58.908022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:12:58.908022Z digest=sha256:500e29526432ff1f747d173365b66157f6fe161b3e23708a8292b66c7b1ff119

Observation 5a67142b-43ef-49a4-a05e-f83c9a3c7073 · inbound

Self-Reflective Reinforcement Learning for Diffusion-based Image Reasoning Generation cites this paper.

Self-Reflective Reinforcement Learning for Diffusion-based Image Reasoning Generation MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T13:14:40.804899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:14:40.804899Z digest=sha256:2062d9164be868004bd6bc6a1c49af033fa3e0295ffd36ae5e20dc69826e27c0

Observation b142e3ae-f505-4f10-8d0c-6687061a00a4 · inbound

Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning cites this paper.

Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:34:12.657363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:34:12.657363Z digest=sha256:89c2e40c5fb040bb6462e5cfc4191d6dbc1197be98544ea31958836721d10000

Observation 8b0f0c64-a046-495b-afaf-bc5aacb038bb · inbound

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning cites this paper.

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:37:14.947419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T10:34:48.849524Z digest=sha256:fa20fe471fe4c9e5530191877670d33650becaa26bad07033684f34af198818a

Observation 7abe790c-2990-4cb5-8685-08c7b2b55b60 · inbound

Multilingual Multimodal Software Developer for Code Generation cites this paper.

Multilingual Multimodal Software Developer for Code Generation MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:58.402131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:58.402131Z digest=sha256:983b57bb38163b0082c45b72792d7084fdcf94c1ef9a186c31213a200ee10749

Observation ebeb8b8e-d759-4132-9dd0-3cd68ed37aac · inbound

SyncLoop: A Multimodal Dual-Loop Framework for Self-Improving Mathematical Reasoning cites this paper.

SyncLoop: A Multimodal Dual-Loop Framework for Self-Improving Mathematical Reasoning MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T15:12:02.048255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:12:02.048255Z digest=sha256:685d9b9d476f69e7ada52cd71d02e6b5faa8cd9516842718c12617f094c20cd9

Observation 1c842d2f-a386-4325-bfd3-8d7b4aa80f9d · inbound

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation cites this paper.

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:18.084915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:44:18.084915Z digest=sha256:00b97d2aa2a1a97d8587a4994eda95060a85c012a42feb30a808b152ddf3bdc0

Observation 9f00765b-351e-4bde-8836-8bd28789fc4a · inbound

NoisyGRPO: Incentivizing Multimodal CoT Reasoning via Noise Injection and Bayesian Estimation cites this paper.

NoisyGRPO: Incentivizing Multimodal CoT Reasoning via Noise Injection and Bayesian Estimation MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:40:53.011695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T04:39:58.296388Z digest=sha256:f6a49b0a33242e401f6aaee4eeb153a47c0f678cc98293aab3e59df14c1a60e5

Observation fc019f4c-96cb-4dad-b159-2c8c190cc679 · inbound

MICo-150K: A Comprehensive Dataset Advancing Multi-Image Composition cites this paper.

MICo-150K: A Comprehensive Dataset Advancing Multi-Image Composition MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:21:23.392627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T00:20:58.483350Z digest=sha256:e1ec69262860ed5724e9c91dd75a85b6c582a4262889ccdd7773e0fc54aa2e4b

Observation 638c9f27-1926-4981-a4f8-e3dc2dd532dd · inbound

Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning cites this paper.

Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T06:18:05.302009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T06:16:47.650748Z digest=sha256:126e9ac7371c906012babb561e3d282b286cbcdcd74f1a17b8a11da1cc182a26

Observation 913fe041-bf1c-4595-8295-43f8808509d4 · inbound

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery cites this paper.

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:47:26.007669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T18:39:44.696961Z digest=sha256:a2c2309bb435cb0025d07e12c0fb010edac95559de53f1cf5b4a3d3a73ced2b0

Observation 664e61c1-603b-4fdd-9fe9-9deef7b40e75 · inbound

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery cites this paper.

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-02T12:05:09.795947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:05:09.795947Z digest=sha256:d1c6ab2c0eebabc8eac6b2ec90d758e7e629f035aaafcfff0aeb0d43bc3a30e8

Observation 8ea9a2ae-b86a-4381-8109-45b8302c42c7 · inbound

FormalAnalyticGeo: A Neural-Symbolic Based Framework for Multimodal Analytic Geometry Problem Generation cites this paper.

FormalAnalyticGeo: A Neural-Symbolic Based Framework for Multimodal Analytic Geometry Problem Generation MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T06:15:36.702227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:15:36.702227Z digest=sha256:6a0f3faa4df5cac82ddb0ff4697d07f33f93cabf061e0ca54f98a0426c421d3d