Pith. sign in

Paper Citation Record · LEDGER

Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2501.05767.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.05767 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:52.218083Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T15:09:55.026823Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 85eec8d3-62e5-4e27-9f0f-024909f90687 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models

Reference 135

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.209391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:836c55bc574418797e247b86da2702d1bf929bc4f4e1d53ff674463b3f8c7e83

Observation 48ecc93c-134f-4f3b-bb96-9ec944c1ce57 · inbound

UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning cites this paper.

UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:52.218083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:52.218083Z digest=sha256:faf346f1798d53998781b844beaf475340083af270f17b50fe3b38f9c981f0c5

Observation 2274912a-62dd-498c-9c8a-40ea3cd393f8 · inbound

PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning cites this paper.

PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:13:34.937970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:13:34.937970Z digest=sha256:b027581af43d0c246b9630ebb2ef7c6686165f425bbc7f4c061df1802cbd22c7

Observation 9940576a-4182-421d-b421-9bfacf7d6405 · inbound

Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning cites this paper.

Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-19T06:52:08.087721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T06:50:02.607136Z digest=sha256:a8425e3552171c7280d5560c52ea5efd0581e2263df84305cc796edfab9acfab

Observation 63fd2520-a011-4839-b7eb-25fdabc0b8cb · inbound

RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension cites this paper.

RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T18:15:05.561502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:15:05.561502Z digest=sha256:fcc54daefd46c42adc4f42b875aee2ed977bef51d0cc0eec61b471669017ca0d

Observation 98d23dba-4d73-42f7-b1e8-54a3b0e3bf51 · inbound

Training Multi-Image Vision Agents via End2End Reinforcement Learning cites this paper.

Training Multi-Image Vision Agents via End2End Reinforcement Learning Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:01:24.328661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T00:59:28.618477Z digest=sha256:0977f6dd6d3dabc72d2ecced1078159c570328f85000e8d77ea214821e2a9316

Observation a417a851-0566-48ec-9742-9f2b06906b2f · inbound

CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding cites this paper.

CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:16:08.325270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T12:26:01.568507Z digest=sha256:91d397beaa84e87fd762d7316c38eeca18ea049f5a1fb715e93218e07c33a89f

Observation 2da49c6c-3e9a-4b2f-929b-abefa8092792 · inbound

Does Seeing More Mean Knowing More? Mono-Anchored Advantage Normalization for Multi-Source Visual Reasoning cites this paper.

Does Seeing More Mean Knowing More? Mono-Anchored Advantage Normalization for Multi-Source Visual Reasoning Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:00.616287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T22:53:55.407461Z digest=sha256:22b10a134c0bedebb9c45f561e33ac3fb5df62212066a20656fabbe1d3a6ad94

Observation ba7be77d-951f-473d-bd51-a39bc7e96394 · inbound

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models cites this paper.

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models

Reference 159

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:09:55.029033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T01:50:54.242508Z digest=sha256:a38662d11a4b2bf9c07ada946068ceeba9450a80ee3db1ef9663ce07b868c520

Observation 9aaaefa8-de82-4fd4-af0d-07d8ec5d2d9a · inbound

DiCoBench: Benchmarking Multi-Image Fine-Grained Perception via Differential and Commonality Visual Cues cites this paper.

DiCoBench: Benchmarking Multi-Image Fine-Grained Perception via Differential and Commonality Visual Cues Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:19:50.963105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T05:18:01.931929Z digest=sha256:500b900bd3586f93f99e1367f545925f2ee6e8df074b4b8cb751fdf704999704