Pith. sign in

Paper Citation Record · LEDGER

Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2403.06199.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.06199 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:11:47.724821Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T17:56:25.072862Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ece8428b-604a-409e-aed0-10c15b07b89e · inbound

HyperSeg: Towards Universal Visual Segmentation with Large Language Model cites this paper.

HyperSeg: Towards Universal Visual Segmentation with Large Language Model Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-12T12:03:54.256115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:03:54.256115Z digest=sha256:a84509ea7688fbfe0de9d38d0da3c1898f6f41b68750760924d3de0312dbc59f

Observation 6d31a72e-f5cf-467c-bda0-e859e0ea21ac · inbound

FlashSloth: Lightning Multimodal Large Language Models via Embedded Visual Compression cites this paper.

FlashSloth: Lightning Multimodal Large Language Models via Embedded Visual Compression Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T21:37:08.427822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:37:08.427822Z digest=sha256:21e8815352dfe9166e7799a382fd6f157bc63ef04a2e1bfbdc19f1b785021531

Observation 0c855a13-d7ce-447e-969e-9181ab7ba09d · inbound

LinVT: Empower Your Image-level Large Language Model to Understand Videos cites this paper.

LinVT: Empower Your Image-level Large Language Model to Understand Videos Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-11T20:54:16.675248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:54:16.675248Z digest=sha256:604c9227e98ec671fe0cae1753e6c920acf83ff48ef994c63bff6e13ccc05ab9

Observation ae40ddf9-7382-4fc5-9dce-e0a364bc2a19 · inbound

InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models cites this paper.

InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T12:37:59.687407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:37:59.687407Z digest=sha256:5accd9fb222655fe58221fab6e1c44f84db878c61dbc42f61320c45b5364047a

Observation 5b2ec4d8-7180-4bd7-bb9c-7b913907ced9 · inbound

LOP: Learning Optimal Pruning for Efficient On-Demand MLLMs Scaling cites this paper.

LOP: Learning Optimal Pruning for Efficient On-Demand MLLMs Scaling Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T20:11:47.724821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:11:47.724821Z digest=sha256:9a23a836cd01daf72c019f93b326fc55ab8becfbd3bcc8167433b6784aa6209c

Observation 80e71a35-495e-4fb4-979e-48df0d34a593 · inbound

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study cites this paper.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:48.525847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:48.525847Z digest=sha256:5920a0d9716d6f7d9ffcd6a91d2fafd6361e453979e0cb3ffa8124cdabbc1413

Observation 945a6bac-08f4-462f-a088-0a581fb5ec00 · inbound

Mema: Memory-Augmented Adapter for Enhanced Vision-Language Understanding cites this paper.

Mema: Memory-Augmented Adapter for Enhanced Vision-Language Understanding Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:56:25.076636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T17:53:29.860496Z digest=sha256:94fbbe7a2d58e72bd8bb74e0edc11f3039150b20a5b9670fcee9ce54994d2ce9

Observation a9ffd270-bbf9-4bb7-a8ec-bf4040c1881b · inbound

Structural Pruning of Large Vision Language Models: A Comprehensive Study on Pruning Dynamics, Recovery, and Data Efficiency cites this paper.

Structural Pruning of Large Vision Language Models: A Comprehensive Study on Pruning Dynamics, Recovery, and Data Efficiency Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:56:14.035868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-08T03:47:38.100037Z digest=sha256:3ab23702ea34fd9897fff8590af13550bc0ca6ae3c229536cd3784f0798577d0