Pith. sign in

Paper Citation Record · LEDGER

Matryoshka Multimodal Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2405.17430.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.17430 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:29:47.998290Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T11:28:04.133973Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0f3e4acc-c4f0-4034-9e3f-43b4e2628177 · inbound

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding cites this paper.

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding Matryoshka Multimodal Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:47.998290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:47.998290Z digest=sha256:19f8b023b0400a6a3f0cb85448693f40f421e3d95aa68586be8cb7164c7a6951

Observation 2f5c417c-18a9-4dc0-9b18-847fb8334728 · inbound

MagicVL-2B: Empowering Vision-Language Models on Mobile Devices with Lightweight Visual Encoders via Curriculum Learning cites this paper.

MagicVL-2B: Empowering Vision-Language Models on Mobile Devices with Lightweight Visual Encoders via Curriculum Learning Matryoshka Multimodal Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T05:36:16.252253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:36:16.252253Z digest=sha256:817405353828c847fe24b2b5d41b210144d152f53be064cbcc8b43734aa89a18

Observation 012a663d-0fad-46a8-b757-735387d93d37 · inbound

LLaVA-RE: Binary Image-Text Relevancy Evaluation with Multimodal Large Language Model cites this paper.

LLaVA-RE: Binary Image-Text Relevancy Evaluation with Multimodal Large Language Model Matryoshka Multimodal Models

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:55.784910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:55.784910Z digest=sha256:e72971bc18d18462f2a8b95c582501f709e7ac3ef3fb74c8478e7ae7da5b5eba

Observation c4b0a532-c0a0-469a-bf15-d1cc363d3756 · inbound

Elastic ViTs from Pretrained Models without Retraining cites this paper.

Elastic ViTs from Pretrained Models without Retraining Matryoshka Multimodal Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T09:03:03.616876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:03:03.616876Z digest=sha256:208b4bdab967919a6a5dcacbd3f1bf31a9d016127f868ee4a278433158676f7d

Observation 137d494d-9b77-4c58-8529-a2532223e982 · inbound

HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models cites this paper.

HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models Matryoshka Multimodal Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:01:04.856070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T17:20:19.653806Z digest=sha256:2a4c305fc4885d845f08270841c75642bbd4766635b2f7e69f1bf2dc6c7600fb

Observation e95de2c0-608d-4bf0-a353-9540adf2de9e · inbound

MIPIC: Matryoshka Representation Learning via Self-Distilled Intra-Relational and Progressive Information Chaining cites this paper.

MIPIC: Matryoshka Representation Learning via Self-Distilled Intra-Relational and Progressive Information Chaining Matryoshka Multimodal Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:12.550666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T03:35:44.435355Z digest=sha256:a0a533b2c0261804e5df43ede122ecaa1f65aed086ac9e6f07589b95b84fd418

Observation 37e8af78-8e33-4093-8fc4-bf7954f83427 · inbound

WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization cites this paper.

WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization Matryoshka Multimodal Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:36:08.098510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T16:06:26.450483Z digest=sha256:2132abbc28a3a2dce9db972509c78d53f891ebd4690dbcb1bf84b492565c3ca0

Observation 0949e3d4-0f25-4283-997e-f9e0daa42f83 · inbound

EvoGround: Self-Evolving Video Agents for Video Temporal Grounding cites this paper.

EvoGround: Self-Evolving Video Agents for Video Temporal Grounding Matryoshka Multimodal Models

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:32:52.455738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T19:29:47.356665Z digest=sha256:b2e32eaccabb7415f960f92a0d325a8c5e11ff2fe7b8be5df17c87f52ba3fbfb

Observation b0437ff3-6d98-4c11-9273-f5f072516862 · inbound

GraSP-VL: Length as a Semantic Granularity Interface for Vision-Language Representations cites this paper.

GraSP-VL: Length as a Semantic Granularity Interface for Vision-Language Representations Matryoshka Multimodal Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:52:48.226001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-19T21:50:27.380019Z digest=sha256:6cf53b49853c2365bbe830233979b77151ec92ebd170714cbf61aaadc5d74872

Observation a0dfdd1a-0b2b-4c64-a053-2d452deb8f67 · inbound

Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models cites this paper.

Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models Matryoshka Multimodal Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:28:04.135442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T09:35:24.118536Z digest=sha256:cf0d8eecb7d516cdb77082a6cf8c7f817aa29dd51ce5ea5a09ee2271dbbaedf9

Observation d659ef2f-90cb-45e5-99d1-923c05e3f84b · inbound

UltraViT: Latency-Optimized On-device Vision Encoder for Large Vision-Language Models cites this paper.

UltraViT: Latency-Optimized On-device Vision Encoder for Large Vision-Language Models Matryoshka Multimodal Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-31T00:01:33.328665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T00:01:33.328665Z digest=sha256:46eddd3e7c8bc256aa9f3219aeab35e0b6e512b03a58eed1769240ded773de1f