Pith. sign in

Paper Citation Record · LEDGER

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R

As of 9 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 2 inbound Pith citation observations for arXiv:2507.08505.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08505 v2

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:20:53.640322Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T15:09:25.818449Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T15:10:16.528858Z

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 75fc37fc-9c21-4d22-829f-a0592d761a66 · outbound

This paper cites DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.520305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.520305Z digest=sha256:9c040be4bc3e9ce224f49d8bfbbeb3753dbc883007c9b4c0e34d64bc154a66ab

Observation 6f975698-f196-4643-939a-8f482f3ce123 · outbound

This paper cites Large Multimodal Agents: A Survey.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Large Multimodal Agents: A Survey

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.530368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.530368Z digest=sha256:97b81333f36ce83a3428403b1b000bb609ed58f6e9da458da173af4ed74ff16a

Observation 0b428182-fc7c-427b-a729-c682db3579a5 · outbound

This paper cites PowerInfer-2: Fast Large Language Model Inference on a Smartphone.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R PowerInfer-2: Fast Large Language Model Inference on a Smartphone

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.546644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.546644Z digest=sha256:e923ad76f2fdccfab38d5393d61b16db8eaef93de9c08844c5f713f837560ee6

Observation 1a78480c-e864-413e-b8bf-b0e5b0bb2613 · outbound

This paper cites Fast On-device LLM Inference with NPUs.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Fast On-device LLM Inference with NPUs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.553489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.553489Z digest=sha256:2a2469b8f89fc1d6d05c28806dbcbc9e6fb60ed4ef053e3c4eaa4b9de4a9ab0d

Observation c42b0a87-959e-45bf-aacd-6126a805075a · outbound

This paper cites SwapMoE: Serving Off-the-shelf MoE-based Large Language Models with Tunable Memory Budget.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R SwapMoE: Serving Off-the-shelf MoE-based Large Language Models with Tunable Memory Budget

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.562080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.562080Z digest=sha256:3f7308d4482a7f5163aaeb99e1ae04273e47f1be3584d296dd27e4d84bed6587

Observation aec358dd-da61-41c0-857c-663e8702bc61 · outbound

This paper cites llama.cpp: Efficient inference of llama models.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R llama.cpp: Efficient inference of llama models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:54.099816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T18:20:53.569560Z digest=sha256:b259d17d71b4d9a6fa6d50f54fcb10f79677bfafc26d4fe64b7e81d8e56d14ee

Observation 8bad181b-2ad6-464e-a59e-c7f9cce8d3bc · outbound

This paper cites an unresolved cited work.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:54.069315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T18:20:53.577311Z digest=sha256:add49faee01b000d05748f55a28bd7e30e6bf4e69b970d65eca57b8124eac490

Observation b54822bb-a506-4592-a8f7-908cc4f01434 · outbound

This paper cites mllm: On-device multimodal llm inference framework.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R mllm: On-device multimodal llm inference framework

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:54.027213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T18:20:53.583039Z digest=sha256:72e3d7ade5af0fdead65daaa297334197de750be30c3d38ad141db87f3880bcd

Observation 0f6c3968-16de-4113-812a-3430a7501f25 · outbound

This paper cites Visual Instruction Tuning.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Visual Instruction Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.589256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.589256Z digest=sha256:ad2a8d99735acca15f911a99bfe788df5a15a54494ed43aac335dfce0a666379

Observation 50ae9b4b-2809-41e9-8dd6-8d760fc22283 · outbound

This paper cites Mobilevlm: An efficient vision-language model for mobile devices.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Mobilevlm: An efficient vision-language model for mobile devices

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:54.000329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T18:20:53.594658Z digest=sha256:d7761147d7fed5461f43df0a7881a66d91777e078353ce3d1f84d8f12b395e3b

Observation 5371ee33-436e-4348-8600-c6c0b5f10cda · outbound

This paper cites Imp: Highly Capable Large Multimodal Models for Mobile Devices.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Imp: Highly Capable Large Multimodal Models for Mobile Devices

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.600924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.600924Z digest=sha256:7163fd6a0c7983bd422abfb5f07a83a3fc32d1c6d56df0d1f4f91e1ced7ee278

Observation 8cecdc0f-a379-4c9b-ab4d-abd5e6baebde · outbound

This paper cites Gptq: Accurate post-training quantization for generative pretrained transformers.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Gptq: Accurate post-training quantization for generative pretrained transformers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:53.974844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T18:20:53.607060Z digest=sha256:44f181b0634467642716eee05acc8a1f888297faab39c1452cd268122ccf2a36

Observation 9904750c-2caf-4e56-9628-c8946d734eb3 · outbound

This paper cites AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.614221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.614221Z digest=sha256:0f1b7c22c783e184e82fb6e987d8ebf7c0be2e404b9b9fe382ad9648c993e073

Observation 5427bdbc-b356-4fc1-9b1e-151c1803ddb2 · outbound

This paper cites Learning both weights and connections for efficient neural networks.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Learning both weights and connections for efficient neural networks

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:53.950668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T18:20:53.621874Z digest=sha256:3f24cfba734a230061f274d756eee0cc3841ebd4de8384250128d1f183e1660d

Observation 619e70b8-35e7-429d-b039-91618cb9a3cf · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Distilling the Knowledge in a Neural Network

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.627325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.627325Z digest=sha256:986572c683d942e8310f134ca4d4147de63b2cfdeb699ed2e6a51fe2bf6dd622

Observation e1d43d0f-ccd4-4afc-984e-af0d9b6b9c7c · outbound

This paper cites Fu, Stefano Ermon, Atri Rudra, and Christopher Ré.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Fu, Stefano Ermon, Atri Rudra, and Christopher Ré

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:53.920708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T18:20:53.633849Z digest=sha256:8fb479e1d3f10f9555ee8c0e579e693e66a18911f8a2d4bfff2986c3e343050b

Observation 0cf8702f-7a58-4cd4-9568-d9fee0740a68 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.640322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.640322Z digest=sha256:2ad960dae73dc15261e2c42a2ca65912287885333f74df0ff4269e1ab9aaeeca

Pith citing papers

Observation e49eea97-df3d-4f89-bd72-36e15061fbb6 · inbound

On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression cites this paper.

On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-21T15:10:16.531308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T15:09:25.818449Z digest=sha256:699bd4b90825824d74b76aaa518bb6ce9f5b236e5f33ee79aeeeaab6860268cc

Observation c821ae7c-d33c-4712-a469-b2b649ef2ddd · inbound

Progressive Semantic Communication for Efficient Edge-Cloud Vision-Language Models cites this paper.

Progressive Semantic Communication for Efficient Edge-Cloud Vision-Language Models Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T09:31:25.578099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T10:45:20.829708Z digest=sha256:3613a34f190a88754cedbb08e701284cdb192bfeda35312f10a6d264e80a2c8e