Pith. sign in

Paper Citation Record · LEDGER

MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2409.14818.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.14818 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:05:09.305960Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T23:01:20.549286Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ac52c24c-3ca5-416f-8f7b-1a5271ba8a31 · inbound

Safe + Safe = Unsafe? Exploring How Safe Images Can Be Exploited to Jailbreak Large Vision-Language Models cites this paper.

Safe + Safe = Unsafe? Exploring How Safe Images Can Be Exploited to Jailbreak Large Vision-Language Models MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T18:32:12.123346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:32:12.123346Z digest=sha256:da6da7aee02b8570e099fd565bcac46a85e7fd455b3264a468ad8da80f4206ad

Observation a3101dcc-14ae-417c-b28b-5c12cf8a2975 · inbound

TransBench: Breaking Barriers for Transferable Graphical User Interface Agents in Dynamic Digital Environments cites this paper.

TransBench: Breaking Barriers for Transferable Graphical User Interface Agents in Dynamic Digital Environments MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:17.299342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:46:17.299342Z digest=sha256:5846eb2bfd0259236d4bb1005886f648c90b80dd94b1190a241d36fc4006fbca

Observation ccabe1f6-81d6-4c7e-9296-bd3d434b7c5b · inbound

Da Yu: Towards USV-Based Image Captioning for Waterway Surveillance and Scene Understanding cites this paper.

Da Yu: Towards USV-Based Image Captioning for Waterway Surveillance and Scene Understanding MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T23:12:56.998779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:12:56.998779Z digest=sha256:04cd4fd71de9d0b9d6df8ccb4f69520e6369f55b160f57bf94e80d0ed622465d

Observation 54489867-d679-4ce6-863a-775806b8c29f · inbound

OS-MAP: How Far Can Computer-Using Agents Go in Breadth and Depth? cites this paper.

OS-MAP: How Far Can Computer-Using Agents Go in Breadth and Depth? MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T18:05:09.305960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:05:09.305960Z digest=sha256:6e74ea94105a915db50c0fef5c8b2f4b5d8e54a7ab5833130dc3b9cdba0cc896

Observation c663a255-1c6e-4b5c-9bae-bed7463e0359 · inbound

Agentic Services Computing cites this paper.

Agentic Services Computing MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding

Reference 189

Resolution
unresolved
no resolver link, observed 2026-08-04T14:41:50.948851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:41:50.948851Z digest=sha256:515bfed32479b1990b76d412c50059da917d181df2bb7ccbb08531ccd4304ca7

Observation bee5e058-a1c4-465d-b067-36467276915e · inbound

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents cites this paper.

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:01:20.551673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T22:59:16.413568Z digest=sha256:5635ad5e79f1fbda75bb4e2c872bfa7cd00d4e7d2ddf806babe5d566d3a216d7

Observation 158ccabb-312e-4251-9b5f-d3293f0093f3 · inbound

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents cites this paper.

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T16:38:36.380317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:38:36.380317Z digest=sha256:e4657b83e5402aa93deaf0e9b625fc35657beb5944a1ec6fe9a48ed43900d9b3

Observation 254986cb-92e9-4b3f-888f-f47901f887d2 · inbound

Firebolt-VL: Efficient Vision-Language Understanding with Cross-Modality Modulation cites this paper.

Firebolt-VL: Efficient Vision-Language Understanding with Cross-Modality Modulation MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:35:50.681570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T19:01:18.334371Z digest=sha256:98242c525d137aaafe1cb2183798f5cb6da6bc488e6a802575913902d41e68e6