Pith. sign in

Paper Citation Record · LEDGER

VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2403.20213.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.20213 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T20:54:45.697450Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T13:56:59.147746Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ed5819a0-0e38-4cf2-b49a-7cd5d38da620 · inbound

LHRS-Bot-Nova: Improved Multimodal Large Language Model for Remote Sensing Vision-Language Interpretation cites this paper.

LHRS-Bot-Nova: Improved Multimodal Large Language Model for Remote Sensing Vision-Language Interpretation VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T20:54:45.697450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:54:45.697450Z digest=sha256:026863ce698cc78a433de67c78ee986fed4a61e5874cc6172a00b256fda5ee43

Observation 1b0766c5-dfac-4e36-8aae-06d285a67953 · inbound

Large Vision-Language Models for Remote Sensing Visual Question Answering cites this paper.

Large Vision-Language Models for Remote Sensing Visual Question Answering VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T19:15:40.721915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:15:40.721915Z digest=sha256:650a7dfa28acf9f08f0f3470785034062a4495c8994e0706767da1a750b546c3

Observation bd64a8d8-7697-4ecf-95cd-9f43f1febd21 · inbound

GeoGround: A Unified Large Vision-Language Model for Remote Sensing Visual Grounding cites this paper.

GeoGround: A Unified Large Vision-Language Model for Remote Sensing Visual Grounding VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T19:27:57.531117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:27:57.531117Z digest=sha256:a8a7ed3d9e72bb6f21260bd8d9e91fe3baf18ad00904fbf231741d39f28ad951

Observation 5a950a6c-8ddb-4127-a94d-01f6292bd8c4 · inbound

RSUniVLM: A Unified Vision Language Model for Remote Sensing via Granularity-oriented Mixture of Experts cites this paper.

RSUniVLM: A Unified Vision Language Model for Remote Sensing via Granularity-oriented Mixture of Experts VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T20:32:31.701539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:32:31.701539Z digest=sha256:7d775cbe5e2448e7104ced9938747d41c705c977c5c2c925d4eaaa594ef4e5a1

Observation 466843bb-5fd1-4272-8138-5d817e886da3 · inbound

REO-VLM: Transforming VLM to Meet Regression Challenges in Earth Observation cites this paper.

REO-VLM: Transforming VLM to Meet Regression Challenges in Earth Observation VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T10:30:49.646578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:30:49.646578Z digest=sha256:6e059ec608c9aa7bb2da70c30c7bb27e2a64d6cdb2c3599c5004a23954c9b499

Observation 7ce4e8d6-01fa-4f8f-914d-5a5e49a61172 · inbound

Visual Large Language Models for Generalized and Specialized Applications cites this paper.

Visual Large Language Models for Generalized and Specialized Applications VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 277

Resolution
unresolved
no resolver link, observed 2026-08-10T22:08:09.973416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:08:09.973416Z digest=sha256:d2138cbf6bd99ae8df776c7ab7bc736f24f0e5b734fa62f99d9fa4d4579ec1f1

Observation ee450c8c-b169-474f-9419-a843b0de736a · inbound

GeoPixel: Pixel Grounding Large Multimodal Model in Remote Sensing cites this paper.

GeoPixel: Pixel Grounding Large Multimodal Model in Remote Sensing VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T15:32:12.360409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:32:12.360409Z digest=sha256:c72d98352454900f6822a194b0fc154d6ea4db4322ee83c0c3553a81386911c4

Observation 3f4aace0-37a0-4b07-b11b-ba9fbb559ba7 · inbound

VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs cites this paper.

VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:12.316595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:46:12.316595Z digest=sha256:31f2ba4f5d1b944c4c9f34c55dd3f3cebc4b2f5f1d0671498f6efc9cf19366c7

Observation f1be23c9-f869-48ca-a5ac-204dcf3a7b54 · inbound

GeoMag: A Vision-Language Model for Pixel-level Fine-Grained Remote Sensing Image Parsing cites this paper.

GeoMag: A Vision-Language Model for Pixel-level Fine-Grained Remote Sensing Image Parsing VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:13.264931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:13.264931Z digest=sha256:ebd36b3f692d0fae0b2d8d0743ebf8306b5f2fae008d4145e719f809c1b75812

Observation 898e6e55-730e-4c46-b48c-0116bab51854 · inbound

Annotation-Free Open-Vocabulary Segmentation for Remote-Sensing Images cites this paper.

Annotation-Free Open-Vocabulary Segmentation for Remote-Sensing Images VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T16:41:43.149382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:41:43.149382Z digest=sha256:16e03fe2388bd13083d6e98c8240424e51bcfd1e31c9ab030fe68c5b29e3d9d6

Observation 466528d3-ba0d-4c11-837a-69181a9ef3de · inbound

Semantic-Geometric Dual Compression: Training-Free Visual Token Reduction for Ultra-High-Resolution Remote Sensing Understanding cites this paper.

Semantic-Geometric Dual Compression: Training-Free Visual Token Reduction for Ultra-High-Resolution Remote Sensing Understanding VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:26:01.418886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T15:29:49.681399Z digest=sha256:ce6ea481cf0352b33732877fa93f4d29262c3218e5f4d7936c49687cf3025c29

Observation 59d02767-e132-4130-980e-c818e6987d9d · inbound

GeoSearcher: Anchor-Guided Progressive Reasoning for Remote Sensing Visual Grounding with Process Supervision cites this paper.

GeoSearcher: Anchor-Guided Progressive Reasoning for Remote Sensing Visual Grounding with Process Supervision VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:56:59.149262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-02T13:52:29.466227Z digest=sha256:784d5db9e1bfdf48185a321d32c812b72758bf231d4eecaefb3776598c86d8a5

Observation bfeebafc-3298-4522-b84b-cc55166125ef · inbound

Multimodal Large Language Models for Remote Sensing Image Understanding: Domain-Specific or General-Purpose? cites this paper.

Multimodal Large Language Models for Remote Sensing Image Understanding: Domain-Specific or General-Purpose? VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T10:20:56.785879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T10:20:56.785879Z digest=sha256:c017232bd01323f81d25f9360701a69a08ea1a6e0258e615a16a1b88e209ded9

Observation 6e171e44-4f04-46e4-97f5-1d0ae0d5def9 · inbound

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing cites this paper.

Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T00:59:20.770802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:59:20.770802Z digest=sha256:39c0f9073eb6896889be169da85c23511c8ed9be22539917f237118fe50260d2