Pith. sign in

Paper Citation Record · LEDGER

RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2406.12479.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.12479 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:38:46.888234Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:27:39.992089Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1c232cf3-5270-4340-81cf-8ad1dd03c2ae · inbound

Vision-Language Modeling Meets Remote Sensing: Models, Datasets and Perspectives cites this paper.

Vision-Language Modeling Meets Remote Sensing: Models, Datasets and Perspectives RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:38:46.888234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:38:46.888234Z digest=sha256:8b578c618ee77a884c3f03fdb076b3404cd1bc8f90fe28005567bf2034b96c95

Observation e1d15310-df64-4800-afd8-219bfce622f3 · inbound

Multimodal Mathematical Reasoning Embedded in Aerial Vehicle Imagery: Benchmarking, Analysis, and Exploration cites this paper.

Multimodal Mathematical Reasoning Embedded in Aerial Vehicle Imagery: Benchmarking, Analysis, and Exploration RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-04T18:17:17.851335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:17:17.851335Z digest=sha256:5202c644c5f1264bab66a091f944fd2d34a843557f016541f57807a6aff66192

Observation 9b2b62d2-1cd9-4ea9-9d7e-55e2b0cd535e · inbound

MMLANDMARKS: a Cross-View Instance-Level Benchmark for Geo-Spatial Understanding cites this paper.

MMLANDMARKS: a Cross-View Instance-Level Benchmark for Geo-Spatial Understanding RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:51:15.286707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T20:49:11.961079Z digest=sha256:26d553e296edf83bd9900df2b3964cd4e87d6932587129590e93c4a15d69492b

Observation b9240cff-dbb0-47e1-ab6e-3b89350f566b · inbound

Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap cites this paper.

Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 139

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:10:29.793391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T13:48:08.135538Z digest=sha256:31d1f0bebffbec7db8c7af05decd472ed603d53eb10b8bc666ff2715e6db5dff

Observation 8dd208b8-a452-4e44-8c97-cd91f28824fe · inbound

Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs cites this paper.

Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:15:56.904525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:52:51.777228Z digest=sha256:191522cea8c0a9c8e474484243882553e85c4a5230a3e9e566fc03fcad2088d7

Observation c3b7e2df-fb52-4088-8464-e9f1477270ee · inbound

Earth-OneVision: Extending Remote Sensing Multimodal Large Language Models to More Sensor Modalities and Tasks cites this paper.

Earth-OneVision: Extending Remote Sensing Multimodal Large Language Models to More Sensor Modalities and Tasks RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:27:39.993530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T13:16:12.768793Z digest=sha256:d22d8d7821b1f59e3af4f2453ddc21acdef57b34dfbd1b6d43201830127534af

Observation 47114dda-ffff-4cdb-9ab6-8f15fcc4680d · inbound

Multimodal Large Language Models for Remote Sensing Image Understanding: Domain-Specific or General-Purpose? cites this paper.

Multimodal Large Language Models for Remote Sensing Image Understanding: Domain-Specific or General-Purpose? RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T10:20:58.536227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T10:20:58.536227Z digest=sha256:4e601434370b96704e477876bc203a27cb302845dbc8f6f91ca0ab4c28e52d4f