Pith. sign in

Paper Citation Record · LEDGER

RSGPT: A Remote Sensing Vision Language Model and Benchmark

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2307.15266.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.15266 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T20:54:45.651757Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T17:05:51.453799Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a9774f71-66ff-4897-97db-3c0f802a7730 · inbound

LHRS-Bot-Nova: Improved Multimodal Large Language Model for Remote Sensing Vision-Language Interpretation cites this paper.

LHRS-Bot-Nova: Improved Multimodal Large Language Model for Remote Sensing Vision-Language Interpretation RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T20:54:45.651757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:54:45.651757Z digest=sha256:119ac70b418d7f5f9d7d39ebaff6901e90140edaf99f505b6536e6fb8e084c7f

Observation 8f5e3d6f-4ffe-41f9-8388-d5f0bf429695 · inbound

Large Vision-Language Models for Remote Sensing Visual Question Answering cites this paper.

Large Vision-Language Models for Remote Sensing Visual Question Answering RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T19:15:40.732785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:15:40.732785Z digest=sha256:9d1350043e1c6ff757287b1dbb4d6eb6d4abcfe3527025a40cf294cc57074414

Observation ba962f70-11da-476b-bbda-9c765070f5a4 · inbound

MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs cites this paper.

MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 138

Resolution
unresolved
no resolver link, observed 2026-08-12T14:31:37.266735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:31:37.266735Z digest=sha256:844618f6f381e9fbd79faf30788d70d4561c5003c8ec0ce7cede93558494a20b

Observation 2e5381a7-c03e-4aa2-8af6-376c51d4163c · inbound

GEOBench-VLM: Benchmarking Vision-Language Models for Geospatial Tasks cites this paper.

GEOBench-VLM: Benchmarking Vision-Language Models for Geospatial Tasks RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T10:21:22.491536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:21:22.491536Z digest=sha256:390e467e068e4c6c70c839f2ad5e8bcede81d6837234f3476271c035820d0dfa

Observation 1632509a-76ec-42c0-b2ad-5f63ca53df58 · inbound

RSUniVLM: A Unified Vision Language Model for Remote Sensing via Granularity-oriented Mixture of Experts cites this paper.

RSUniVLM: A Unified Vision Language Model for Remote Sensing via Granularity-oriented Mixture of Experts RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T20:32:31.516181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:32:31.516181Z digest=sha256:e260f2fe28cab9e282eaf129ae420f41709c5ed6a50c3d839334aa6893537d0b

Observation 0f3179b1-9005-4bdd-8ab2-f8ec1cf0bc58 · inbound

EarthDial: Turning Multi-sensory Earth Observations to Interactive Dialogues cites this paper.

EarthDial: Turning Multi-sensory Earth Observations to Interactive Dialogues RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T11:36:35.263849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:36:35.263849Z digest=sha256:da328d117c06b546f99a8d8c046ffe3591ac99af473f1e17c78a337de6e7577d

Observation e1fa586c-e91f-4078-82c5-69d5a2d4deed · inbound

REO-VLM: Transforming VLM to Meet Regression Challenges in Earth Observation cites this paper.

REO-VLM: Transforming VLM to Meet Regression Challenges in Earth Observation RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T10:30:49.600576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:30:49.600576Z digest=sha256:20b2a788ee4738322f028720d0f6190a82e35bf161d8155f8e0b0c03c8d8e73d

Observation 0221ae97-370c-4d17-9822-1dfbd82cf01e · inbound

UniRS: Unifying Multi-temporal Remote Sensing Tasks through Vision Language Models cites this paper.

UniRS: Unifying Multi-temporal Remote Sensing Tasks through Vision Language Models RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T23:18:21.294085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:18:21.294085Z digest=sha256:b0e10ebdf37e7f1b6fac4d8f4d2ca2043b5323a019e2f76c174a449af9c91c18

Observation 73413d2c-6b90-4240-b15c-a994ab71d390 · inbound

Visual Large Language Models for Generalized and Specialized Applications cites this paper.

Visual Large Language Models for Generalized and Specialized Applications RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 124

Resolution
unresolved
no resolver link, observed 2026-08-10T22:08:09.363894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:08:09.363894Z digest=sha256:0e8d5ce21cf883327f9955abdf5a4083253dc9eaad9558bd904b8c1e29204c9d

Observation b57f24db-3cc5-4524-b60a-e8a3273e966d · inbound

GeoPix: Multi-Modal Large Language Model for Pixel-level Image Understanding in Remote Sensing cites this paper.

GeoPix: Multi-Modal Large Language Model for Pixel-level Image Understanding in Remote Sensing RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T20:53:47.960477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:53:47.960477Z digest=sha256:b20455aff193213250b0be6e5f7be64c0ee03ff249ccf5b5ca2eab55b38a6cd9

Observation 76172d4a-8c94-4492-81bd-414cd2f905d0 · inbound

GeoPixel: Pixel Grounding Large Multimodal Model in Remote Sensing cites this paper.

GeoPixel: Pixel Grounding Large Multimodal Model in Remote Sensing RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T15:32:12.314531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:32:12.314531Z digest=sha256:56226405e0c6315912aea4f4daaaf1e00d966bb508c2caeedd99cd97582070b7

Observation 230bbcc1-05db-493b-b1c1-5f64b364b39c · inbound

Multi-Agent Geospatial Copilots for Remote Sensing Workflows cites this paper.

Multi-Agent Geospatial Copilots for Remote Sensing Workflows RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T13:41:49.043349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:41:49.043349Z digest=sha256:08ae6eeb3ccaa7073bb71c2155ed0ea41065b76edb7dfd9a4bbd58ec4005ae6d

Observation 52a68107-ddfe-483d-9ca3-3aee66aa6a90 · inbound

SARChat-Bench-2M: A Multi-Task Vision-Language Benchmark for SAR Image Interpretation cites this paper.

SARChat-Bench-2M: A Multi-Task Vision-Language Benchmark for SAR Image Interpretation RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T10:14:30.133950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:14:30.133950Z digest=sha256:a8f2a40c9565e9a002b9d299f75fb5e123d2e887c5f964e1092eeda2589983b3

Observation 4ea8d0c7-8626-4ed6-97ff-c88b562bda53 · inbound

Chain-of-Talkers (CoTalk): Fast Human Annotation of Dense Image Captions cites this paper.

Chain-of-Talkers (CoTalk): Fast Human Annotation of Dense Image Captions RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:54.615881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:54.615881Z digest=sha256:ed8df319dae654424e08fd8bbe496b10c55142cf0e9888a1e566598b29a710d6

Observation 89386da8-4e18-4d7c-801e-a4d5714a0262 · inbound

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding cites this paper.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.928081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.928081Z digest=sha256:e51c6296218ff965efbc0e8bd33a0e9760ae8ac1fd68fbf1a78344c6d6cf351b

Observation 3edb0cda-0e2e-45c5-b5c1-bc48b5709486 · inbound

GeoProg3D: Compositional Visual Reasoning for City-Scale 3D Language Fields cites this paper.

GeoProg3D: Compositional Visual Reasoning for City-Scale 3D Language Fields RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:49:43.574175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:49:43.574175Z digest=sha256:ee36669233a6db386b11deca435c913b305ba90cfc43bd4d5eff4cd91a8adf2e

Observation d91cd1ef-41d9-4f02-9fbf-215499192c50 · inbound

A Satellite-Ground Synergistic Large Vision-Language Model System for Earth Observation cites this paper.

A Satellite-Ground Synergistic Large Vision-Language Model System for Earth Observation RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T19:24:45.014869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:24:45.014869Z digest=sha256:af8b3de6945966e0ff2eda29579f1d48835bb7bed8c05a076c89bba70bdc033d

Observation 4e9ac6fd-9126-4258-8354-b240f88382ba · inbound

GeoMag: A Vision-Language Model for Pixel-level Fine-Grained Remote Sensing Image Parsing cites this paper.

GeoMag: A Vision-Language Model for Pixel-level Fine-Grained Remote Sensing Image Parsing RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:13.189322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:13.189322Z digest=sha256:89f606758c1ab1f3b3ac9dc74cdf5f6b3b3630933f7288dba297641ecabe1706

Observation 88f2835b-5272-4196-83ea-6c24caab3b65 · inbound

Enhancing Remote Sensing Vision-Language Models Through MLLM and LLM-Based High-Quality Image-Text Dataset Generation cites this paper.

Enhancing Remote Sensing Vision-Language Models Through MLLM and LLM-Based High-Quality Image-Text Dataset Generation RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:07:32.895752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:07:32.895752Z digest=sha256:aef2ff6cd353fafbc46dc219fc374a745ba2469065ad7810f05a1452c8958b6c

Observation 4fc331a4-3056-4123-b880-b41c7de012e1 · inbound

Few-Shot Vision-Language Reasoning for Satellite Imagery via Verifiable Rewards cites this paper.

Few-Shot Vision-Language Reasoning for Satellite Imagery via Verifiable Rewards RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T12:29:37.161294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:29:37.161294Z digest=sha256:82fce63aa4dcaf37896ae6e583c3e232ac6ed4527eee892cb407210918054755

Observation 84425ee0-fc77-48b2-ba75-c0e776e709ec · inbound

WildfireVLM: AI-powered Analysis for Early Wildfire Detection and Risk Assessment Using Satellite Imagery cites this paper.

WildfireVLM: AI-powered Analysis for Early Wildfire Detection and Risk Assessment Using Satellite Imagery RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:17:22.437778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T05:14:25.778239Z digest=sha256:014682e1350a9baff058f29590dbdb2dba8a3f30d23d8e711601eccb5ec5728c

Observation fdae9e8b-69a3-4b7a-b2d3-c0fbc8a65ea0 · inbound

Geo2Sound: A Scalable Geo-Aligned Framework for Soundscape Generation from Satellite Imagery cites this paper.

Geo2Sound: A Scalable Geo-Aligned Framework for Soundscape Generation from Satellite Imagery RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:59:03.522206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T09:52:35.741400Z digest=sha256:6e26fbddf0dd126611db0279def9cd922e8a8076bc1d8d0fd8278d78250e3439

Observation 149d635b-41a2-4a53-be7b-bcbc09121d74 · inbound

ChangeQuery: Advancing Remote Sensing Change Analysis for Natural and Human-Induced Disasters from Visual Detection to Semantic Understanding cites this paper.

ChangeQuery: Advancing Remote Sensing Change Analysis for Natural and Human-Induced Disasters from Visual Detection to Semantic Understanding RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:06:09.807361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T12:41:38.571997Z digest=sha256:31da9037ca0d32f54e8230142af8ccb854557833dd7db6526235bf6cb0ab5994

Observation fcf196cd-2b09-4b70-bd31-e1c570e8d72e · inbound

Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs cites this paper.

Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:15:56.985551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-11T01:52:51.777228Z digest=sha256:6bfcb7da52fb17fac421b3bdf61ec278b0c499260b928c63298c82144869f191

Observation 8dcd6151-2a25-462d-a736-66e18c306fc0 · inbound

GeoVista: Visually Grounded Active Perception for Ultra-High-Resolution Remote Sensing Understanding cites this paper.

GeoVista: Visually Grounded Active Perception for Ultra-High-Resolution Remote Sensing Understanding RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:53:33.974344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-15T02:39:56.666424Z digest=sha256:ccaed2ba50c37cbf92d9d2c7d009fe2c161fdbe7bb4cb8d60a3481c232898def

Observation 4419f07d-7ece-4962-831b-b018f9877d56 · inbound

OmniCD: A Foundational Framework for Remote Sensing Image Change Detection Guided by Multimodal Semantics cites this paper.

OmniCD: A Foundational Framework for Remote Sensing Image Change Detection Guided by Multimodal Semantics RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T08:13:16.011552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T08:03:27.451288Z digest=sha256:0d236603fca425021ca8f73e6b68dee9dd9f9706f19603623ed4197b9b8dbcc6

Observation 905bb316-54d7-40f5-a2d2-cd856d52b5c4 · inbound

RSICCLLM: A Multimodal Large Language Model for Remote Sensing Image Change Captioning cites this paper.

RSICCLLM: A Multimodal Large Language Model for Remote Sensing Image Change Captioning RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T17:05:51.455236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T04:09:32.397341Z digest=sha256:c197021d761ff4255632ddbccd58826770c04491908f417b02822053de576e87

Observation 702be6a7-ebeb-4edb-9787-cc5e61a184f9 · inbound

WeaveEarth: Structured Evidence Construction and Reasoning for Training-Free UHR Remote Sensing Understanding cites this paper.

WeaveEarth: Structured Evidence Construction and Reasoning for Training-Free UHR Remote Sensing Understanding RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-14T14:09:30.395518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T14:09:30.395518Z digest=sha256:2000948b1fd544dd137dff1a0548028618848334bf81bb4f48fa3fef1dc6224b