Pith. sign in

Paper Citation Record · LEDGER

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks

As of 10 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.07411.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07411 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T05:03:11.987987Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact6
  • verified fuzzy2
  • unresolved17
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b29a3ab4-ba31-4f82-86ba-5941050d75d7 · outbound

This paper cites Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.859067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.859067Z digest=sha256:21f2c27af004cd51cc050ede607a39dc3791673c779aade0882b7e276dcbfa6c

Observation 1243e9bb-a124-4ac9-9f85-12dba1f655bb · outbound

This paper cites 2026.GeoBenchmark: Probing Large Language Models for Geo-Spatial Knowledge.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks 2026.GeoBenchmark: Probing Large Language Models for Geo-Spatial Knowledge

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T05:03:12.950320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.865139Z digest=sha256:62dc35d2b973d05345af2607ec4f2419d758f0bc40298807f04e815f79c7c3da

Observation 0ae1e18d-e2c5-4cb3-9464-d3cc6cdf0ae0 · outbound

This paper cites MS MARCO: A Human Generated MAchine Reading COmprehension Dataset.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks MS MARCO: A Human Generated MAchine Reading COmprehension Dataset

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.870231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.870231Z digest=sha256:7840c3201ccaf84c7fd62c27a67c8d92f47ece553483fe6badc55608b958eb07

Observation d1f12f62-fd5e-4890-aa91-28c67c241c34 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.935439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.875298Z digest=sha256:4a6866acf9107fd3e3e742e0d5566fe09a7dce47d6bbc0b83d26b42696b1c9ab

Observation f31abbd8-dc67-471e-bf29-3069bb1c2fd1 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.920975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.884565Z digest=sha256:d0af987d7620ca6513334356aa517cc11700867d44c341fc004189b80dcc21eb

Observation 898014ab-fabe-47cc-b51d-c82713aab38d · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 6

Resolution
malformed identifier
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.767628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.888820Z digest=sha256:06ab9df8b312fab7c11f3856a1ec1cc919572bd6009d6ab009e3faca761229a4

Observation 18b06fc2-66e8-44c4-8bd7-8663da8aafe8 · outbound

This paper cites Kummerfeld, Li Zhang, Karthik Ra- manathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Kummerfeld, Li Zhang, Karthik Ra- manathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev

Reference 7

Resolution
verified exact
doi, observed 2026-08-10T05:03:12.036948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.893227Z digest=sha256:1c2e9116dd4084e916b3be38471eae8954357493a6005e32e8ac276fc236c1cb

Observation 41bfce68-ec68-495a-a591-8bbe64bf0386 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.906588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.898261Z digest=sha256:c57eabbbd5278aeda38e3692d9991fbdc3f43aada3b1ddadd5660a6820327ebf

Observation b4143ca2-f8a7-42b8-acbf-72e82efc93c6 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 9

Resolution
malformed identifier
no resolver link, observed 2026-08-10T05:03:11.902705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.902705Z digest=sha256:fa7b1cc57eba795fba19a03f91bcd33e542704ffc1b7b6d7c073c8cb612afebb

Observation d1ddcc79-2481-45d6-8f98-b7a3d5e64b9a · outbound

This paper cites When Retriever-Reader Meets Scenario-Based Multiple-Choice Questions.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks When Retriever-Reader Meets Scenario-Based Multiple-Choice Questions

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.594227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.906911Z digest=sha256:e48287fbe5854a79b36c3fb82936e0082cdd9925b8b16cd62c37b50d8477e4b6

Observation 461d081f-8558-4617-bf5d-40bef1bc4ec3 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.882941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.911376Z digest=sha256:c0c452624d1c70acf062efd268d1ee146b345f8946231cc6a980591147ef8279

Observation 193be87d-49b2-4449-b6d3-764a7e3b3fce · outbound

This paper cites Location Aware Modular Biencoder for Tourism Question Answering.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Location Aware Modular Biencoder for Tourism Question Answering

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.571961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.916168Z digest=sha256:73d2b443a4bd6188f886b07cdf0444fdb0f0c1e07f012bc096b26bb0118c5861

Observation 2e7a74a1-2726-4083-98aa-b16793dbc806 · outbound

This paper cites GridRoute: A Benchmark for LLM-Based Route Planning with Cardinal Movement in Grid Environments.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks GridRoute: A Benchmark for LLM-Based Route Planning with Cardinal Movement in Grid Environments

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.920484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.920484Z digest=sha256:809ac9ff27a12c08896fa779651269d0b40228f1ec7c6c3cc1fb27d3762ad105

Observation d3a9abd7-488d-4076-a13c-b9e7f241dc99 · outbound

This paper cites STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.925122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.925122Z digest=sha256:b052d5589703a97f321e2601c56600a7b3355373a237240a71b21bf36f48bec5

Observation a34b83a5-0c74-4b30-b27d-1620ab96a38a · outbound

This paper cites MapQA: Open-domain Geospatial Question Answering on Map Data.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks MapQA: Open-domain Geospatial Question Answering on Map Data

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.929941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.929941Z digest=sha256:3f9d2b8410faa9af9235f837cee233bc20a6ff263da6135ac24e78ab4d359b61

Observation 80dbe486-ce64-400d-b954-2c3ec0deb3bb · outbound

This paper cites Geographic Question Answering: Challenges, Uniqueness, Classification, and Future Directions.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Geographic Question Answering: Challenges, Uniqueness, Classification, and Future Directions

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.503759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.934820Z digest=sha256:700ec28e093d16d00386d5c036bd087c96be2e22d8afe0ce8716491c573afaa1

Observation c21d67c9-5d70-4df1-b96e-3d8154952b83 · outbound

This paper cites GeoLLM: Extracting Geospatial Knowledge from Large Language Models.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks GeoLLM: Extracting Geospatial Knowledge from Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.939699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.939699Z digest=sha256:492108a2cfd06650f8dceecbcb976e3e8285014a1ae4b42a648ff7657fc02204

Observation 95dcee66-3548-4287-b9a8-b9a604c62c60 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 18

Resolution
malformed identifier
raw_fallback, observed 2026-08-10T05:03:12.868619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.944499Z digest=sha256:fa6d358228d78943b8cddd979f391f1c6379d86daf80d5bda83083728d6018fb

Observation 09bc2716-bf8c-4f47-91b4-0ffe1cc5b046 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.949570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.949570Z digest=sha256:6d564bac667ab2685e3b929091826cf1c746a9a6208a15b2e2e257c98e595ae5

Observation e7fcb46f-6059-432e-91fe-444767dbe26e · outbound

This paper cites Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.954622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.954622Z digest=sha256:24cc614c2660f305a60406f5e1571474754286ed0925387bcf1f9231a2382a73

Observation c9526de0-46b9-4c11-8933-710b74330e2f · outbound

This paper cites Evaluating Large Language Models on Spatial Tasks: A Multi-Task Benchmarking Study.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Evaluating Large Language Models on Spatial Tasks: A Multi-Task Benchmarking Study

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.959631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.959631Z digest=sha256:c8c54834507cb6a7633ead37deda42bedb09e52522ddd9b68090d2dfb598df9e

Observation 78f93cbe-31e3-45f4-9e16-1c6ee9a1e4a4 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.964369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.964369Z digest=sha256:3ae18475077f3d3059e919fd774334009f39df2c935df59113a17d97245cb18c

Observation fcda4f81-077c-4bca-aa95-1c4558a5ddb9 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 23

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.434184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.973919Z digest=sha256:a97ff3ac05e84956cbb30e476ba563261ca06093bfa93d4210f573ff6d5bf721

Observation 8624ffda-299e-49f4-bfac-82692c81a839 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 24

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.252810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.978541Z digest=sha256:b2ba10e70cbb50d9e26ec913a9c02141944759dd62a5f6403007cdb615f29a90

Observation 6c35c5d7-cc75-42eb-97ce-9a8d03fe1a13 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.828565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.983283Z digest=sha256:d7e4bb7921de22df4ece6157e5d9b8cbd15428025f0583cb051a2504d5af8c34

Observation 6697a866-c5c8-489a-adca-4d75d1c6d205 · outbound

This paper cites Zelle and Raymond J.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Zelle and Raymond J

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T05:03:12.813608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.987987Z digest=sha256:fd28d074a7718e2a72aba74b68fd8fd5955ae293d884ba8e82aa5df6dee9587a

Observation cd7dade0-df3f-4d1a-adb4-09abed50e8a3 · outbound

This paper cites In Proceedings of the 30th ACM International Conference on Information & Knowledge Management(Virtual Event, Queensland, Australia)(CIKM ’21).

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks In Proceedings of the 30th ACM International Conference on Information & Knowledge Management(Virtual Event, Queensland, Australia)(CIKM ’21)

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.880072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.880072Z digest=sha256:7760afd4e16cec2624af7934634597b3e325d21ba0b320447af354c0d4aff973

Observation c8cc3774-ee36-469c-af82-f0fe177c1202 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.844144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.969114Z digest=sha256:1ff4b0e47fe3b3914920d3de1160b1283c39cda64bc2388b5345b7bc4d974ce6

Pith citing papers

No inbound Pith citation observations are available.