Pith. sign in

Paper Citation Record · LEDGER

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks

As of 14 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.07411.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07411 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T05:03:11.987987Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact6
  • verified fuzzy2
  • unresolved17
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b29a3ab4-ba31-4f82-86ba-5941050d75d7 · outbound

This paper cites Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.859067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.859067Z digest=sha256:36e1a0128b827d23253bed208428531d0b07d8985f905ea73743ac3a691a0d89

Observation 1243e9bb-a124-4ac9-9f85-12dba1f655bb · outbound

This paper cites 2026.GeoBenchmark: Probing Large Language Models for Geo-Spatial Knowledge.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks 2026.GeoBenchmark: Probing Large Language Models for Geo-Spatial Knowledge

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T05:03:12.950320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.865139Z digest=sha256:955a68c41d67fe6dfc5d41e440cfa697b1036905398391d9800cee3629627abc

Observation 0ae1e18d-e2c5-4cb3-9464-d3cc6cdf0ae0 · outbound

This paper cites MS MARCO: A Human Generated MAchine Reading COmprehension Dataset.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks MS MARCO: A Human Generated MAchine Reading COmprehension Dataset

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.870231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.870231Z digest=sha256:9c07cc5ac3cc9c968e18fcf20d96f9f72f24de5287d2f0720b10bbc0a45d5fa3

Observation d1f12f62-fd5e-4890-aa91-28c67c241c34 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.935439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.875298Z digest=sha256:2f984acf7c1c54fa8b6f3d7a79c3209399020e7d074f8e2d0c4a57c9bb863e35

Observation f31abbd8-dc67-471e-bf29-3069bb1c2fd1 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.920975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.884565Z digest=sha256:89f6bff6241ac6850e49f07d1f5f3867fd5d055bf0bd92e960b874c08f8af1f3

Observation 898014ab-fabe-47cc-b51d-c82713aab38d · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 6

Resolution
malformed identifier
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.767628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.888820Z digest=sha256:da217f113478718deac0e3fef4cc1999d4872a9b3c4fd27f795ad97f040a29b5

Observation 18b06fc2-66e8-44c4-8bd7-8663da8aafe8 · outbound

This paper cites Kummerfeld, Li Zhang, Karthik Ra- manathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Kummerfeld, Li Zhang, Karthik Ra- manathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev

Reference 7

Resolution
verified exact
doi, observed 2026-08-10T05:03:12.036948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.893227Z digest=sha256:f0629864a3f2febe6f688043d5c056b65859fe9e32cb50ad1ed97ba542acfaeb

Observation 41bfce68-ec68-495a-a591-8bbe64bf0386 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.906588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.898261Z digest=sha256:81ab55ec5ff3d83f1729305e1960fcc5644070d12190ed43ddb4271490e6e7bb

Observation b4143ca2-f8a7-42b8-acbf-72e82efc93c6 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 9

Resolution
malformed identifier
no resolver link, observed 2026-08-10T05:03:11.902705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.902705Z digest=sha256:e60bb8a0cc6e00b28ebc1e9c30782e81486732f96b73b3a98888697d5da0603d

Observation d1ddcc79-2481-45d6-8f98-b7a3d5e64b9a · outbound

This paper cites When Retriever-Reader Meets Scenario-Based Multiple-Choice Questions.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks When Retriever-Reader Meets Scenario-Based Multiple-Choice Questions

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.594227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.906911Z digest=sha256:f22e849382e8ebf8de2d31730495d2d2e5beb9a69fe7d90e60f453b58a5f3239

Observation 461d081f-8558-4617-bf5d-40bef1bc4ec3 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.882941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.911376Z digest=sha256:c2139afb5086328cbbfc9802bf6ea3cb8234f71f2be1616b179978f88eb6b032

Observation 193be87d-49b2-4449-b6d3-764a7e3b3fce · outbound

This paper cites Location Aware Modular Biencoder for Tourism Question Answering.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Location Aware Modular Biencoder for Tourism Question Answering

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.571961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.916168Z digest=sha256:c2718f47864e4e70f3583b96d93bfcaa392bdac130b342a73a604887b5609008

Observation 2e7a74a1-2726-4083-98aa-b16793dbc806 · outbound

This paper cites GridRoute: A Benchmark for LLM-Based Route Planning with Cardinal Movement in Grid Environments.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks GridRoute: A Benchmark for LLM-Based Route Planning with Cardinal Movement in Grid Environments

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.920484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.920484Z digest=sha256:c34614ae0fd404b64090981b156327520f8ba27c66f0d56d8c655ffd41764cd5

Observation d3a9abd7-488d-4076-a13c-b9e7f241dc99 · outbound

This paper cites STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.925122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.925122Z digest=sha256:eee068a08775d5f2a0418d45fdfc8b4f4303861dc8b10789bef5fb7ca03a0389

Observation a34b83a5-0c74-4b30-b27d-1620ab96a38a · outbound

This paper cites MapQA: Open-domain Geospatial Question Answering on Map Data.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks MapQA: Open-domain Geospatial Question Answering on Map Data

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.929941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.929941Z digest=sha256:f6b5553a701a56178400299ccf6aedb6977c410767c188ec751c7b5de12e65d5

Observation 80dbe486-ce64-400d-b954-2c3ec0deb3bb · outbound

This paper cites Geographic Question Answering: Challenges, Uniqueness, Classification, and Future Directions.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Geographic Question Answering: Challenges, Uniqueness, Classification, and Future Directions

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.503759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.934820Z digest=sha256:d246fcb2d99a94db0bb3fd116c16013e56d97c7e8d1baf7f4b2e02bae0fa2932

Observation c21d67c9-5d70-4df1-b96e-3d8154952b83 · outbound

This paper cites GeoLLM: Extracting Geospatial Knowledge from Large Language Models.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks GeoLLM: Extracting Geospatial Knowledge from Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.939699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.939699Z digest=sha256:e07787ce4240766c26374ef7854389f54cf284f253ed9a2f2ff586d4c61040ec

Observation 95dcee66-3548-4287-b9a8-b9a604c62c60 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 18

Resolution
malformed identifier
raw_fallback, observed 2026-08-10T05:03:12.868619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.944499Z digest=sha256:59c1f3b845ebbcb1266f4e4fd19b5d40a9708e3106fccf5751a0def192d070db

Observation 09bc2716-bf8c-4f47-91b4-0ffe1cc5b046 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.949570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.949570Z digest=sha256:c37c38d7fb9369c9743d8aa503cf9e25c5d685426cea434de576b625bf6eaaeb

Observation e7fcb46f-6059-432e-91fe-444767dbe26e · outbound

This paper cites Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.954622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.954622Z digest=sha256:32d95c0e0c765ab44630802ae1868abef0fca5acea68fdf753b8b501f9c416a3

Observation c9526de0-46b9-4c11-8933-710b74330e2f · outbound

This paper cites Evaluating Large Language Models on Spatial Tasks: A Multi-Task Benchmarking Study.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Evaluating Large Language Models on Spatial Tasks: A Multi-Task Benchmarking Study

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.959631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.959631Z digest=sha256:d944cfaca6625d17a7ab243874eb4d6b34e4083fd7bbc706f044c59f6bfe5d83

Observation 78f93cbe-31e3-45f4-9e16-1c6ee9a1e4a4 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.964369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.964369Z digest=sha256:b70b4bb186f2a9481062567641578f5b0f5fd734704521a690e8e69283456c86

Observation fcda4f81-077c-4bca-aa95-1c4558a5ddb9 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 23

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.434184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.973919Z digest=sha256:52ce34accfc0c13e1c8d77c8358aa8f830c70fc2d9905d56371267a6a02590bf

Observation 8624ffda-299e-49f4-bfac-82692c81a839 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 24

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.252810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.978541Z digest=sha256:8f39d0561b0b987ab41f41a17a98711f0cb0414a1e65b613b29260c3a96c889e

Observation 6c35c5d7-cc75-42eb-97ce-9a8d03fe1a13 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.828565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.983283Z digest=sha256:49ad64096a2480d1b4c5ae65d95d9097642045940d67b02589d957f574ad19ea

Observation 6697a866-c5c8-489a-adca-4d75d1c6d205 · outbound

This paper cites Zelle and Raymond J.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Zelle and Raymond J

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T05:03:12.813608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.987987Z digest=sha256:23a1bb2913fca2765c514053864bfb1aceec0dbcfa8b565192e83d78b6d2b715

Observation cd7dade0-df3f-4d1a-adb4-09abed50e8a3 · outbound

This paper cites In Proceedings of the 30th ACM International Conference on Information & Knowledge Management(Virtual Event, Queensland, Australia)(CIKM ’21).

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks In Proceedings of the 30th ACM International Conference on Information & Knowledge Management(Virtual Event, Queensland, Australia)(CIKM ’21)

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.880072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.880072Z digest=sha256:d1d5dbc3571d7860d6d11288c1bff6178e94585bb91c3cad4632c5695e7ef43b

Observation c8cc3774-ee36-469c-af82-f0fe177c1202 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.844144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T05:03:11.969114Z digest=sha256:f952fcf354ee13dcb46402fb14f7bfbf409947c26fb213b923a0504e26b80577

Pith citing papers

No inbound Pith citation observations are available.