Pith. sign in

Paper Citation Record · LEDGER

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents

As of 22 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2505.17572.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17572 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:47:19.524779Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-22T05:47:29.031123Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T05:51:08.879906Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact2
  • verified fuzzy40
  • unresolved35
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d3bad177-8d16-4566-a713-eac1b4519d0a · outbound

This paper cites Smart sustainable cities of the future: An extensive interdisciplinary literature review.Sustainable cities and society, 31:183–212, 2017.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Smart sustainable cities of the future: An extensive interdisciplinary literature review.Sustainable cities and society, 31:183–212, 2017

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:30.053205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:12.046737Z digest=sha256:a93f41832b29873b19514c67b7186776fec3c82f83dad9324a2f2918156e9266

Observation 9885d6b5-dcee-466e-a819-85deb337d722 · outbound

This paper cites Box and jenkins: time series analysis, forecasting and control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Box and jenkins: time series analysis, forecasting and control

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.739759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:12.109721Z digest=sha256:7ffdd8ff2798f528d52a2eb072b05aef62dd9119a3e60ff51a63a8fcd79cf9c3

Observation bf73a06d-d4d3-45e8-a8e0-c00a25460d23 · outbound

This paper cites TEMPO: Prompt-based generative pre-trained transformer for time series forecasting.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents TEMPO: Prompt-based generative pre-trained transformer for time series forecasting

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.470939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:12.210665Z digest=sha256:56430076c21c86fc1f906a8bace1a3a2ece2cf80c7afbffcef38aafa755217a2

Observation c255d701-a394-4e5c-be7d-30674f53e250 · outbound

This paper cites Agentboard: An analytical evaluation board of multi-turn llm agents.Advances in Neural Information Processing Systems, 37:74325–74362, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Agentboard: An analytical evaluation board of multi-turn llm agents.Advances in Neural Information Processing Systems, 37:74325–74362, 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.217389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:12.317306Z digest=sha256:8698a9ae5068ddee4045e9cb2fe2093062a68492283e1ff5cb92798b00784954

Observation ad43525f-18d8-4cd5-943f-5c15873e9543 · outbound

This paper cites Graphwiz: An instruction-following language model for graph computational problems.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Graphwiz: An instruction-following language model for graph computational problems

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.443472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.443472Z digest=sha256:0fb2f3d219b45bf24848ec137fd09d589f99cdc9f7c97fe02a13141372cf53f7

Observation 6c75bcdb-df9a-46e9-9849-480d16ce3eba · outbound

This paper cites Deeptransport: Learning spatial-temporal dependency for traffic condition forecasting.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Deeptransport: Learning spatial-temporal dependency for traffic condition forecasting

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.987240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:12.538621Z digest=sha256:b56f0c98a8ca9fd94d8d1f1fe3771f15e9f1e65a4705bf9facb0a62a34bdbeb0

Observation a3e9f2f7-2aa0-4acc-8083-40f2d7188618 · outbound

This paper cites TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.611434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.611434Z digest=sha256:b9eaa9b44c8b4575671d618179be0f15d76e91b1a8d6fc708618d6c2c5652984

Observation c8fe3724-6bf0-4dbf-8be1-42ebe63c9e00 · outbound

This paper cites On the evolution of random graphs.Publ.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents On the evolution of random graphs.Publ

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.735372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:12.700990Z digest=sha256:48ac3e77fb29ad0abc4dbd476f5b9e4be52012cd67ab917e321c7499174b196c

Observation a59d186c-7fc3-48c5-9aa1-90e8e9f54f9b · outbound

This paper cites Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.799685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.799685Z digest=sha256:9b648ad0b49c2f3a11639feefcd727149c8f3d17c43be8c88366586857ddc7ce

Observation 90ac15c2-34fb-41be-9f7e-e7ba2f5f5eb2 · outbound

This paper cites CityGPT: Empowering Urban Spatial Cognition of Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CityGPT: Empowering Urban Spatial Cognition of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.923658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.923658Z digest=sha256:c44104db49c1cd00096233eaca2329df33bff7c12b9cb2f2127d8de8584b7feb

Observation b05f7949-d8c9-498a-9b12-096d40ba262b · outbound

This paper cites CityBench: Evaluating the Capabilities of Large Language Models for Urban Tasks.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CityBench: Evaluating the Capabilities of Large Language Models for Urban Tasks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.027010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.027010Z digest=sha256:1177a03a50b27126ce6bd6a6e050b6e7dae7fba04991b4adf74d1f24532fa072

Observation 56f07dc6-d6a3-4acb-8701-9bc858e7f096 · outbound

This paper cites Pygad: An intuitive genetic algorithm python library.Multimedia tools and applications, 83(20):58029–58042, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Pygad: An intuitive genetic algorithm python library.Multimedia tools and applications, 83(20):58029–58042, 2024

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.524138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:13.132409Z digest=sha256:a432e6e6c91b995612b43fd94af6dda4b242587f3a163a6f2f466bc3d2901238

Observation 301f1204-0906-4602-ba2f-615260bd1b80 · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.204895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.204895Z digest=sha256:b8c16160c2de9a426088e296aa0ab7676add9a4be1cd31a5fc6836e129833f83

Observation b1aca464-8f36-4df2-b1a8-765071066833 · outbound

This paper cites The Llama 3 Herd of Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents The Llama 3 Herd of Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.286626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.286626Z digest=sha256:e269952e42b2a81f56404436459ebc2d6d3974086af994cda5bfcd0c2df4a704

Observation d35c513c-5a74-454e-882a-19d19dea1460 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.387728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.387728Z digest=sha256:8af6e2e52538f2059d5fb58891886e96560035d71c4535a080282e02e9efa0e6

Observation 88da6acb-25fb-4d14-bf4a-95722d3d2e20 · outbound

This paper cites Language Models Represent Space and Time.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Language Models Represent Space and Time

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.477756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.477756Z digest=sha256:c6b3e8f2a0c85254e4200b70a7234cd0432207410752e396b19110e23f8dff6d

Observation 06b0d6f0-dd10-4535-9cb6-71fdfe62f4c8 · outbound

This paper cites The scoot on-line traffic signal optimisation technique.Traffic Engineering & Control, 23(4), 1982.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents The scoot on-line traffic signal optimisation technique.Traffic Engineering & Control, 23(4), 1982

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.285556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:13.575875Z digest=sha256:8d3cf1efeaac870a1802ce19b1b0d9c4e2ef456879a8e91e3b7fa29f77d48ee5

Observation 50b9e666-f72a-4045-93a3-4acfb89323e9 · outbound

This paper cites GPT-4o System Card.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents GPT-4o System Card

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.695425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.695425Z digest=sha256:3e5ce2652338d369c3010f3687cb017282ce8d3c453821c02475d4c5f81476a9

Observation 49476d7d-4ca7-496c-ad4c-af8166894df1 · outbound

This paper cites OpenAI o1 System Card.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents OpenAI o1 System Card

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.795554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.795554Z digest=sha256:af382d3f1b0aea3f68136e6730022166c4e92fcc733712682285e142b33fd8d5

Observation e88a2df5-49c1-4ece-898c-30b2188c82e7 · outbound

This paper cites Towards mitigating LLM hallucination via self reflection.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards mitigating LLM hallucination via self reflection

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.899032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.899032Z digest=sha256:61ca96858ab5f9e5007b9b5e30eafd9a55f29ad8271facbaa0c2b77a607140ca

Observation c9424195-67a4-4265-9689-cd4d5e9eb84f · outbound

This paper cites Llmlight: Large language models as traffic signal control agents.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Llmlight: Large language models as traffic signal control agents

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.013525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:14.009911Z digest=sha256:9a8658da4cca9ee2dc3b2a6f3508f5b1bacad94b6033a9273e2a0fce468038f2

Observation ecf904e5-b7a4-4121-929f-8c83c7c34bcc · outbound

This paper cites Reframing Spatial Reasoning Evaluation in Language Models: A Real-World Simulation Benchmark for Qualitative Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Reframing Spatial Reasoning Evaluation in Language Models: A Real-World Simulation Benchmark for Qualitative Reasoning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:20.582052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:14.089548Z digest=sha256:6097102cfab54ae5a27fdfda2202145510356cb60b5657ef178d53735f02f8dc

Observation ae512c5d-51a9-4e84-99e8-024664a580cc · outbound

This paper cites Repetition in repetition out: Towards understanding neural text degeneration from the data perspective.Advances in Neural Information Processing Systems, 36:72888–72903, 2023.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Repetition in repetition out: Towards understanding neural text degeneration from the data perspective.Advances in Neural Information Processing Systems, 36:72888–72903, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.169851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.169851Z digest=sha256:81196890c90fb6ca1b1fbe6fecde1b140b08a0601cad1adc364b2b9b9bdecc3f

Observation cde91bcb-a5c6-4f1a-9bdf-dc48f39c0cee · outbound

This paper cites Towards alleviating traffic congestion: Optimal route planning for massive-scale trips.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards alleviating traffic congestion: Optimal route planning for massive-scale trips

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.808459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:14.238835Z digest=sha256:ee2c5f3333215be70d0ac98c5535c9d91eef8aa929fa970cbb1c929b9ac1ee0f

Observation bf669141-44b7-4e7e-ab00-e559ef889780 · outbound

This paper cites STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.331330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.331330Z digest=sha256:475b5b38af3335c33c24e79b67227846c63264a66fc4f4536d9a7c33000afaf0

Observation e5fbb148-bc8a-42b0-a3c2-d058aaf4748a · outbound

This paper cites Urbangpt: Spatio-temporal large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Urbangpt: Spatio-temporal large language models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.625467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:14.400820Z digest=sha256:6792d9d60e214728ab80573d394ac8b75c30301266dc8b87879f64bf8498ed6c

Observation 10788f50-7666-48cf-b82f-6b5c41040519 · outbound

This paper cites Timecma: Towards llm-empowered multivariate time series forecasting via cross-modality alignment.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Timecma: Towards llm-empowered multivariate time series forecasting via cross-modality alignment

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.509291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.509291Z digest=sha256:4e7597f375a5dcdcd761e03d706ab8946c55044ef16aa5577d514e665acede22

Observation a0ec181d-3d0c-40c4-a3e0-f98ee4e6e9a3 · outbound

This paper cites Knowledge-infused contrastive learning for urban imagery-based socioeconomic prediction.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Knowledge-infused contrastive learning for urban imagery-based socioeconomic prediction

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.411860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:14.582881Z digest=sha256:4f88dc2ebac4a8e0fb68918cf0bcb3573acfa364a787852375e05c89e713b89d

Observation 422597ef-7eb4-488b-b0b3-4004cea8b6ad · outbound

This paper cites Simulation of urban mobility (sumo), February 4 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Simulation of urban mobility (sumo), February 4 2025

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.198446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:14.661472Z digest=sha256:7cc67cfd0de147e8e225e7140b3e082f78d532f4fdddf097935ec3f64b4a5cbb

Observation e3b83036-12e7-4dff-ab0a-8f649ede33e4 · outbound

This paper cites Scats, sydney co-ordinated adaptive traffic system: A traffic responsive method of controlling urban traffic.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Scats, sydney co-ordinated adaptive traffic system: A traffic responsive method of controlling urban traffic

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.016511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:14.750231Z digest=sha256:4d1b9d2e137394a934b6a4e9632a445a2c627f30f93e079116b98a4068a15a7d

Observation f747e7a6-3ced-4808-ac9c-b3c9e30c2a34 · outbound

This paper cites SpartQA: : A Textual Question Answering Benchmark for Spatial Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents SpartQA: : A Textual Question Answering Benchmark for Spatial Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.883698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.883698Z digest=sha256:b9759dc97e732a01335e3458c277fcdf8973892772ffdb1cb5a5efbcfaa507fd

Observation d259570f-b2c7-4498-b981-7de7b44aa08a · outbound

This paper cites Transfer Learning with Synthetic Corpora for Spatial Role Labeling and Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Transfer Learning with Synthetic Corpora for Spatial Role Labeling and Reasoning

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:20.222720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:14.953677Z digest=sha256:ebe0bbfc8dfb601a3b3ccda3d0ae85d133b51136a6bb5b130b895aa55c82f2ef

Observation 4828f170-3d74-4f7d-b4a9-e47b21946184 · outbound

This paper cites Towards understanding the spatial literacy of chatgpt.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards understanding the spatial literacy of chatgpt

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.760392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:15.053874Z digest=sha256:45635269489f154ca61d80ad56bcacbf0ba63b24afd956c89e6ae222fb8e727f

Observation 14f51320-84f5-495e-a4e6-d249685ad8dc · outbound

This paper cites Tlc trip record data, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Tlc trip record data, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.551616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:15.128952Z digest=sha256:4ed4340659ca853341b7c07ba6a9bf2612bf68df082783d5c73121def570ad21

Observation 4cc97432-e7f0-474a-9398-f6319ecbaac2 · outbound

This paper cites Dima: An llm-powered ride-hailing assistant at didi.arXiv preprint arXiv:2503.04768, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Dima: An llm-powered ride-hailing assistant at didi.arXiv preprint arXiv:2503.04768, 2025

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.195897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.195897Z digest=sha256:e0825fa97b716d38627028dc94e145e07d97cca9b5afcf9057fe24561a53e58c

Observation dec75734-d12d-4434-b73d-9313191c4b88 · outbound

This paper cites UrbanKGent: A Unified Large Language Model Agent Framework for Urban Knowledge Graph Construction.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents UrbanKGent: A Unified Large Language Model Agent Framework for Urban Knowledge Graph Construction

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.268020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.268020Z digest=sha256:768f3a934a1c08dc52919f42f565de721fe1ccf47ecb9ed1c948ac32dafff09b

Observation cb413b5a-1f9e-4f3c-9391-efd7c0e6903d · outbound

This paper cites Openstreetmap planet data, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Openstreetmap planet data, 2025

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.285455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:15.353419Z digest=sha256:30bf67b18855b1a173c2874860ac84b4011268d44b61b6767703eb73ccad6c8f

Observation 6ad461f8-920f-470c-84de-08b57ba8d8c7 · outbound

This paper cites Self-Reflection in LLM Agents: Effects on Problem-Solving Performance.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Self-Reflection in LLM Agents: Effects on Problem-Solving Performance

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.428482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.428482Z digest=sha256:7ba5a1d0c5e57906cd917561209fdd5f5599ab4ba89aeea5e74a80ca0747bd51

Observation 723dd069-13af-447a-a128-06ea71470315 · outbound

This paper cites Sparc and sparp: Spatial reasoning char- acterization and path generation for understanding spatial reasoning capability of large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Sparc and sparp: Spatial reasoning char- acterization and path generation for understanding spatial reasoning capability of large language models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.080108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:15.517387Z digest=sha256:a9d4c421dba96fae2214de7e14c06070330de9981fc91a24142977b75dddbffe

Observation a69353f3-9ea3-4fe3-bebd-270f11f1adc1 · outbound

This paper cites Proximal Policy Optimization Algorithms.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Proximal Policy Optimization Algorithms

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.585405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.585405Z digest=sha256:681fd3b9f2020fa602cdf998a6b1acb311993b590e4a0f2c72afd690899a5701

Observation 4daf5592-2ac2-4e64-8a55-200258dc8fc8 · outbound

This paper cites Stepgame: A new benchmark for robust multi- hop spatial reasoning in texts.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Stepgame: A new benchmark for robust multi- hop spatial reasoning in texts

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.851533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:15.674222Z digest=sha256:61e24474019bf4ea9125152caf69e1c14953e6629da585fe168ffbf8561b14ec

Observation c251b9f7-f9ca-43fe-b35f-5a38e669c87a · outbound

This paper cites Towards Benchmarking and Improving the Temporal Reasoning Capability of Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards Benchmarking and Improving the Temporal Reasoning Capability of Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.798080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.798080Z digest=sha256:15797480486937f4bc0321cfa8efba34618838233544943ba7b1af6f48b132f4

Observation 0f1326f0-41f0-49e4-8514-d1ebba27bd56 · outbound

This paper cites Cityflow: A city-scale benchmark for multi-target multi-camera vehicle tracking and re-identification.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Cityflow: A city-scale benchmark for multi-target multi-camera vehicle tracking and re-identification

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.623904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:15.887239Z digest=sha256:fe5224f9ccd99d48ec152814ac59d1997f535cb59472cc76349db2d04cb5bf23

Observation fa4bfccf-fd74-4b0e-a578-54a27441a953 · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Qwq-32b: Embracing the power of reinforcement learning, 2025

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.303061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:15.978560Z digest=sha256:38ab740e402b8ca000f6816e61f95a0c74cf0706c5a95d320f8363b8f2960236

Observation f4af302b-8d6b-4e6b-8601-3ad39f849674 · outbound

This paper cites Air quality prediction with physics-guided dual neural odes in open systems.ICLR, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Air quality prediction with physics-guided dual neural odes in open systems.ICLR, 2025

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.077903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:16.075275Z digest=sha256:ee863863143f7ff5807f4acf5273487cb08d281730fd26639decfe3a8d4fc8a1

Observation 2552d402-3c96-4f5b-ad90-9fbdb48000f7 · outbound

This paper cites Applications of artificial intelligence and machine learning in smart cities.Computer Communications, 154:313– 323, 2020.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Applications of artificial intelligence and machine learning in smart cities.Computer Communications, 154:313– 323, 2020

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:24.803848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:16.158597Z digest=sha256:67349a4aae0519a6f537849c63c5dee473c3e7f12395afbd439cebe6ee44e69a

Observation af1e7ce4-b8e8-4356-a1e3-b3f04352730e · outbound

This paper cites Robust extrema features for time-series data analysis.IEEE transactions on pattern analysis and machine intelligence, 35(6):1464–1479, 2012.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Robust extrema features for time-series data analysis.IEEE transactions on pattern analysis and machine intelligence, 35(6):1464–1479, 2012

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:24.526996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:16.246214Z digest=sha256:bf1110bd658d2cd530bd5f83f9921a5f844436e664fd4bb8e3b11b8b385d3ada

Observation e2e9300c-94fa-4f4f-b587-ac87034a2781 · outbound

This paper cites Reinforcement learning-based placement of charging stations in urban road networks.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Reinforcement learning-based placement of charging stations in urban road networks

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:24.223000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:16.325026Z digest=sha256:d3d1b1fbcfbbdfd2a777f12912ec3e74488fc11434efbd92db4d5100aa3511f6

Observation 19ad2d6f-34f2-4d09-9f88-daae5c60538b · outbound

This paper cites A survey on large language model based autonomous agents.Frontiers of Computer Science, 18(6):186345, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents A survey on large language model based autonomous agents.Frontiers of Computer Science, 18(6):186345, 2024

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.401542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.401542Z digest=sha256:858f0e5d15247e565b2d14cfca1cd3fba00734f5aa82f9a083adebd3217391c7

Observation 90c123ca-ef52-4e9c-982c-e2b79595fb2c · outbound

This paper cites Global gridded gdp data set consistent with the shared socioeco- nomic pathways.Scientific data, 9(1):221, 2022.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Global gridded gdp data set consistent with the shared socioeco- nomic pathways.Scientific data, 9(1):221, 2022

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.916545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:16.480492Z digest=sha256:05382d30873da6e45d630f06f4159737860f2a73c4a0998d37c72f930bf5aded

Observation bc30ab37-8be1-49a4-99da-2284ce663a53 · outbound

This paper cites Where Would I Go Next? Large Language Models as Human Mobility Predictors.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Where Would I Go Next? Large Language Models as Human Mobility Predictors

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.540853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.540853Z digest=sha256:9ee9c2985bf836998f2d95ae13b43e1a7afe706f1ffcc3d7d10b029e9f22412b

Observation 631cb0a6-a581-461d-bb63-c6f357f4aa12 · outbound

This paper cites From news to forecast: Integrating event analysis in llm-based time series forecasting with reflection.Advances in Neural Information Processing Systems, 37:58118–58153, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents From news to forecast: Integrating event analysis in llm-based time series forecasting with reflection.Advances in Neural Information Processing Systems, 37:58118–58153, 2024

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.622503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.622503Z digest=sha256:eb48aa8ceab7da7a8b02e0d32f78fd66cd6dc473e9b1c507158735ffe8410bd3

Observation d99b94fc-7b0c-45fb-994c-9ed5ba95a5ac · outbound

This paper cites Tram: Benchmarking temporal reasoning for large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Tram: Benchmarking temporal reasoning for large language models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.698546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:16.712157Z digest=sha256:19345942a4b845723790475144c5b337eab85b89424246f4fc83bcd91e993ad8

Observation 3d1eb627-a4e6-4049-aa82-7a20f09596ef · outbound

This paper cites Colight: Learning network-level cooperation for traffic signal control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Colight: Learning network-level cooperation for traffic signal control

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.475729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:16.786439Z digest=sha256:7f4346d76fc413e7aaf8c191b134903ebb73bf3b42027bf325ea9ef6b0e37f7d

Observation 92fe0cfb-22d7-44ae-9710-dde0885ad2fb · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Chain-of-thought prompting elicits reasoning in large language models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.887823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.887823Z digest=sha256:906978fed1efc386ae457f0599d889167a985df9bd12e350b28dfa661252dd21

Observation 505bf480-0890-478b-b977-1c17a5ad767a · outbound

This paper cites Coverage location models: alternatives, approximation, and uncertainty.International Regional Science Review, 39(1):48–76, 2016.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Coverage location models: alternatives, approximation, and uncertainty.International Regional Science Review, 39(1):48–76, 2016

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.247621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:16.946618Z digest=sha256:841c2ce48faaa84cd28e77c32f3d66dfce5bfc36c4863bf697364c1bfe0aab2c

Observation 0096260e-70f7-423c-baae-ebacf7b67b57 · outbound

This paper cites Worldpop hub, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Worldpop hub, 2025

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.019759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:17.010800Z digest=sha256:cf9e64bd0298fb88023afc26d7bf160e31ae23d94aa8874155cd8cc3bfd45e89

Observation 65fc1a15-7390-4d78-90d8-35e328c4f2ab · outbound

This paper cites Large language models can learn temporal reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Large language models can learn temporal reasoning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.045214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.045214Z digest=sha256:ea972378c839e4fb8d9590ac4794b1591666b0a1ff7464d90737956fac353957

Observation 63d0c6b9-0757-42fd-99dc-eea7f6aee440 · outbound

This paper cites Evaluating Spatial Understanding of Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Evaluating Spatial Understanding of Large Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.109005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.109005Z digest=sha256:692c74b6b967efd78c2fda7cf9b79da5bf3196329355a867b81b29d925a459f7

Observation 50dc80f6-64a3-494a-bcaa-7b0934374277 · outbound

This paper cites Qwen2.5 Technical Report.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Qwen2.5 Technical Report

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.170094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.170094Z digest=sha256:255cd82e203fa243848c245ed7674caac10c3299decf4a4a53f7486cbccadf7b

Observation a7da1f14-7825-47e0-8d35-858f2469cd3b · outbound

This paper cites Foursquare dataset.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Foursquare dataset

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.831190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:17.230652Z digest=sha256:3dacc33b5d9c594a37f7317540486f3959688edcf839216ac4ecf8c748144ede

Observation 0947063f-e531-4e33-b803-ac5d06090025 · outbound

This paper cites Unist: A prompt-empowered universal model for urban spatio-temporal prediction.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Unist: A prompt-empowered universal model for urban spatio-temporal prediction

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.582495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:17.315597Z digest=sha256:6e9acbcc533a64257741d84329c089b330108020e1250720cafab784479c5d81

Observation 534d8ca3-9355-478c-be5e-d2617d2ce9cb · outbound

This paper cites CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal Control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal Control

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.318918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.318918Z digest=sha256:ac871c6f74070bbf7336f8922991a4a6ff6d21358f706c17794d7116a22f3b8a

Observation a3a05314-e66e-49aa-a2ab-2a28e89ac2ae · outbound

This paper cites AgentTuning: Enabling Generalized Agent Abilities for LLMs.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents AgentTuning: Enabling Generalized Agent Abilities for LLMs

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.428562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.428562Z digest=sha256:123fc216d508366ef070c1dd2ecdf3a2415129a8ac16670c7622f544a222b802

Observation 48950252-710c-4907-94f0-437c9c7b95d3 · outbound

This paper cites Open3dvqa: A benchmark for comprehensive spatial reasoning with multimodal large language model in open space.arXiv preprint arXiv:2503.11094, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Open3dvqa: A benchmark for comprehensive spatial reasoning with multimodal large language model in open space.arXiv preprint arXiv:2503.11094, 2025

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.608625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.608625Z digest=sha256:8631015aedb2f96353f5c7e1f470626606da80a9463dd616bf8f9bcb328edd4d

Observation 2d326e74-c59f-4b3e-8183-5e397f5281bb · outbound

This paper cites Reinforcement learning for traffic signal control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Reinforcement learning for traffic signal control

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.404837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:17.786637Z digest=sha256:1b339eb4142bcac1e1ee143ecd0f5410ed4bb854255b0f9690de8ff0add5a8e7

Observation fabf270e-123f-4fc7-9ce1-332b15430add · outbound

This paper cites Urbanvideo-bench: Benchmarking vision- language models on embodied intelligence with video data in urban spaces.arXiv preprint arXiv:2503.06157, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Urbanvideo-bench: Benchmarking vision- language models on embodied intelligence with video data in urban spaces.arXiv preprint arXiv:2503.06157, 2025

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.931121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.931121Z digest=sha256:0e37bd8ebbcc1514ac96b0c5ed2e008932a9599a33e831628d3ec82a0315f260

Observation 0d75fb69-a602-4e7e-8442-29c2235cce72 · outbound

This paper cites Where to go next: A spatio-temporal gated network for next poi recommendation.IEEE Transactions on Knowledge and Data Engineering, 34(5):2512–2524, 2020.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Where to go next: A spatio-temporal gated network for next poi recommendation.IEEE Transactions on Knowledge and Data Engineering, 34(5):2512–2524, 2020

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.238763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:18.112245Z digest=sha256:88547ca9e9e107e2c77bdf78e933d37ef7f091fbaf1e6f372e62c66f47c439cc

Observation 41cab77f-526e-4780-82bd-3ba195d55daf · outbound

This paper cites CityEQA: A Hierarchical LLM Agent on Embodied Question Answering Benchmark in City Space.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CityEQA: A Hierarchical LLM Agent on Embodied Question Answering Benchmark in City Space

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:18.279741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:18.279741Z digest=sha256:f9385d16c5b5fc22430fdbfb427cfa760bb9ad3d321e920a4eac0eaf19c90215

Observation 6832a9e1-703e-4f49-b092-d7d0f99a8b58 · outbound

This paper cites Llamafactory: Unified efficient fine-tuning of 100+ language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Llamafactory: Unified efficient fine-tuning of 100+ language models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:18.409485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:18.409485Z digest=sha256:d90e6cb748ba1cf5acb085db9fd8990631947324731771bf09387da4154f4c8c

Observation 19360793-b3c5-4839-83e3-15ec9d74624c · outbound

This paper cites Spatial planning of urban communities via deep reinforcement learning.Nature Computational Science, 3(9):748– 762, 2023.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Spatial planning of urban communities via deep reinforcement learning.Nature Computational Science, 3(9):748– 762, 2023

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.005609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:18.533561Z digest=sha256:360950da30f930016a2ba1e6d1e28d875903bb5c6bb8540d4972cfcdcfa3816e

Observation be5f3e12-424a-4851-9e95-73b5a71994bf · outbound

This paper cites UrbanPlanBench: A Comprehensive Urban Planning Benchmark for Evaluating Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents UrbanPlanBench: A Comprehensive Urban Planning Benchmark for Evaluating Large Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:18.662366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:18.662366Z digest=sha256:48ecba39b606b78d04f4b1973b9369b2b8cb00c4ed176f8c08a4828fe2ed0a3f

Observation c5dd61e9-0ffb-464b-97af-6e169aee06af · outbound

This paper cites Road planning for slums via deep reinforcement learning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Road planning for slums via deep reinforcement learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.838077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:18.808980Z digest=sha256:d1a44a5d60c976c0b398717f08f89307555600e61d09ae91154427fe8c6f21db

Observation 61ce2295-ac49-4efe-a6f4-b1a2f38a06a2 · outbound

This paper cites going on a vacation.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents going on a vacation

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.632527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:18.941064Z digest=sha256:a6aa0ab3b0b82ecb2283e4535bbe05b2c7c9f372dd52e86ea7fe37b57ad666ce

Observation 8b91d25d-f6ec-49d2-9dac-c2c543ed9e90 · outbound

This paper cites Large Language Model for Participatory Urban Planning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Large Language Model for Participatory Urban Planning

Reference 75

Resolution
malformed identifier
no resolver link, observed 2026-08-07T14:47:19.072276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:19.072276Z digest=sha256:4b1b08a2e46475764e1ae01fc12b7eecfb7ba42ed4ebd9a3974ec26d079b0354

Observation 25e769a7-7f59-4d94-8167-be3f432879db · outbound

This paper cites answer":.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents answer":

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.417259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:19.242745Z digest=sha256:40c79bf87f5e491791ef37ede4f59b7da5bc0ae33b060be1366c1a1ee4811607

Observation 0a1940f9-9f5b-4dcf-a6ce-05b98748b6fb · outbound

This paper cites Miscellaneous Shop.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Miscellaneous Shop

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.155850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:19.378369Z digest=sha256:90323e053e83a3112fc2538f8bf9b18bc9005bcddef1358b58f186c06b4ae55a

Observation 682dc218-38ba-4888-9e43-f4757dff36fb · outbound

This paper cites answer":.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents answer":

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:20.945046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T14:47:19.524779Z digest=sha256:f5041a16e33cfd930c66fd66da78622a19e65c87a7d17c06159c241f0751a7d2

Pith citing papers

Observation 12238777-3eda-4da9-85cd-1e1a49d2e174 · inbound

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation cites this paper.

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:51:08.882918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-22T05:47:29.031123Z digest=sha256:d5aa968623978597f2505d1ffb9d87c7c6d60457be3d8d4f13734b9136e17f8c