Pith. sign in

Paper Citation Record · LEDGER

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents

As of 9 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2505.17572.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17572 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:47:19.524779Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-22T05:47:29.031123Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T05:51:08.879906Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact2
  • verified fuzzy40
  • unresolved35
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d3bad177-8d16-4566-a713-eac1b4519d0a · outbound

This paper cites Smart sustainable cities of the future: An extensive interdisciplinary literature review.Sustainable cities and society, 31:183–212, 2017.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Smart sustainable cities of the future: An extensive interdisciplinary literature review.Sustainable cities and society, 31:183–212, 2017

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:30.053205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:12.046737Z digest=sha256:97c373444531be5fb4d200e1195911b9c734031d84778200290e9109e579b6ff

Observation 9885d6b5-dcee-466e-a819-85deb337d722 · outbound

This paper cites Box and jenkins: time series analysis, forecasting and control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Box and jenkins: time series analysis, forecasting and control

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.739759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:12.109721Z digest=sha256:f621b61ad9ea422feb2f64c9bcb6d97af6ee3737483fcb9e3b19190ecd61317a

Observation bf73a06d-d4d3-45e8-a8e0-c00a25460d23 · outbound

This paper cites TEMPO: Prompt-based generative pre-trained transformer for time series forecasting.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents TEMPO: Prompt-based generative pre-trained transformer for time series forecasting

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.470939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:12.210665Z digest=sha256:2b39643fae9ba310e393b3cb6a6cc0fe063aeffcfc7e9965cebdfe67fbf1c3ba

Observation c255d701-a394-4e5c-be7d-30674f53e250 · outbound

This paper cites Agentboard: An analytical evaluation board of multi-turn llm agents.Advances in Neural Information Processing Systems, 37:74325–74362, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Agentboard: An analytical evaluation board of multi-turn llm agents.Advances in Neural Information Processing Systems, 37:74325–74362, 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.217389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:12.317306Z digest=sha256:bd7f5090a8da532bab3bb375cb7af7c5f7152747ceb259ecf41bfbd0da86cba8

Observation ad43525f-18d8-4cd5-943f-5c15873e9543 · outbound

This paper cites Graphwiz: An instruction-following language model for graph computational problems.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Graphwiz: An instruction-following language model for graph computational problems

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.443472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.443472Z digest=sha256:efdd6ba7a48173445436af3d5835ffded10f3d1f48659721e3793562ed9ea314

Observation 6c75bcdb-df9a-46e9-9849-480d16ce3eba · outbound

This paper cites Deeptransport: Learning spatial-temporal dependency for traffic condition forecasting.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Deeptransport: Learning spatial-temporal dependency for traffic condition forecasting

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.987240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:12.538621Z digest=sha256:b544107f9953ecee7f502a1d5d002a5f003749212326e3f606e811ba3c8d120f

Observation a3e9f2f7-2aa0-4acc-8083-40f2d7188618 · outbound

This paper cites TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.611434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.611434Z digest=sha256:41bd588d98936c84788ec218156993ee59176437e5b2e3e254875f9d414ca0ca

Observation c8fe3724-6bf0-4dbf-8be1-42ebe63c9e00 · outbound

This paper cites On the evolution of random graphs.Publ.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents On the evolution of random graphs.Publ

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.735372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:12.700990Z digest=sha256:dc0a685df63f476f4dadf801e4e9ba8d5dd252a152e50ef87abc8e1ea41d94b9

Observation a59d186c-7fc3-48c5-9aa1-90e8e9f54f9b · outbound

This paper cites Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.799685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.799685Z digest=sha256:2cb2a9a201169492f658727ea68ab42ad6213404c5d6ebd68b7e52031d67a9b4

Observation 90ac15c2-34fb-41be-9f7e-e7ba2f5f5eb2 · outbound

This paper cites CityGPT: Empowering Urban Spatial Cognition of Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CityGPT: Empowering Urban Spatial Cognition of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.923658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.923658Z digest=sha256:8a678976eaff1782e6354cf4e4e18d8450f6f69e03b1ab802ed67483b94183e2

Observation b05f7949-d8c9-498a-9b12-096d40ba262b · outbound

This paper cites CityBench: Evaluating the Capabilities of Large Language Models for Urban Tasks.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CityBench: Evaluating the Capabilities of Large Language Models for Urban Tasks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.027010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.027010Z digest=sha256:6453d570262fc2d949f86ec556b5362a8eb0c7af5d76a4f976b698a38736b3d7

Observation 56f07dc6-d6a3-4acb-8701-9bc858e7f096 · outbound

This paper cites Pygad: An intuitive genetic algorithm python library.Multimedia tools and applications, 83(20):58029–58042, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Pygad: An intuitive genetic algorithm python library.Multimedia tools and applications, 83(20):58029–58042, 2024

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.524138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:13.132409Z digest=sha256:896d0f573a38a44fb7f0d023294dff10c91b9f09eaece81f04ef2b63235e22ad

Observation 301f1204-0906-4602-ba2f-615260bd1b80 · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.204895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.204895Z digest=sha256:6149ea7486e40ebf7c6f2fa142485c2e47fc46ef37719d237d3f676d23d3b8f3

Observation b1aca464-8f36-4df2-b1a8-765071066833 · outbound

This paper cites The Llama 3 Herd of Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents The Llama 3 Herd of Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.286626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.286626Z digest=sha256:2ab5d6c9f32a42b3b23dc47925f70dee89ad5e9c3e8f46aca6b363dd88f3330e

Observation d35c513c-5a74-454e-882a-19d19dea1460 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.387728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.387728Z digest=sha256:6d3b2c2db474dddf4273e6c1e77e95303d0f6593044d4aa25971f3b2a502501b

Observation 88da6acb-25fb-4d14-bf4a-95722d3d2e20 · outbound

This paper cites Language Models Represent Space and Time.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Language Models Represent Space and Time

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.477756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.477756Z digest=sha256:e68c2e43dba754af3e70d64a30f59ae9c02c9842f586ca6d481c50d9f129cefb

Observation 06b0d6f0-dd10-4535-9cb6-71fdfe62f4c8 · outbound

This paper cites The scoot on-line traffic signal optimisation technique.Traffic Engineering & Control, 23(4), 1982.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents The scoot on-line traffic signal optimisation technique.Traffic Engineering & Control, 23(4), 1982

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.285556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:13.575875Z digest=sha256:445fb89c3d53fc2e351f5d725e2b5f631ba4610760edb0bae01dab1f74203d86

Observation 50b9e666-f72a-4045-93a3-4acfb89323e9 · outbound

This paper cites GPT-4o System Card.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents GPT-4o System Card

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.695425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.695425Z digest=sha256:9869bf65d7138573c33f706cd98fbc6551fe953a1843b921551a564e803e5f9b

Observation 49476d7d-4ca7-496c-ad4c-af8166894df1 · outbound

This paper cites OpenAI o1 System Card.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents OpenAI o1 System Card

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.795554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.795554Z digest=sha256:9d63f3763ecc3541cbe9bb5989f06363799fe0989dc854b1f8b63c0479c6d15f

Observation e88a2df5-49c1-4ece-898c-30b2188c82e7 · outbound

This paper cites Towards mitigating LLM hallucination via self reflection.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards mitigating LLM hallucination via self reflection

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.899032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.899032Z digest=sha256:36c14a57658cf811157efcccc22bdfee4dca30e74842e065c1bc1f889ed5dc9c

Observation c9424195-67a4-4265-9689-cd4d5e9eb84f · outbound

This paper cites Llmlight: Large language models as traffic signal control agents.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Llmlight: Large language models as traffic signal control agents

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.013525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:14.009911Z digest=sha256:27ffe68ef6ff243c785c84a90d514188ca9967f1fc4a96476085b713ce9441e9

Observation ecf904e5-b7a4-4121-929f-8c83c7c34bcc · outbound

This paper cites Reframing Spatial Reasoning Evaluation in Language Models: A Real-World Simulation Benchmark for Qualitative Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Reframing Spatial Reasoning Evaluation in Language Models: A Real-World Simulation Benchmark for Qualitative Reasoning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:20.582052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:14.089548Z digest=sha256:c0bee44075d46aff3b860f269129b0023216910cc909361c61abf8c7f8d9f903

Observation ae512c5d-51a9-4e84-99e8-024664a580cc · outbound

This paper cites Repetition in repetition out: Towards understanding neural text degeneration from the data perspective.Advances in Neural Information Processing Systems, 36:72888–72903, 2023.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Repetition in repetition out: Towards understanding neural text degeneration from the data perspective.Advances in Neural Information Processing Systems, 36:72888–72903, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.169851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.169851Z digest=sha256:0b0bc7b1faf27d7d689e1ae1b23dd5521f6318952cc41d15e8cc64d4459f754a

Observation cde91bcb-a5c6-4f1a-9bdf-dc48f39c0cee · outbound

This paper cites Towards alleviating traffic congestion: Optimal route planning for massive-scale trips.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards alleviating traffic congestion: Optimal route planning for massive-scale trips

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.808459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:14.238835Z digest=sha256:28b60a080279cd6a40aa5b1bf7b8811a8ea9e517231be9d1d6e6b0f8f908be09

Observation bf669141-44b7-4e7e-ab00-e559ef889780 · outbound

This paper cites STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.331330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.331330Z digest=sha256:a8ed83d09afb0b4ef07161e002dd2ce951df3ca66db574ab283947b966ca7b55

Observation e5fbb148-bc8a-42b0-a3c2-d058aaf4748a · outbound

This paper cites Urbangpt: Spatio-temporal large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Urbangpt: Spatio-temporal large language models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.625467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:14.400820Z digest=sha256:e31099faa16f37b14eac4b601b344640ffb9f51a23c2ef3f30652dd83db723fd

Observation 10788f50-7666-48cf-b82f-6b5c41040519 · outbound

This paper cites Timecma: Towards llm-empowered multivariate time series forecasting via cross-modality alignment.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Timecma: Towards llm-empowered multivariate time series forecasting via cross-modality alignment

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.509291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.509291Z digest=sha256:281d3629ba756e34138b57b1aeac8895c611ada42e16496a57b187c70417cfd8

Observation a0ec181d-3d0c-40c4-a3e0-f98ee4e6e9a3 · outbound

This paper cites Knowledge-infused contrastive learning for urban imagery-based socioeconomic prediction.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Knowledge-infused contrastive learning for urban imagery-based socioeconomic prediction

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.411860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:14.582881Z digest=sha256:460c31b2ac2b0723bc40a8bc570618d09b4a90290b2920c8b13ec623a7ddfd86

Observation 422597ef-7eb4-488b-b0b3-4004cea8b6ad · outbound

This paper cites Simulation of urban mobility (sumo), February 4 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Simulation of urban mobility (sumo), February 4 2025

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.198446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:14.661472Z digest=sha256:ff8d685c491549a3292dddfbb1fb68e6a9b52d824dd4e1a1186c77d6c97a53ab

Observation e3b83036-12e7-4dff-ab0a-8f649ede33e4 · outbound

This paper cites Scats, sydney co-ordinated adaptive traffic system: A traffic responsive method of controlling urban traffic.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Scats, sydney co-ordinated adaptive traffic system: A traffic responsive method of controlling urban traffic

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.016511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:14.750231Z digest=sha256:a73c494ea6a55d7c712e946f4869208467201344f475f411b916f3bccc29ec50

Observation f747e7a6-3ced-4808-ac9c-b3c9e30c2a34 · outbound

This paper cites SpartQA: : A Textual Question Answering Benchmark for Spatial Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents SpartQA: : A Textual Question Answering Benchmark for Spatial Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.883698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.883698Z digest=sha256:208d1388fcda406242321bc8d0da2a729789a8faa8d24d840f9e6fe2cee64511

Observation d259570f-b2c7-4498-b981-7de7b44aa08a · outbound

This paper cites Transfer Learning with Synthetic Corpora for Spatial Role Labeling and Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Transfer Learning with Synthetic Corpora for Spatial Role Labeling and Reasoning

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:20.222720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:14.953677Z digest=sha256:f7650355e33820ba5c6d647c95edd22ca00645a1a48ce18a84bbf68c49b4a203

Observation 4828f170-3d74-4f7d-b4a9-e47b21946184 · outbound

This paper cites Towards understanding the spatial literacy of chatgpt.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards understanding the spatial literacy of chatgpt

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.760392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:15.053874Z digest=sha256:02666d610526ea674aff751bc5ff2bccf843641be114f319bc9e3cd9728e0df3

Observation 14f51320-84f5-495e-a4e6-d249685ad8dc · outbound

This paper cites Tlc trip record data, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Tlc trip record data, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.551616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:15.128952Z digest=sha256:837fc90fcb6e8a386f31f30a2624d4ca2c9d860c8cda17f5f01ca6b02863bf84

Observation 4cc97432-e7f0-474a-9398-f6319ecbaac2 · outbound

This paper cites Dima: An llm-powered ride-hailing assistant at didi.arXiv preprint arXiv:2503.04768, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Dima: An llm-powered ride-hailing assistant at didi.arXiv preprint arXiv:2503.04768, 2025

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.195897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.195897Z digest=sha256:18d8b27a5b0a15102ae3adcf6a3bd158791227313c634f01d6d61e0f4f443012

Observation dec75734-d12d-4434-b73d-9313191c4b88 · outbound

This paper cites UrbanKGent: A Unified Large Language Model Agent Framework for Urban Knowledge Graph Construction.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents UrbanKGent: A Unified Large Language Model Agent Framework for Urban Knowledge Graph Construction

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.268020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.268020Z digest=sha256:519dd9269e547cab22f5a27c4869d1214d9eaa7eb989ab8ea69c1a07d935fb2c

Observation cb413b5a-1f9e-4f3c-9391-efd7c0e6903d · outbound

This paper cites Openstreetmap planet data, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Openstreetmap planet data, 2025

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.285455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:15.353419Z digest=sha256:cb4d8d08f8ddf4e43433f66acbc86d2c430efc11e762acdaaf4168d2db8c79b6

Observation 6ad461f8-920f-470c-84de-08b57ba8d8c7 · outbound

This paper cites Self-Reflection in LLM Agents: Effects on Problem-Solving Performance.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Self-Reflection in LLM Agents: Effects on Problem-Solving Performance

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.428482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.428482Z digest=sha256:6656dfc2b860731dae0e9971705ea85d95a690f5b3e4389c6b0b41c4fcadd46b

Observation 723dd069-13af-447a-a128-06ea71470315 · outbound

This paper cites Sparc and sparp: Spatial reasoning char- acterization and path generation for understanding spatial reasoning capability of large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Sparc and sparp: Spatial reasoning char- acterization and path generation for understanding spatial reasoning capability of large language models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.080108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:15.517387Z digest=sha256:ad0bbf1b902d7633d1b81230543d3dfac447cbba8fb27225801a4fe4e518d841

Observation a69353f3-9ea3-4fe3-bebd-270f11f1adc1 · outbound

This paper cites Proximal Policy Optimization Algorithms.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Proximal Policy Optimization Algorithms

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.585405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.585405Z digest=sha256:a2af3c50dbf096f76070519e68b56ece06c27d44ba2f11552f81a26ac8fef9e2

Observation 4daf5592-2ac2-4e64-8a55-200258dc8fc8 · outbound

This paper cites Stepgame: A new benchmark for robust multi- hop spatial reasoning in texts.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Stepgame: A new benchmark for robust multi- hop spatial reasoning in texts

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.851533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:15.674222Z digest=sha256:9d5915d34a0b2978ca204269662d5677702b27db987bd1558e5f802139a0ce22

Observation c251b9f7-f9ca-43fe-b35f-5a38e669c87a · outbound

This paper cites Towards Benchmarking and Improving the Temporal Reasoning Capability of Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards Benchmarking and Improving the Temporal Reasoning Capability of Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.798080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.798080Z digest=sha256:6a0be7329ae169c94ea66fd1aa670ee7942540f6c0b86274dea3b996550f204d

Observation 0f1326f0-41f0-49e4-8514-d1ebba27bd56 · outbound

This paper cites Cityflow: A city-scale benchmark for multi-target multi-camera vehicle tracking and re-identification.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Cityflow: A city-scale benchmark for multi-target multi-camera vehicle tracking and re-identification

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.623904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:15.887239Z digest=sha256:ee5d2c45974d04e561f95e76073668c837a98d84ee52f1c3968beb1902fcfcdf

Observation fa4bfccf-fd74-4b0e-a578-54a27441a953 · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Qwq-32b: Embracing the power of reinforcement learning, 2025

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.303061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:15.978560Z digest=sha256:b5ee5fd1ca11c3b29b14699f6f4e0066d40381251ed6a70c827ce1cfa8f4bf59

Observation f4af302b-8d6b-4e6b-8601-3ad39f849674 · outbound

This paper cites Air quality prediction with physics-guided dual neural odes in open systems.ICLR, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Air quality prediction with physics-guided dual neural odes in open systems.ICLR, 2025

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.077903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:16.075275Z digest=sha256:7db7adab940b0343d9f8e0eacb76202a92b0d2a570e90d80a113716f29d479f8

Observation 2552d402-3c96-4f5b-ad90-9fbdb48000f7 · outbound

This paper cites Applications of artificial intelligence and machine learning in smart cities.Computer Communications, 154:313– 323, 2020.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Applications of artificial intelligence and machine learning in smart cities.Computer Communications, 154:313– 323, 2020

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:24.803848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:16.158597Z digest=sha256:581d1f46abb95e8ce4afb0aa4401d0c4d51bd35889f390b3d0713ff71a3fbce0

Observation af1e7ce4-b8e8-4356-a1e3-b3f04352730e · outbound

This paper cites Robust extrema features for time-series data analysis.IEEE transactions on pattern analysis and machine intelligence, 35(6):1464–1479, 2012.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Robust extrema features for time-series data analysis.IEEE transactions on pattern analysis and machine intelligence, 35(6):1464–1479, 2012

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:24.526996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:16.246214Z digest=sha256:e98f0ea187235ed46b0bf595b22885d76ff6530938cfa990e3cba3fc15d5c15e

Observation e2e9300c-94fa-4f4f-b587-ac87034a2781 · outbound

This paper cites Reinforcement learning-based placement of charging stations in urban road networks.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Reinforcement learning-based placement of charging stations in urban road networks

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:24.223000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:16.325026Z digest=sha256:b46ce75bb9de3c9a300206d1e0c0954c42f4620416d922f95151821a034bbdef

Observation 19ad2d6f-34f2-4d09-9f88-daae5c60538b · outbound

This paper cites A survey on large language model based autonomous agents.Frontiers of Computer Science, 18(6):186345, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents A survey on large language model based autonomous agents.Frontiers of Computer Science, 18(6):186345, 2024

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.401542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.401542Z digest=sha256:87429319f742dbd808d8de406f2f2da64d09ccecdb1ea96e938b4c651d6646ae

Observation 90c123ca-ef52-4e9c-982c-e2b79595fb2c · outbound

This paper cites Global gridded gdp data set consistent with the shared socioeco- nomic pathways.Scientific data, 9(1):221, 2022.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Global gridded gdp data set consistent with the shared socioeco- nomic pathways.Scientific data, 9(1):221, 2022

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.916545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:16.480492Z digest=sha256:d4abed25d30469e6da960c6446ff6014a19cc1343eff5345bedaa10b1aa349b7

Observation bc30ab37-8be1-49a4-99da-2284ce663a53 · outbound

This paper cites Where Would I Go Next? Large Language Models as Human Mobility Predictors.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Where Would I Go Next? Large Language Models as Human Mobility Predictors

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.540853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.540853Z digest=sha256:bb4ff2cf69cc45fe9d990d9537d14e8df364bbd309988dfbe7c04ed4b9d64b3e

Observation 631cb0a6-a581-461d-bb63-c6f357f4aa12 · outbound

This paper cites From news to forecast: Integrating event analysis in llm-based time series forecasting with reflection.Advances in Neural Information Processing Systems, 37:58118–58153, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents From news to forecast: Integrating event analysis in llm-based time series forecasting with reflection.Advances in Neural Information Processing Systems, 37:58118–58153, 2024

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.622503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.622503Z digest=sha256:c8a40d9577a63feb5c4c2f92fcabe04aac064b2163ab4c116d25954de5e3f35c

Observation d99b94fc-7b0c-45fb-994c-9ed5ba95a5ac · outbound

This paper cites Tram: Benchmarking temporal reasoning for large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Tram: Benchmarking temporal reasoning for large language models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.698546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:16.712157Z digest=sha256:70f8f424e96709b1231303a5720faf11197d96c98c8a00e0e8c425d98cd16a27

Observation 3d1eb627-a4e6-4049-aa82-7a20f09596ef · outbound

This paper cites Colight: Learning network-level cooperation for traffic signal control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Colight: Learning network-level cooperation for traffic signal control

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.475729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:16.786439Z digest=sha256:9e1b642c1ffe71600db3560bd8ec85252042ef77f95951a04abda129e23fcd6b

Observation 92fe0cfb-22d7-44ae-9710-dde0885ad2fb · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Chain-of-thought prompting elicits reasoning in large language models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.887823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.887823Z digest=sha256:a05af73d75c8486d727e09fa6b372ee40907896b7007dffc0b9adffb313a8ff4

Observation 505bf480-0890-478b-b977-1c17a5ad767a · outbound

This paper cites Coverage location models: alternatives, approximation, and uncertainty.International Regional Science Review, 39(1):48–76, 2016.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Coverage location models: alternatives, approximation, and uncertainty.International Regional Science Review, 39(1):48–76, 2016

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.247621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:16.946618Z digest=sha256:4ae894e647823f21900db5e83edca1a7522504e7bb3b16cb79d6224f63dd54c5

Observation 0096260e-70f7-423c-baae-ebacf7b67b57 · outbound

This paper cites Worldpop hub, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Worldpop hub, 2025

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.019759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:17.010800Z digest=sha256:a6fb176fcc867a35d2cd013ee7a60bf66cd9695481c27b230b13103e4102e849

Observation 65fc1a15-7390-4d78-90d8-35e328c4f2ab · outbound

This paper cites Large language models can learn temporal reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Large language models can learn temporal reasoning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.045214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.045214Z digest=sha256:3d4dcd2dd7fd062950f89097461bb0291a8ae60d786b98182918377655092f37

Observation 63d0c6b9-0757-42fd-99dc-eea7f6aee440 · outbound

This paper cites Evaluating Spatial Understanding of Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Evaluating Spatial Understanding of Large Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.109005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.109005Z digest=sha256:c27de56f11cfe05db1bcc3172ae15088fdaf4e35a9df6fa61614bd1852a61c87

Observation 50dc80f6-64a3-494a-bcaa-7b0934374277 · outbound

This paper cites Qwen2.5 Technical Report.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Qwen2.5 Technical Report

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.170094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.170094Z digest=sha256:f267394785eeb4d42797fc09f329e53646ff5c7af5c0049884e86ba7f3aa9182

Observation a7da1f14-7825-47e0-8d35-858f2469cd3b · outbound

This paper cites Foursquare dataset.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Foursquare dataset

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.831190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:17.230652Z digest=sha256:bbd268e48c23d7435015a46328c997df685d8244c3501c8ecdb5525420839b16

Observation 0947063f-e531-4e33-b803-ac5d06090025 · outbound

This paper cites Unist: A prompt-empowered universal model for urban spatio-temporal prediction.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Unist: A prompt-empowered universal model for urban spatio-temporal prediction

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.582495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:17.315597Z digest=sha256:bfb89a3419335c8af65089b550cd3f26da6401fa46579689dd4f33b9bbfc6873

Observation 534d8ca3-9355-478c-be5e-d2617d2ce9cb · outbound

This paper cites CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal Control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal Control

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.318918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.318918Z digest=sha256:01afef30c0a3ca1d206e1c096dc900a3f6116a8bfe3dc2ed6045a3fa7da5bc96

Observation a3a05314-e66e-49aa-a2ab-2a28e89ac2ae · outbound

This paper cites AgentTuning: Enabling Generalized Agent Abilities for LLMs.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents AgentTuning: Enabling Generalized Agent Abilities for LLMs

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.428562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.428562Z digest=sha256:275fafd4392acb11f4092f138e91481a6d34bf50012fa5c47dcc490cd3c89c55

Observation 48950252-710c-4907-94f0-437c9c7b95d3 · outbound

This paper cites Open3dvqa: A benchmark for comprehensive spatial reasoning with multimodal large language model in open space.arXiv preprint arXiv:2503.11094, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Open3dvqa: A benchmark for comprehensive spatial reasoning with multimodal large language model in open space.arXiv preprint arXiv:2503.11094, 2025

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.608625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.608625Z digest=sha256:c5cfe7f5286ede6e27158e9ce5b29ece3bb655ad630532292d8ddbcdb0bb1a84

Observation 2d326e74-c59f-4b3e-8183-5e397f5281bb · outbound

This paper cites Reinforcement learning for traffic signal control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Reinforcement learning for traffic signal control

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.404837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:17.786637Z digest=sha256:937e3984cd76a00a0d8270d44e4a962237e5c9cae0037abed360b60bc62170aa

Observation fabf270e-123f-4fc7-9ce1-332b15430add · outbound

This paper cites Urbanvideo-bench: Benchmarking vision- language models on embodied intelligence with video data in urban spaces.arXiv preprint arXiv:2503.06157, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Urbanvideo-bench: Benchmarking vision- language models on embodied intelligence with video data in urban spaces.arXiv preprint arXiv:2503.06157, 2025

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.931121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.931121Z digest=sha256:3ee9f8ec50492afe488eb81d61545da983f9826f5f5c0c276409f5975b32a144

Observation 0d75fb69-a602-4e7e-8442-29c2235cce72 · outbound

This paper cites Where to go next: A spatio-temporal gated network for next poi recommendation.IEEE Transactions on Knowledge and Data Engineering, 34(5):2512–2524, 2020.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Where to go next: A spatio-temporal gated network for next poi recommendation.IEEE Transactions on Knowledge and Data Engineering, 34(5):2512–2524, 2020

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.238763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:18.112245Z digest=sha256:5b156d7f69bf0f765d5f003bf4b8c5a0ea05181db9709a2d56d75e7b54fd70aa

Observation 41cab77f-526e-4780-82bd-3ba195d55daf · outbound

This paper cites CityEQA: A Hierarchical LLM Agent on Embodied Question Answering Benchmark in City Space.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CityEQA: A Hierarchical LLM Agent on Embodied Question Answering Benchmark in City Space

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:18.279741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:18.279741Z digest=sha256:09fa9fa2f9215e35b1b38b47dfc53aa38f8c8791a93826291ac5795b62f94825

Observation 6832a9e1-703e-4f49-b092-d7d0f99a8b58 · outbound

This paper cites Llamafactory: Unified efficient fine-tuning of 100+ language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Llamafactory: Unified efficient fine-tuning of 100+ language models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:18.409485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:18.409485Z digest=sha256:fd0506f05f77bb50212fa0b6ce05e65311fa3983e8f071f30f8c26bb5cb25c00

Observation 19360793-b3c5-4839-83e3-15ec9d74624c · outbound

This paper cites Spatial planning of urban communities via deep reinforcement learning.Nature Computational Science, 3(9):748– 762, 2023.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Spatial planning of urban communities via deep reinforcement learning.Nature Computational Science, 3(9):748– 762, 2023

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.005609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:18.533561Z digest=sha256:dda76ef102c593edbdaeaf9286d674ac8f79ee832abdcb043c354392b975d780

Observation be5f3e12-424a-4851-9e95-73b5a71994bf · outbound

This paper cites UrbanPlanBench: A Comprehensive Urban Planning Benchmark for Evaluating Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents UrbanPlanBench: A Comprehensive Urban Planning Benchmark for Evaluating Large Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:18.662366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:18.662366Z digest=sha256:7ad76410de72cda5525c44b6313d1d7875c48e4365bf69468a0e59b75b963455

Observation c5dd61e9-0ffb-464b-97af-6e169aee06af · outbound

This paper cites Road planning for slums via deep reinforcement learning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Road planning for slums via deep reinforcement learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.838077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:18.808980Z digest=sha256:d87c9830f0f110a97e5682ea8542ea8ec859cc9df235e80c2a4327eaf4ed5a62

Observation 61ce2295-ac49-4efe-a6f4-b1a2f38a06a2 · outbound

This paper cites going on a vacation.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents going on a vacation

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.632527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:18.941064Z digest=sha256:dac5fb47adf53c700c27642f38693ce123e196469f283d4812c7db957589bd9b

Observation 8b91d25d-f6ec-49d2-9dac-c2c543ed9e90 · outbound

This paper cites Large Language Model for Participatory Urban Planning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Large Language Model for Participatory Urban Planning

Reference 75

Resolution
malformed identifier
no resolver link, observed 2026-08-07T14:47:19.072276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:19.072276Z digest=sha256:c883d0ebc57f3805afff711957856736f32b2de5759d3825ad24e831d270ad94

Observation 25e769a7-7f59-4d94-8167-be3f432879db · outbound

This paper cites answer":.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents answer":

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.417259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:19.242745Z digest=sha256:112610cca6152d7cc8ae265aa62efb272d2e7f660b121d949c541af2a8ef7118

Observation 0a1940f9-9f5b-4dcf-a6ce-05b98748b6fb · outbound

This paper cites Miscellaneous Shop.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Miscellaneous Shop

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.155850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:19.378369Z digest=sha256:6b5774c8269caf4f675137d600fe1c48a95b82dbceeae3f88bb3a141a61c195d

Observation 682dc218-38ba-4888-9e43-f4757dff36fb · outbound

This paper cites answer":.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents answer":

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:20.945046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:19.524779Z digest=sha256:76095d2a68866fcd637a649606eb602b8cc65333482f5a2a7f5a8933084c6d48

Pith citing papers

Observation 12238777-3eda-4da9-85cd-1e1a49d2e174 · inbound

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation cites this paper.

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:51:08.882918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T05:47:29.031123Z digest=sha256:896efbcf24557599ad2adb7b8e3eb5bf36ed69f694902d066077b0c00e5d00bd