Pith. sign in

Paper Citation Record · LEDGER

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents

As of 7 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2505.17572.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17572 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:47:19.524779Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-22T05:47:29.031123Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T05:51:08.879906Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact2
  • verified fuzzy40
  • unresolved35
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d3bad177-8d16-4566-a713-eac1b4519d0a · outbound

This paper cites Smart sustainable cities of the future: An extensive interdisciplinary literature review.Sustainable cities and society, 31:183–212, 2017.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Smart sustainable cities of the future: An extensive interdisciplinary literature review.Sustainable cities and society, 31:183–212, 2017

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:30.053205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:12.046737Z digest=sha256:26001aeca94810f0b8355022b9cef1af505f157d49ef3671247a76401757ed79

Observation 9885d6b5-dcee-466e-a819-85deb337d722 · outbound

This paper cites Box and jenkins: time series analysis, forecasting and control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Box and jenkins: time series analysis, forecasting and control

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.739759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:12.109721Z digest=sha256:8c633a18ecbb98d2f973cd1a9941c16bb12f676415815164c2429db51b0a65d4

Observation bf73a06d-d4d3-45e8-a8e0-c00a25460d23 · outbound

This paper cites TEMPO: Prompt-based generative pre-trained transformer for time series forecasting.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents TEMPO: Prompt-based generative pre-trained transformer for time series forecasting

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.470939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:12.210665Z digest=sha256:dca791b94b035c46f8fd2cd0c0899714a0a28c0f642b3c4e8b73f4cc15bf06fc

Observation c255d701-a394-4e5c-be7d-30674f53e250 · outbound

This paper cites Agentboard: An analytical evaluation board of multi-turn llm agents.Advances in Neural Information Processing Systems, 37:74325–74362, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Agentboard: An analytical evaluation board of multi-turn llm agents.Advances in Neural Information Processing Systems, 37:74325–74362, 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.217389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:12.317306Z digest=sha256:9ca72573d9e12fcacf937f03c566e0ca9c0046765fb4d237a09d1fb7c097bd15

Observation ad43525f-18d8-4cd5-943f-5c15873e9543 · outbound

This paper cites Graphwiz: An instruction-following language model for graph computational problems.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Graphwiz: An instruction-following language model for graph computational problems

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.443472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.443472Z digest=sha256:30adf413fb56c8411bded90a2f9126fb11e78055a313a2a42ad2ff4fd6b41412

Observation 6c75bcdb-df9a-46e9-9849-480d16ce3eba · outbound

This paper cites Deeptransport: Learning spatial-temporal dependency for traffic condition forecasting.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Deeptransport: Learning spatial-temporal dependency for traffic condition forecasting

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.987240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:12.538621Z digest=sha256:88b5400424c0386ff6390733f20088f8f26eb02d78f3a05b999839fe41c5b175

Observation a3e9f2f7-2aa0-4acc-8083-40f2d7188618 · outbound

This paper cites TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.611434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.611434Z digest=sha256:d277741c19e4fdb6ecc6058c27eafde1c8fd90ccf2ec117f1599bcb2df77af84

Observation c8fe3724-6bf0-4dbf-8be1-42ebe63c9e00 · outbound

This paper cites On the evolution of random graphs.Publ.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents On the evolution of random graphs.Publ

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.735372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:12.700990Z digest=sha256:d6267a97a10bdb9ce5232fa01c8045e001e113854f37141efd7c3084852f844e

Observation a59d186c-7fc3-48c5-9aa1-90e8e9f54f9b · outbound

This paper cites Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.799685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.799685Z digest=sha256:d62b4af4d3136a1b7ec0401f42493bad8f0838a7243bb471a8dc6346a17ca1e6

Observation 90ac15c2-34fb-41be-9f7e-e7ba2f5f5eb2 · outbound

This paper cites CityGPT: Empowering Urban Spatial Cognition of Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CityGPT: Empowering Urban Spatial Cognition of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.923658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.923658Z digest=sha256:d5b87684b71d6da68a6de2329b92e1b8d34ac96163f99c57f438c9816e66a801

Observation b05f7949-d8c9-498a-9b12-096d40ba262b · outbound

This paper cites CityBench: Evaluating the Capabilities of Large Language Models for Urban Tasks.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CityBench: Evaluating the Capabilities of Large Language Models for Urban Tasks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.027010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.027010Z digest=sha256:50716f4a4d26458d9d579df8f46dc52c9e1f0aa012e5350b2eaa4e263f860399

Observation 56f07dc6-d6a3-4acb-8701-9bc858e7f096 · outbound

This paper cites Pygad: An intuitive genetic algorithm python library.Multimedia tools and applications, 83(20):58029–58042, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Pygad: An intuitive genetic algorithm python library.Multimedia tools and applications, 83(20):58029–58042, 2024

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.524138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:13.132409Z digest=sha256:eac2f6c521a75c68efa80cbf17f79fb1dc728dbad08268802ac485660066eae7

Observation 301f1204-0906-4602-ba2f-615260bd1b80 · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.204895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.204895Z digest=sha256:36190b2dd2b778324c265409e0a8f282fd55ee157f80dd9f3153c82bd6338c14

Observation b1aca464-8f36-4df2-b1a8-765071066833 · outbound

This paper cites The Llama 3 Herd of Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents The Llama 3 Herd of Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.286626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.286626Z digest=sha256:ada77d8f1db9d80bf8dd5b86e986323e08a7dc0993c3795b26c6129c4456af4d

Observation d35c513c-5a74-454e-882a-19d19dea1460 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.387728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.387728Z digest=sha256:645a35fd1cce44f036725df4679f3bd1b205e871dfb5979e93cd3d9adeb2fbed

Observation 88da6acb-25fb-4d14-bf4a-95722d3d2e20 · outbound

This paper cites Language Models Represent Space and Time.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Language Models Represent Space and Time

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.477756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.477756Z digest=sha256:069d4e9cab8cf1640361fd4f307c503e42f551bf033459efe2d314062a20b218

Observation 06b0d6f0-dd10-4535-9cb6-71fdfe62f4c8 · outbound

This paper cites The scoot on-line traffic signal optimisation technique.Traffic Engineering & Control, 23(4), 1982.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents The scoot on-line traffic signal optimisation technique.Traffic Engineering & Control, 23(4), 1982

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.285556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:13.575875Z digest=sha256:ad52b33ea596f451ab655646e1f29bf178fd502ddaaeeab7fffd988ebcbf0b13

Observation 50b9e666-f72a-4045-93a3-4acfb89323e9 · outbound

This paper cites GPT-4o System Card.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents GPT-4o System Card

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.695425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.695425Z digest=sha256:a5cd2d7e4978fdaa19b9744141867a917201bc66a1f7baecfd02859c9254c043

Observation 49476d7d-4ca7-496c-ad4c-af8166894df1 · outbound

This paper cites OpenAI o1 System Card.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents OpenAI o1 System Card

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.795554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.795554Z digest=sha256:2d710b569d3310b465e43b3e13a11702616957bc4fc140774241cf156d2718b4

Observation e88a2df5-49c1-4ece-898c-30b2188c82e7 · outbound

This paper cites Towards mitigating LLM hallucination via self reflection.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards mitigating LLM hallucination via self reflection

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.899032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.899032Z digest=sha256:662250b272b1aeae86d7823245df7e538ac56f153725ed138d9fb399a81e7bad

Observation c9424195-67a4-4265-9689-cd4d5e9eb84f · outbound

This paper cites Llmlight: Large language models as traffic signal control agents.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Llmlight: Large language models as traffic signal control agents

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.013525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:14.009911Z digest=sha256:8973dc166ab487c867285873ff0db28f6a22bb677532b753759c84fec188d706

Observation ecf904e5-b7a4-4121-929f-8c83c7c34bcc · outbound

This paper cites Reframing Spatial Reasoning Evaluation in Language Models: A Real-World Simulation Benchmark for Qualitative Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Reframing Spatial Reasoning Evaluation in Language Models: A Real-World Simulation Benchmark for Qualitative Reasoning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:20.582052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:14.089548Z digest=sha256:93c19782cda3b4ce2d3e5244a289a1b3079bafbea336ec3c4ea26ddd02090980

Observation ae512c5d-51a9-4e84-99e8-024664a580cc · outbound

This paper cites Repetition in repetition out: Towards understanding neural text degeneration from the data perspective.Advances in Neural Information Processing Systems, 36:72888–72903, 2023.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Repetition in repetition out: Towards understanding neural text degeneration from the data perspective.Advances in Neural Information Processing Systems, 36:72888–72903, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.169851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.169851Z digest=sha256:f444eb7ff9996f59eccf12f59a1fea46f815a000c7112332450ba9a7b3a35e43

Observation cde91bcb-a5c6-4f1a-9bdf-dc48f39c0cee · outbound

This paper cites Towards alleviating traffic congestion: Optimal route planning for massive-scale trips.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards alleviating traffic congestion: Optimal route planning for massive-scale trips

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.808459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:14.238835Z digest=sha256:f48e052e980c048657d4e79488f7823cfa86a81749cc3175e503947d370242b7

Observation bf669141-44b7-4e7e-ab00-e559ef889780 · outbound

This paper cites STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.331330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.331330Z digest=sha256:f35a70e289b2ce3e985739af05783500f321f1e1026d2b562d18bf7a647956a5

Observation e5fbb148-bc8a-42b0-a3c2-d058aaf4748a · outbound

This paper cites Urbangpt: Spatio-temporal large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Urbangpt: Spatio-temporal large language models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.625467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:14.400820Z digest=sha256:b6dbefe8696f0fc78371b25ac98c6afbb4c2713b229344c71889abf473182015

Observation 10788f50-7666-48cf-b82f-6b5c41040519 · outbound

This paper cites Timecma: Towards llm-empowered multivariate time series forecasting via cross-modality alignment.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Timecma: Towards llm-empowered multivariate time series forecasting via cross-modality alignment

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.509291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.509291Z digest=sha256:1aee46fbfcf928654d13dbf0d8a1657d5177b988615a47ad9c4ceabb649709b0

Observation a0ec181d-3d0c-40c4-a3e0-f98ee4e6e9a3 · outbound

This paper cites Knowledge-infused contrastive learning for urban imagery-based socioeconomic prediction.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Knowledge-infused contrastive learning for urban imagery-based socioeconomic prediction

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.411860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:14.582881Z digest=sha256:29bdb8c2e3043c8b4e45e74d89aa809589708a44da70b68a70162cebe9fbb86a

Observation 422597ef-7eb4-488b-b0b3-4004cea8b6ad · outbound

This paper cites Simulation of urban mobility (sumo), February 4 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Simulation of urban mobility (sumo), February 4 2025

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.198446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:14.661472Z digest=sha256:b2f388370bce4554212b5ae9dc58f72cb0b326acee2317ca2dceadd51715befb

Observation e3b83036-12e7-4dff-ab0a-8f649ede33e4 · outbound

This paper cites Scats, sydney co-ordinated adaptive traffic system: A traffic responsive method of controlling urban traffic.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Scats, sydney co-ordinated adaptive traffic system: A traffic responsive method of controlling urban traffic

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.016511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:14.750231Z digest=sha256:c502c2bf863b41682b850fc02b7fc5bd7c78a38b89bbfcee4382d315d4197311

Observation f747e7a6-3ced-4808-ac9c-b3c9e30c2a34 · outbound

This paper cites SpartQA: : A Textual Question Answering Benchmark for Spatial Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents SpartQA: : A Textual Question Answering Benchmark for Spatial Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.883698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.883698Z digest=sha256:37f4b569fac20cde3d68b5ed572b6a334757c8a47fecdb4e1a337101d1b85cbb

Observation d259570f-b2c7-4498-b981-7de7b44aa08a · outbound

This paper cites Transfer Learning with Synthetic Corpora for Spatial Role Labeling and Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Transfer Learning with Synthetic Corpora for Spatial Role Labeling and Reasoning

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:20.222720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:14.953677Z digest=sha256:89aab34dfa160926569efa7749820dc0d3677c21b0d15a9c0bba336f076934f9

Observation 4828f170-3d74-4f7d-b4a9-e47b21946184 · outbound

This paper cites Towards understanding the spatial literacy of chatgpt.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards understanding the spatial literacy of chatgpt

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.760392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:15.053874Z digest=sha256:404fec5e138abc4ec483e30232fa41619298f613b8673f9330fd8dfe9965d558

Observation 14f51320-84f5-495e-a4e6-d249685ad8dc · outbound

This paper cites Tlc trip record data, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Tlc trip record data, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.551616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:15.128952Z digest=sha256:5809d233fe1044d8aa22fe4e45ae82585ef3b1c2525fd642331b0c03f0251aff

Observation 4cc97432-e7f0-474a-9398-f6319ecbaac2 · outbound

This paper cites Dima: An llm-powered ride-hailing assistant at didi.arXiv preprint arXiv:2503.04768, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Dima: An llm-powered ride-hailing assistant at didi.arXiv preprint arXiv:2503.04768, 2025

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.195897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.195897Z digest=sha256:f4921a2fb359250d1c588619316ac264b366a4829a63739252cd2126793359dd

Observation dec75734-d12d-4434-b73d-9313191c4b88 · outbound

This paper cites UrbanKGent: A Unified Large Language Model Agent Framework for Urban Knowledge Graph Construction.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents UrbanKGent: A Unified Large Language Model Agent Framework for Urban Knowledge Graph Construction

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.268020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.268020Z digest=sha256:c4ab3cb5d2b73217bb5a958bd79f9ebbc135bc9b1ade75254bcaf4bb7ce8e214

Observation cb413b5a-1f9e-4f3c-9391-efd7c0e6903d · outbound

This paper cites Openstreetmap planet data, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Openstreetmap planet data, 2025

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.285455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:15.353419Z digest=sha256:0b23df490cba049f1c4c6e85e33554fa5ae6fac6f0182b4cdd6e4df9064cec55

Observation 6ad461f8-920f-470c-84de-08b57ba8d8c7 · outbound

This paper cites Self-Reflection in LLM Agents: Effects on Problem-Solving Performance.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Self-Reflection in LLM Agents: Effects on Problem-Solving Performance

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.428482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.428482Z digest=sha256:0775f825482682211579c27b57aae144430a88a050a5bcb5ed7076aa13b95a8a

Observation 723dd069-13af-447a-a128-06ea71470315 · outbound

This paper cites Sparc and sparp: Spatial reasoning char- acterization and path generation for understanding spatial reasoning capability of large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Sparc and sparp: Spatial reasoning char- acterization and path generation for understanding spatial reasoning capability of large language models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.080108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:15.517387Z digest=sha256:50fbc80eef4367d7a853c9c82eea2c01919ce46b0b9d7570d18c2785b5ce970c

Observation a69353f3-9ea3-4fe3-bebd-270f11f1adc1 · outbound

This paper cites Proximal Policy Optimization Algorithms.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Proximal Policy Optimization Algorithms

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.585405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.585405Z digest=sha256:6f3aef03079bc8a51c8d607b4c8409b16741332b25b314dc3375b2ef7bcca319

Observation 4daf5592-2ac2-4e64-8a55-200258dc8fc8 · outbound

This paper cites Stepgame: A new benchmark for robust multi- hop spatial reasoning in texts.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Stepgame: A new benchmark for robust multi- hop spatial reasoning in texts

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.851533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:15.674222Z digest=sha256:ef84a70849e55004a19d5cbe5d349ef41b695e62fb836debc19ea689182cb213

Observation c251b9f7-f9ca-43fe-b35f-5a38e669c87a · outbound

This paper cites Towards Benchmarking and Improving the Temporal Reasoning Capability of Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards Benchmarking and Improving the Temporal Reasoning Capability of Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.798080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.798080Z digest=sha256:15becff924477fc44a2793bb87b8b9ed3deaac43530a3435b453a74037e73ef1

Observation 0f1326f0-41f0-49e4-8514-d1ebba27bd56 · outbound

This paper cites Cityflow: A city-scale benchmark for multi-target multi-camera vehicle tracking and re-identification.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Cityflow: A city-scale benchmark for multi-target multi-camera vehicle tracking and re-identification

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.623904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:15.887239Z digest=sha256:aa0022021b99d8d4c0950be243b810b9bee51842055d3afbdefa9c32e4874f46

Observation fa4bfccf-fd74-4b0e-a578-54a27441a953 · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Qwq-32b: Embracing the power of reinforcement learning, 2025

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.303061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:15.978560Z digest=sha256:548ef9536a0f5110e2289a5d8c4acc4138a63a5de37a73187d44970f90a61374

Observation f4af302b-8d6b-4e6b-8601-3ad39f849674 · outbound

This paper cites Air quality prediction with physics-guided dual neural odes in open systems.ICLR, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Air quality prediction with physics-guided dual neural odes in open systems.ICLR, 2025

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.077903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:16.075275Z digest=sha256:a6e66c2d939f184f3bc5b3dd040b1d032d0adf6e09a3ac1291f276df601760bb

Observation 2552d402-3c96-4f5b-ad90-9fbdb48000f7 · outbound

This paper cites Applications of artificial intelligence and machine learning in smart cities.Computer Communications, 154:313– 323, 2020.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Applications of artificial intelligence and machine learning in smart cities.Computer Communications, 154:313– 323, 2020

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:24.803848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:16.158597Z digest=sha256:601ee8555a7d22c6603957f808c4d67de79a7dd9a9be303ed16840ad16cf7fbb

Observation af1e7ce4-b8e8-4356-a1e3-b3f04352730e · outbound

This paper cites Robust extrema features for time-series data analysis.IEEE transactions on pattern analysis and machine intelligence, 35(6):1464–1479, 2012.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Robust extrema features for time-series data analysis.IEEE transactions on pattern analysis and machine intelligence, 35(6):1464–1479, 2012

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:24.526996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:16.246214Z digest=sha256:47f3543aaa540092d2aa7cd34918b89dcef0800eb868a2c92cd14bd5e3d27f51

Observation e2e9300c-94fa-4f4f-b587-ac87034a2781 · outbound

This paper cites Reinforcement learning-based placement of charging stations in urban road networks.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Reinforcement learning-based placement of charging stations in urban road networks

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:24.223000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:16.325026Z digest=sha256:cc7e18623c2033a05df0e749868915cc3171f2adeb814cffee96880596afb75b

Observation 19ad2d6f-34f2-4d09-9f88-daae5c60538b · outbound

This paper cites A survey on large language model based autonomous agents.Frontiers of Computer Science, 18(6):186345, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents A survey on large language model based autonomous agents.Frontiers of Computer Science, 18(6):186345, 2024

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.401542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.401542Z digest=sha256:97b8b8cc3155cd273bb02a4c8b4b63351d51a2d544446eda9c97a445bb4130d8

Observation 90c123ca-ef52-4e9c-982c-e2b79595fb2c · outbound

This paper cites Global gridded gdp data set consistent with the shared socioeco- nomic pathways.Scientific data, 9(1):221, 2022.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Global gridded gdp data set consistent with the shared socioeco- nomic pathways.Scientific data, 9(1):221, 2022

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.916545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:16.480492Z digest=sha256:49ea7f69024411ed89054c6aafa4fdfeb678fb774684d1dbbcb89941a8982703

Observation bc30ab37-8be1-49a4-99da-2284ce663a53 · outbound

This paper cites Where Would I Go Next? Large Language Models as Human Mobility Predictors.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Where Would I Go Next? Large Language Models as Human Mobility Predictors

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.540853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.540853Z digest=sha256:c297a63641c4bf55183763ebbf37d9e428d48bd74775be51dafb28297afef1b8

Observation 631cb0a6-a581-461d-bb63-c6f357f4aa12 · outbound

This paper cites From news to forecast: Integrating event analysis in llm-based time series forecasting with reflection.Advances in Neural Information Processing Systems, 37:58118–58153, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents From news to forecast: Integrating event analysis in llm-based time series forecasting with reflection.Advances in Neural Information Processing Systems, 37:58118–58153, 2024

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.622503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.622503Z digest=sha256:34556bbbb1f98aaeb84de6e0d476bf0df31b7bbe7b961c5b4a1fe40638e0af61

Observation d99b94fc-7b0c-45fb-994c-9ed5ba95a5ac · outbound

This paper cites Tram: Benchmarking temporal reasoning for large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Tram: Benchmarking temporal reasoning for large language models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.698546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:16.712157Z digest=sha256:9afb55c2c6fce29d39e70dd887715e14a4327626a7e0e1f34bbab3edd36575fa

Observation 3d1eb627-a4e6-4049-aa82-7a20f09596ef · outbound

This paper cites Colight: Learning network-level cooperation for traffic signal control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Colight: Learning network-level cooperation for traffic signal control

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.475729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:16.786439Z digest=sha256:3c08392595ea4f2b9ebb4e32350373e3972bc24763e36aaec9a355400c861a2d

Observation 92fe0cfb-22d7-44ae-9710-dde0885ad2fb · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Chain-of-thought prompting elicits reasoning in large language models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.887823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.887823Z digest=sha256:fdfd341132121cfc108eaef31c7f67daac52d561eabd1ca84ffdbb18e65b6c10

Observation 505bf480-0890-478b-b977-1c17a5ad767a · outbound

This paper cites Coverage location models: alternatives, approximation, and uncertainty.International Regional Science Review, 39(1):48–76, 2016.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Coverage location models: alternatives, approximation, and uncertainty.International Regional Science Review, 39(1):48–76, 2016

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.247621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:16.946618Z digest=sha256:a2e56023672be367815bbe4ddd2df7c6e0e2771c7ab61f72d9262e8ff7df7a52

Observation 0096260e-70f7-423c-baae-ebacf7b67b57 · outbound

This paper cites Worldpop hub, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Worldpop hub, 2025

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.019759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:17.010800Z digest=sha256:abb2c8b8753c1982e9df0d446fc2ed05fd552c7b373da205ac9d6e99cb038736

Observation 65fc1a15-7390-4d78-90d8-35e328c4f2ab · outbound

This paper cites Large language models can learn temporal reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Large language models can learn temporal reasoning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.045214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.045214Z digest=sha256:4f4ded453ee95a1d2921c1022bce8c29b8d577f23e136e60825f8e2f5e56da5e

Observation 63d0c6b9-0757-42fd-99dc-eea7f6aee440 · outbound

This paper cites Evaluating Spatial Understanding of Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Evaluating Spatial Understanding of Large Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.109005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.109005Z digest=sha256:56517cf706bf0795a8242a292401da78077fa5b484b7c56d2ef523bdeb41ca15

Observation 50dc80f6-64a3-494a-bcaa-7b0934374277 · outbound

This paper cites Qwen2.5 Technical Report.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Qwen2.5 Technical Report

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.170094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.170094Z digest=sha256:62c5d13c4d1566807179719d3a3a3e0aaa79d21c7084fbf734eeafafea93b54b

Observation a7da1f14-7825-47e0-8d35-858f2469cd3b · outbound

This paper cites Foursquare dataset.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Foursquare dataset

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.831190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:17.230652Z digest=sha256:70248c00b8da072752c81f91bc9953ffa5e1999dcc3c73c62ff8aa740bb737bb

Observation 0947063f-e531-4e33-b803-ac5d06090025 · outbound

This paper cites Unist: A prompt-empowered universal model for urban spatio-temporal prediction.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Unist: A prompt-empowered universal model for urban spatio-temporal prediction

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.582495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:17.315597Z digest=sha256:9edd9a7f5aeb1f06d1b46ef2dfc172c381c431d1b5f1274d1af0da25960cc0b0

Observation 534d8ca3-9355-478c-be5e-d2617d2ce9cb · outbound

This paper cites CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal Control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal Control

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.318918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.318918Z digest=sha256:5afe37f1fa94e06b90c04edd3de2404f5ec8fd366812fcb88f47a623d1700c1f

Observation a3a05314-e66e-49aa-a2ab-2a28e89ac2ae · outbound

This paper cites AgentTuning: Enabling Generalized Agent Abilities for LLMs.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents AgentTuning: Enabling Generalized Agent Abilities for LLMs

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.428562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.428562Z digest=sha256:5afd635f57fb3ef79f80f31289b8337cb0bb05e9444de8e07032e5180c63ae2a

Observation 48950252-710c-4907-94f0-437c9c7b95d3 · outbound

This paper cites Open3dvqa: A benchmark for comprehensive spatial reasoning with multimodal large language model in open space.arXiv preprint arXiv:2503.11094, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Open3dvqa: A benchmark for comprehensive spatial reasoning with multimodal large language model in open space.arXiv preprint arXiv:2503.11094, 2025

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.608625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.608625Z digest=sha256:cd260726017d98f47f765a91a872c9ad9d39c007d9f96df277b0434023abc76d

Observation 2d326e74-c59f-4b3e-8183-5e397f5281bb · outbound

This paper cites Reinforcement learning for traffic signal control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Reinforcement learning for traffic signal control

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.404837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:17.786637Z digest=sha256:9d949be7a9cf3101660bc30d93aa13f9b7f0d48f7e9db7429d0177059cf9df84

Observation fabf270e-123f-4fc7-9ce1-332b15430add · outbound

This paper cites Urbanvideo-bench: Benchmarking vision- language models on embodied intelligence with video data in urban spaces.arXiv preprint arXiv:2503.06157, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Urbanvideo-bench: Benchmarking vision- language models on embodied intelligence with video data in urban spaces.arXiv preprint arXiv:2503.06157, 2025

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.931121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.931121Z digest=sha256:5bf29062570a599a8fbdeecf4d36b9928ce7819cec64b34e367d1684848b0206

Observation 0d75fb69-a602-4e7e-8442-29c2235cce72 · outbound

This paper cites Where to go next: A spatio-temporal gated network for next poi recommendation.IEEE Transactions on Knowledge and Data Engineering, 34(5):2512–2524, 2020.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Where to go next: A spatio-temporal gated network for next poi recommendation.IEEE Transactions on Knowledge and Data Engineering, 34(5):2512–2524, 2020

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.238763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:18.112245Z digest=sha256:3b3bef78269c490b1409056390df26685b1748888c2ef79d8283a10d7b0a10f8

Observation 41cab77f-526e-4780-82bd-3ba195d55daf · outbound

This paper cites CityEQA: A Hierarchical LLM Agent on Embodied Question Answering Benchmark in City Space.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CityEQA: A Hierarchical LLM Agent on Embodied Question Answering Benchmark in City Space

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:18.279741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:18.279741Z digest=sha256:d74065c97dd274a20d78a5320a2058b31740dd591a80501119d100d3b19d68c7

Observation 6832a9e1-703e-4f49-b092-d7d0f99a8b58 · outbound

This paper cites Llamafactory: Unified efficient fine-tuning of 100+ language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Llamafactory: Unified efficient fine-tuning of 100+ language models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:18.409485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:18.409485Z digest=sha256:12cdbb899ceb04ae3ef612a323cc67efade099d8df84ed692adf52875ba2258f

Observation 19360793-b3c5-4839-83e3-15ec9d74624c · outbound

This paper cites Spatial planning of urban communities via deep reinforcement learning.Nature Computational Science, 3(9):748– 762, 2023.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Spatial planning of urban communities via deep reinforcement learning.Nature Computational Science, 3(9):748– 762, 2023

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.005609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:18.533561Z digest=sha256:04bd305f3fdd4397607f9f78864012b1c2425fad45b7c11f86e44013a6d05d2b

Observation be5f3e12-424a-4851-9e95-73b5a71994bf · outbound

This paper cites UrbanPlanBench: A Comprehensive Urban Planning Benchmark for Evaluating Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents UrbanPlanBench: A Comprehensive Urban Planning Benchmark for Evaluating Large Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:18.662366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:18.662366Z digest=sha256:44ca7ba02860040740c354d69ca860e533e2b62e277cef247b761269eba5d563

Observation c5dd61e9-0ffb-464b-97af-6e169aee06af · outbound

This paper cites Road planning for slums via deep reinforcement learning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Road planning for slums via deep reinforcement learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.838077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:18.808980Z digest=sha256:dd2be4cdb58a52570deee3bb7718cf137054796e8a3342ba7b1c5ca5673e4ecd

Observation 61ce2295-ac49-4efe-a6f4-b1a2f38a06a2 · outbound

This paper cites going on a vacation.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents going on a vacation

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.632527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:18.941064Z digest=sha256:fdc2a7925b7de66311b3257126b9e220fbf38ac957c2a5ec14f5a80521fa2308

Observation 8b91d25d-f6ec-49d2-9dac-c2c543ed9e90 · outbound

This paper cites Large Language Model for Participatory Urban Planning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Large Language Model for Participatory Urban Planning

Reference 75

Resolution
malformed identifier
no resolver link, observed 2026-08-07T14:47:19.072276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:19.072276Z digest=sha256:24443c1a37c485be4dd1cbde0d71a260da2c8571ab5c0a5c23543b1b648a7396

Observation 25e769a7-7f59-4d94-8167-be3f432879db · outbound

This paper cites answer":.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents answer":

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.417259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:19.242745Z digest=sha256:0c9b6875e69c32afb80af6c3afd04b996f403a7e0783aead4e74533f2f312ed2

Observation 0a1940f9-9f5b-4dcf-a6ce-05b98748b6fb · outbound

This paper cites Miscellaneous Shop.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Miscellaneous Shop

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.155850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:19.378369Z digest=sha256:4480a907c0d21ba75222531c636beaa633e4fc40871af5b2970e084d012e3d34

Observation 682dc218-38ba-4888-9e43-f4757dff36fb · outbound

This paper cites answer":.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents answer":

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:20.945046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:47:19.524779Z digest=sha256:baff561fab98a946a14a576c090e67a36da221d8e0e0640381bafe696f95952a

Pith citing papers

Observation 12238777-3eda-4da9-85cd-1e1a49d2e174 · inbound

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation cites this paper.

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:51:08.882918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T05:47:29.031123Z digest=sha256:b5900acd845ead79a07a81703976c5bbc852d5b6ebb3bf05cf6b0e7de088c1f3