Pith. sign in

Paper Citation Record · LEDGER

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents

As of 8 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2505.17572.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17572 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:47:19.524779Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-22T05:47:29.031123Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T05:51:08.879906Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact2
  • verified fuzzy40
  • unresolved35
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d3bad177-8d16-4566-a713-eac1b4519d0a · outbound

This paper cites Smart sustainable cities of the future: An extensive interdisciplinary literature review.Sustainable cities and society, 31:183–212, 2017.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Smart sustainable cities of the future: An extensive interdisciplinary literature review.Sustainable cities and society, 31:183–212, 2017

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:30.053205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:12.046737Z digest=sha256:b4cc1f4ed40df19ef1d06467be41641f33aa0a372341af20ec24e77532390665

Observation 9885d6b5-dcee-466e-a819-85deb337d722 · outbound

This paper cites Box and jenkins: time series analysis, forecasting and control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Box and jenkins: time series analysis, forecasting and control

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.739759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:12.109721Z digest=sha256:97a486daa5868657599af053dfce6a6b67769eb2f716f77aa87080765eeccfbb

Observation bf73a06d-d4d3-45e8-a8e0-c00a25460d23 · outbound

This paper cites TEMPO: Prompt-based generative pre-trained transformer for time series forecasting.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents TEMPO: Prompt-based generative pre-trained transformer for time series forecasting

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.470939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:12.210665Z digest=sha256:b8051c020e40e47ec51ce6a778b67852efaa3f6cc5680f2d863e5cc1f13084ce

Observation c255d701-a394-4e5c-be7d-30674f53e250 · outbound

This paper cites Agentboard: An analytical evaluation board of multi-turn llm agents.Advances in Neural Information Processing Systems, 37:74325–74362, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Agentboard: An analytical evaluation board of multi-turn llm agents.Advances in Neural Information Processing Systems, 37:74325–74362, 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.217389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:12.317306Z digest=sha256:a7e71d9fddfd7c1187b6390216fd16e08b2f2139e54f43809ba0c314a2d20f1b

Observation ad43525f-18d8-4cd5-943f-5c15873e9543 · outbound

This paper cites Graphwiz: An instruction-following language model for graph computational problems.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Graphwiz: An instruction-following language model for graph computational problems

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.443472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.443472Z digest=sha256:efdd6ba7a48173445436af3d5835ffded10f3d1f48659721e3793562ed9ea314

Observation 6c75bcdb-df9a-46e9-9849-480d16ce3eba · outbound

This paper cites Deeptransport: Learning spatial-temporal dependency for traffic condition forecasting.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Deeptransport: Learning spatial-temporal dependency for traffic condition forecasting

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.987240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:12.538621Z digest=sha256:88587309e6bf47507c0e5381be94748fb2d35688c7fd51dab417ea99495388ef

Observation a3e9f2f7-2aa0-4acc-8083-40f2d7188618 · outbound

This paper cites TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.611434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.611434Z digest=sha256:41bd588d98936c84788ec218156993ee59176437e5b2e3e254875f9d414ca0ca

Observation c8fe3724-6bf0-4dbf-8be1-42ebe63c9e00 · outbound

This paper cites On the evolution of random graphs.Publ.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents On the evolution of random graphs.Publ

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.735372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:12.700990Z digest=sha256:fd784cce539181b576f07fce5007cdbf1e0d6a1d92f435229ff701b69f7ea853

Observation a59d186c-7fc3-48c5-9aa1-90e8e9f54f9b · outbound

This paper cites Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.799685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.799685Z digest=sha256:2cb2a9a201169492f658727ea68ab42ad6213404c5d6ebd68b7e52031d67a9b4

Observation 90ac15c2-34fb-41be-9f7e-e7ba2f5f5eb2 · outbound

This paper cites CityGPT: Empowering Urban Spatial Cognition of Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CityGPT: Empowering Urban Spatial Cognition of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:12.923658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:12.923658Z digest=sha256:8a678976eaff1782e6354cf4e4e18d8450f6f69e03b1ab802ed67483b94183e2

Observation b05f7949-d8c9-498a-9b12-096d40ba262b · outbound

This paper cites CityBench: Evaluating the Capabilities of Large Language Models for Urban Tasks.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CityBench: Evaluating the Capabilities of Large Language Models for Urban Tasks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.027010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.027010Z digest=sha256:6453d570262fc2d949f86ec556b5362a8eb0c7af5d76a4f976b698a38736b3d7

Observation 56f07dc6-d6a3-4acb-8701-9bc858e7f096 · outbound

This paper cites Pygad: An intuitive genetic algorithm python library.Multimedia tools and applications, 83(20):58029–58042, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Pygad: An intuitive genetic algorithm python library.Multimedia tools and applications, 83(20):58029–58042, 2024

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.524138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:13.132409Z digest=sha256:2aa5d082f2e345faabc7510d0f4abafc2cd00ff4d040942cd99f03c61445cd8e

Observation 301f1204-0906-4602-ba2f-615260bd1b80 · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.204895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.204895Z digest=sha256:6149ea7486e40ebf7c6f2fa142485c2e47fc46ef37719d237d3f676d23d3b8f3

Observation b1aca464-8f36-4df2-b1a8-765071066833 · outbound

This paper cites The Llama 3 Herd of Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents The Llama 3 Herd of Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.286626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.286626Z digest=sha256:2ab5d6c9f32a42b3b23dc47925f70dee89ad5e9c3e8f46aca6b363dd88f3330e

Observation d35c513c-5a74-454e-882a-19d19dea1460 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.387728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.387728Z digest=sha256:6d3b2c2db474dddf4273e6c1e77e95303d0f6593044d4aa25971f3b2a502501b

Observation 88da6acb-25fb-4d14-bf4a-95722d3d2e20 · outbound

This paper cites Language Models Represent Space and Time.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Language Models Represent Space and Time

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.477756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.477756Z digest=sha256:e68c2e43dba754af3e70d64a30f59ae9c02c9842f586ca6d481c50d9f129cefb

Observation 06b0d6f0-dd10-4535-9cb6-71fdfe62f4c8 · outbound

This paper cites The scoot on-line traffic signal optimisation technique.Traffic Engineering & Control, 23(4), 1982.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents The scoot on-line traffic signal optimisation technique.Traffic Engineering & Control, 23(4), 1982

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.285556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:13.575875Z digest=sha256:ea3950159a9c0426947a74c734cfc41359ed5187a2b0ea873c0e139a6fb9c85c

Observation 50b9e666-f72a-4045-93a3-4acfb89323e9 · outbound

This paper cites GPT-4o System Card.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents GPT-4o System Card

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.695425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.695425Z digest=sha256:9869bf65d7138573c33f706cd98fbc6551fe953a1843b921551a564e803e5f9b

Observation 49476d7d-4ca7-496c-ad4c-af8166894df1 · outbound

This paper cites OpenAI o1 System Card.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents OpenAI o1 System Card

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.795554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.795554Z digest=sha256:9d63f3763ecc3541cbe9bb5989f06363799fe0989dc854b1f8b63c0479c6d15f

Observation e88a2df5-49c1-4ece-898c-30b2188c82e7 · outbound

This paper cites Towards mitigating LLM hallucination via self reflection.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards mitigating LLM hallucination via self reflection

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:13.899032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:13.899032Z digest=sha256:36c14a57658cf811157efcccc22bdfee4dca30e74842e065c1bc1f889ed5dc9c

Observation c9424195-67a4-4265-9689-cd4d5e9eb84f · outbound

This paper cites Llmlight: Large language models as traffic signal control agents.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Llmlight: Large language models as traffic signal control agents

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.013525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:14.009911Z digest=sha256:e0bb9ac078ad6054bd12b5035e5faa058ca8b56b9d598a65eff8f91470299f6b

Observation ecf904e5-b7a4-4121-929f-8c83c7c34bcc · outbound

This paper cites Reframing Spatial Reasoning Evaluation in Language Models: A Real-World Simulation Benchmark for Qualitative Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Reframing Spatial Reasoning Evaluation in Language Models: A Real-World Simulation Benchmark for Qualitative Reasoning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:20.582052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:14.089548Z digest=sha256:a37465f0919e45d19d8dd8d4d37dfa62502f15969e0964a283c1283d412e0d8f

Observation ae512c5d-51a9-4e84-99e8-024664a580cc · outbound

This paper cites Repetition in repetition out: Towards understanding neural text degeneration from the data perspective.Advances in Neural Information Processing Systems, 36:72888–72903, 2023.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Repetition in repetition out: Towards understanding neural text degeneration from the data perspective.Advances in Neural Information Processing Systems, 36:72888–72903, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.169851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.169851Z digest=sha256:0b0bc7b1faf27d7d689e1ae1b23dd5521f6318952cc41d15e8cc64d4459f754a

Observation cde91bcb-a5c6-4f1a-9bdf-dc48f39c0cee · outbound

This paper cites Towards alleviating traffic congestion: Optimal route planning for massive-scale trips.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards alleviating traffic congestion: Optimal route planning for massive-scale trips

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.808459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:14.238835Z digest=sha256:a46c7fc8779ded159b2f993461f1076d7916281669a030e0f12975379048726c

Observation bf669141-44b7-4e7e-ab00-e559ef889780 · outbound

This paper cites STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.331330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.331330Z digest=sha256:a8ed83d09afb0b4ef07161e002dd2ce951df3ca66db574ab283947b966ca7b55

Observation e5fbb148-bc8a-42b0-a3c2-d058aaf4748a · outbound

This paper cites Urbangpt: Spatio-temporal large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Urbangpt: Spatio-temporal large language models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.625467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:14.400820Z digest=sha256:83c8349986a4d41e6d5368af213d846977d737e07a05d6e6ebd9c9e516dfb336

Observation 10788f50-7666-48cf-b82f-6b5c41040519 · outbound

This paper cites Timecma: Towards llm-empowered multivariate time series forecasting via cross-modality alignment.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Timecma: Towards llm-empowered multivariate time series forecasting via cross-modality alignment

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.509291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.509291Z digest=sha256:281d3629ba756e34138b57b1aeac8895c611ada42e16496a57b187c70417cfd8

Observation a0ec181d-3d0c-40c4-a3e0-f98ee4e6e9a3 · outbound

This paper cites Knowledge-infused contrastive learning for urban imagery-based socioeconomic prediction.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Knowledge-infused contrastive learning for urban imagery-based socioeconomic prediction

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.411860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:14.582881Z digest=sha256:017e52cba9ed23fc816ffe9efbc4d0d76662a9646ef16b43bfc28245b2f68caa

Observation 422597ef-7eb4-488b-b0b3-4004cea8b6ad · outbound

This paper cites Simulation of urban mobility (sumo), February 4 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Simulation of urban mobility (sumo), February 4 2025

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.198446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:14.661472Z digest=sha256:4f6155e26f4291daa7168108e13e0433624e3de011da46dd1c846e2d64587c37

Observation e3b83036-12e7-4dff-ab0a-8f649ede33e4 · outbound

This paper cites Scats, sydney co-ordinated adaptive traffic system: A traffic responsive method of controlling urban traffic.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Scats, sydney co-ordinated adaptive traffic system: A traffic responsive method of controlling urban traffic

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.016511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:14.750231Z digest=sha256:a9e254ec921a67f95c3795f96f1d6906a435374f360c7f498ae7afc9b9ba5c5d

Observation f747e7a6-3ced-4808-ac9c-b3c9e30c2a34 · outbound

This paper cites SpartQA: : A Textual Question Answering Benchmark for Spatial Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents SpartQA: : A Textual Question Answering Benchmark for Spatial Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:14.883698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:14.883698Z digest=sha256:208d1388fcda406242321bc8d0da2a729789a8faa8d24d840f9e6fe2cee64511

Observation d259570f-b2c7-4498-b981-7de7b44aa08a · outbound

This paper cites Transfer Learning with Synthetic Corpora for Spatial Role Labeling and Reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Transfer Learning with Synthetic Corpora for Spatial Role Labeling and Reasoning

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:20.222720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:14.953677Z digest=sha256:ef486bd715631a0ffc4e8f15d837d8c11d66caf717726b0ef446d1f0ece04a09

Observation 4828f170-3d74-4f7d-b4a9-e47b21946184 · outbound

This paper cites Towards understanding the spatial literacy of chatgpt.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards understanding the spatial literacy of chatgpt

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.760392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:15.053874Z digest=sha256:80d82b613c68e45c69889eca295b8024dfba824367bfe1a1e4ad86880893b0be

Observation 14f51320-84f5-495e-a4e6-d249685ad8dc · outbound

This paper cites Tlc trip record data, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Tlc trip record data, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.551616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:15.128952Z digest=sha256:432ca3b4db2ce95e62315cde84c448eded30f66916d38448a35757f50d6c5f73

Observation 4cc97432-e7f0-474a-9398-f6319ecbaac2 · outbound

This paper cites Dima: An llm-powered ride-hailing assistant at didi.arXiv preprint arXiv:2503.04768, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Dima: An llm-powered ride-hailing assistant at didi.arXiv preprint arXiv:2503.04768, 2025

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.195897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.195897Z digest=sha256:18d8b27a5b0a15102ae3adcf6a3bd158791227313c634f01d6d61e0f4f443012

Observation dec75734-d12d-4434-b73d-9313191c4b88 · outbound

This paper cites UrbanKGent: A Unified Large Language Model Agent Framework for Urban Knowledge Graph Construction.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents UrbanKGent: A Unified Large Language Model Agent Framework for Urban Knowledge Graph Construction

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.268020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.268020Z digest=sha256:519dd9269e547cab22f5a27c4869d1214d9eaa7eb989ab8ea69c1a07d935fb2c

Observation cb413b5a-1f9e-4f3c-9391-efd7c0e6903d · outbound

This paper cites Openstreetmap planet data, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Openstreetmap planet data, 2025

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.285455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:15.353419Z digest=sha256:9aa000842f18c6eebb72526563ef219980324d8cf6ab7eb1ea70ae25446deb4b

Observation 6ad461f8-920f-470c-84de-08b57ba8d8c7 · outbound

This paper cites Self-Reflection in LLM Agents: Effects on Problem-Solving Performance.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Self-Reflection in LLM Agents: Effects on Problem-Solving Performance

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.428482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.428482Z digest=sha256:f0b332ad9eb05df3381b9137ce859cc61170a5c6afbdb210102ef1465ffcea69

Observation 723dd069-13af-447a-a128-06ea71470315 · outbound

This paper cites Sparc and sparp: Spatial reasoning char- acterization and path generation for understanding spatial reasoning capability of large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Sparc and sparp: Spatial reasoning char- acterization and path generation for understanding spatial reasoning capability of large language models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:26.080108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:15.517387Z digest=sha256:7c404404dce89fc8a7d5d9aafc76092da5603262e51091d76a264a6b3b7fac0a

Observation a69353f3-9ea3-4fe3-bebd-270f11f1adc1 · outbound

This paper cites Proximal Policy Optimization Algorithms.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Proximal Policy Optimization Algorithms

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.585405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.585405Z digest=sha256:a2af3c50dbf096f76070519e68b56ece06c27d44ba2f11552f81a26ac8fef9e2

Observation 4daf5592-2ac2-4e64-8a55-200258dc8fc8 · outbound

This paper cites Stepgame: A new benchmark for robust multi- hop spatial reasoning in texts.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Stepgame: A new benchmark for robust multi- hop spatial reasoning in texts

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.851533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:15.674222Z digest=sha256:7666a88197e67fec076c12ebef46e0a3a7d4bf9e67153ab0f56ad0276c854be5

Observation c251b9f7-f9ca-43fe-b35f-5a38e669c87a · outbound

This paper cites Towards Benchmarking and Improving the Temporal Reasoning Capability of Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Towards Benchmarking and Improving the Temporal Reasoning Capability of Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:15.798080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:15.798080Z digest=sha256:6a0be7329ae169c94ea66fd1aa670ee7942540f6c0b86274dea3b996550f204d

Observation 0f1326f0-41f0-49e4-8514-d1ebba27bd56 · outbound

This paper cites Cityflow: A city-scale benchmark for multi-target multi-camera vehicle tracking and re-identification.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Cityflow: A city-scale benchmark for multi-target multi-camera vehicle tracking and re-identification

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.623904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:15.887239Z digest=sha256:062cdbfd6900102acb7523d18e8ac3acd776e21750a6797b8f35a8a0b2037035

Observation fa4bfccf-fd74-4b0e-a578-54a27441a953 · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Qwq-32b: Embracing the power of reinforcement learning, 2025

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.303061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:15.978560Z digest=sha256:5f9a613d4047d6225271cb29b13c03adeeefc4707865303cc777495012733c4d

Observation f4af302b-8d6b-4e6b-8601-3ad39f849674 · outbound

This paper cites Air quality prediction with physics-guided dual neural odes in open systems.ICLR, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Air quality prediction with physics-guided dual neural odes in open systems.ICLR, 2025

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:25.077903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:16.075275Z digest=sha256:d4f8fc3507bb58c52dcf6f50f00c02e32e38202c507925bb862cb0733944d8aa

Observation 2552d402-3c96-4f5b-ad90-9fbdb48000f7 · outbound

This paper cites Applications of artificial intelligence and machine learning in smart cities.Computer Communications, 154:313– 323, 2020.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Applications of artificial intelligence and machine learning in smart cities.Computer Communications, 154:313– 323, 2020

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:24.803848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:16.158597Z digest=sha256:734d8864315fd50f0a83a658dca05067d15b9eb48bc7415f44bdb3f653130df8

Observation af1e7ce4-b8e8-4356-a1e3-b3f04352730e · outbound

This paper cites Robust extrema features for time-series data analysis.IEEE transactions on pattern analysis and machine intelligence, 35(6):1464–1479, 2012.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Robust extrema features for time-series data analysis.IEEE transactions on pattern analysis and machine intelligence, 35(6):1464–1479, 2012

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:24.526996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:16.246214Z digest=sha256:9326bf57ec00350db166c1db67aedbf269a9861e27abb81b1151ce2c4e2d090a

Observation e2e9300c-94fa-4f4f-b587-ac87034a2781 · outbound

This paper cites Reinforcement learning-based placement of charging stations in urban road networks.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Reinforcement learning-based placement of charging stations in urban road networks

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:24.223000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:16.325026Z digest=sha256:d502509bccdede17ccaff5313b62e4f7550b283f0bfc860c7116333743c3e97f

Observation 19ad2d6f-34f2-4d09-9f88-daae5c60538b · outbound

This paper cites A survey on large language model based autonomous agents.Frontiers of Computer Science, 18(6):186345, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents A survey on large language model based autonomous agents.Frontiers of Computer Science, 18(6):186345, 2024

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.401542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.401542Z digest=sha256:87429319f742dbd808d8de406f2f2da64d09ccecdb1ea96e938b4c651d6646ae

Observation 90c123ca-ef52-4e9c-982c-e2b79595fb2c · outbound

This paper cites Global gridded gdp data set consistent with the shared socioeco- nomic pathways.Scientific data, 9(1):221, 2022.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Global gridded gdp data set consistent with the shared socioeco- nomic pathways.Scientific data, 9(1):221, 2022

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.916545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:16.480492Z digest=sha256:cc953cc2aae00823d00c6c5d3ff22c54be9ea0c06b4a4befce08117c92e643ab

Observation bc30ab37-8be1-49a4-99da-2284ce663a53 · outbound

This paper cites Where Would I Go Next? Large Language Models as Human Mobility Predictors.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Where Would I Go Next? Large Language Models as Human Mobility Predictors

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.540853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.540853Z digest=sha256:bb4ff2cf69cc45fe9d990d9537d14e8df364bbd309988dfbe7c04ed4b9d64b3e

Observation 631cb0a6-a581-461d-bb63-c6f357f4aa12 · outbound

This paper cites From news to forecast: Integrating event analysis in llm-based time series forecasting with reflection.Advances in Neural Information Processing Systems, 37:58118–58153, 2024.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents From news to forecast: Integrating event analysis in llm-based time series forecasting with reflection.Advances in Neural Information Processing Systems, 37:58118–58153, 2024

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.622503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.622503Z digest=sha256:c8a40d9577a63feb5c4c2f92fcabe04aac064b2163ab4c116d25954de5e3f35c

Observation d99b94fc-7b0c-45fb-994c-9ed5ba95a5ac · outbound

This paper cites Tram: Benchmarking temporal reasoning for large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Tram: Benchmarking temporal reasoning for large language models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.698546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:16.712157Z digest=sha256:d01c98df285ec1f9add6e63a86059d92eba255e9ccb02de89e8d74e3445a18e5

Observation 3d1eb627-a4e6-4049-aa82-7a20f09596ef · outbound

This paper cites Colight: Learning network-level cooperation for traffic signal control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Colight: Learning network-level cooperation for traffic signal control

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.475729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:16.786439Z digest=sha256:bdc595671c1182db87315071b6b18f7613188f55d138afc5f72a64110ffd1461

Observation 92fe0cfb-22d7-44ae-9710-dde0885ad2fb · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Chain-of-thought prompting elicits reasoning in large language models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:16.887823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:16.887823Z digest=sha256:a05af73d75c8486d727e09fa6b372ee40907896b7007dffc0b9adffb313a8ff4

Observation 505bf480-0890-478b-b977-1c17a5ad767a · outbound

This paper cites Coverage location models: alternatives, approximation, and uncertainty.International Regional Science Review, 39(1):48–76, 2016.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Coverage location models: alternatives, approximation, and uncertainty.International Regional Science Review, 39(1):48–76, 2016

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.247621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:16.946618Z digest=sha256:468f2521ed7e0b2a51840753cee92ec2c392230bbb787ca00db32e27eddde732

Observation 0096260e-70f7-423c-baae-ebacf7b67b57 · outbound

This paper cites Worldpop hub, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Worldpop hub, 2025

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:23.019759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:17.010800Z digest=sha256:b97e24a1272dff3f97f9229ccccd3380b7361c09de683bf091d3c55c0e15fe5d

Observation 65fc1a15-7390-4d78-90d8-35e328c4f2ab · outbound

This paper cites Large language models can learn temporal reasoning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Large language models can learn temporal reasoning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.045214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.045214Z digest=sha256:3d4dcd2dd7fd062950f89097461bb0291a8ae60d786b98182918377655092f37

Observation 63d0c6b9-0757-42fd-99dc-eea7f6aee440 · outbound

This paper cites Evaluating Spatial Understanding of Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Evaluating Spatial Understanding of Large Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.109005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.109005Z digest=sha256:c27de56f11cfe05db1bcc3172ae15088fdaf4e35a9df6fa61614bd1852a61c87

Observation 50dc80f6-64a3-494a-bcaa-7b0934374277 · outbound

This paper cites Qwen2.5 Technical Report.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Qwen2.5 Technical Report

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.170094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.170094Z digest=sha256:f267394785eeb4d42797fc09f329e53646ff5c7af5c0049884e86ba7f3aa9182

Observation a7da1f14-7825-47e0-8d35-858f2469cd3b · outbound

This paper cites Foursquare dataset.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Foursquare dataset

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.831190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:17.230652Z digest=sha256:c5b31c704f229df20ab4337011caf164d3da778b8bf39b9c5329a0d6cb11bc08

Observation 0947063f-e531-4e33-b803-ac5d06090025 · outbound

This paper cites Unist: A prompt-empowered universal model for urban spatio-temporal prediction.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Unist: A prompt-empowered universal model for urban spatio-temporal prediction

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.582495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:17.315597Z digest=sha256:d36b35ebc59a0fb77ab54709e71ea60f8662ddc38d25fb20377534c5bb00fa27

Observation 534d8ca3-9355-478c-be5e-d2617d2ce9cb · outbound

This paper cites CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal Control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal Control

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.318918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.318918Z digest=sha256:01afef30c0a3ca1d206e1c096dc900a3f6116a8bfe3dc2ed6045a3fa7da5bc96

Observation a3a05314-e66e-49aa-a2ab-2a28e89ac2ae · outbound

This paper cites AgentTuning: Enabling Generalized Agent Abilities for LLMs.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents AgentTuning: Enabling Generalized Agent Abilities for LLMs

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.428562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.428562Z digest=sha256:275fafd4392acb11f4092f138e91481a6d34bf50012fa5c47dcc490cd3c89c55

Observation 48950252-710c-4907-94f0-437c9c7b95d3 · outbound

This paper cites Open3dvqa: A benchmark for comprehensive spatial reasoning with multimodal large language model in open space.arXiv preprint arXiv:2503.11094, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Open3dvqa: A benchmark for comprehensive spatial reasoning with multimodal large language model in open space.arXiv preprint arXiv:2503.11094, 2025

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.608625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.608625Z digest=sha256:c5cfe7f5286ede6e27158e9ce5b29ece3bb655ad630532292d8ddbcdb0bb1a84

Observation 2d326e74-c59f-4b3e-8183-5e397f5281bb · outbound

This paper cites Reinforcement learning for traffic signal control.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Reinforcement learning for traffic signal control

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.404837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:17.786637Z digest=sha256:9fd55296e1256f1ad2eb69f905d62a6c0007725a51d637cff46d25f81dfdc2d9

Observation fabf270e-123f-4fc7-9ce1-332b15430add · outbound

This paper cites Urbanvideo-bench: Benchmarking vision- language models on embodied intelligence with video data in urban spaces.arXiv preprint arXiv:2503.06157, 2025.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Urbanvideo-bench: Benchmarking vision- language models on embodied intelligence with video data in urban spaces.arXiv preprint arXiv:2503.06157, 2025

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:17.931121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:17.931121Z digest=sha256:3ee9f8ec50492afe488eb81d61545da983f9826f5f5c0c276409f5975b32a144

Observation 0d75fb69-a602-4e7e-8442-29c2235cce72 · outbound

This paper cites Where to go next: A spatio-temporal gated network for next poi recommendation.IEEE Transactions on Knowledge and Data Engineering, 34(5):2512–2524, 2020.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Where to go next: A spatio-temporal gated network for next poi recommendation.IEEE Transactions on Knowledge and Data Engineering, 34(5):2512–2524, 2020

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.238763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:18.112245Z digest=sha256:f9650c530c621476263d112ba03c6c509dda07fd016ceefa589b28c3189e298d

Observation 41cab77f-526e-4780-82bd-3ba195d55daf · outbound

This paper cites CityEQA: A Hierarchical LLM Agent on Embodied Question Answering Benchmark in City Space.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents CityEQA: A Hierarchical LLM Agent on Embodied Question Answering Benchmark in City Space

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:18.279741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:18.279741Z digest=sha256:09fa9fa2f9215e35b1b38b47dfc53aa38f8c8791a93826291ac5795b62f94825

Observation 6832a9e1-703e-4f49-b092-d7d0f99a8b58 · outbound

This paper cites Llamafactory: Unified efficient fine-tuning of 100+ language models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Llamafactory: Unified efficient fine-tuning of 100+ language models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:18.409485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:18.409485Z digest=sha256:fd0506f05f77bb50212fa0b6ce05e65311fa3983e8f071f30f8c26bb5cb25c00

Observation 19360793-b3c5-4839-83e3-15ec9d74624c · outbound

This paper cites Spatial planning of urban communities via deep reinforcement learning.Nature Computational Science, 3(9):748– 762, 2023.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Spatial planning of urban communities via deep reinforcement learning.Nature Computational Science, 3(9):748– 762, 2023

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:22.005609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:18.533561Z digest=sha256:72b51707ecbef54801ddf6ea53a490d39e9d7ff0927cdbe9653915ae0d551961

Observation be5f3e12-424a-4851-9e95-73b5a71994bf · outbound

This paper cites UrbanPlanBench: A Comprehensive Urban Planning Benchmark for Evaluating Large Language Models.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents UrbanPlanBench: A Comprehensive Urban Planning Benchmark for Evaluating Large Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:18.662366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:18.662366Z digest=sha256:7ad76410de72cda5525c44b6313d1d7875c48e4365bf69468a0e59b75b963455

Observation c5dd61e9-0ffb-464b-97af-6e169aee06af · outbound

This paper cites Road planning for slums via deep reinforcement learning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Road planning for slums via deep reinforcement learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.838077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:18.808980Z digest=sha256:49ef48c65fa9bcfb1aa180175a5daf5c18b44c68d01e44f3b5decf3f884574eb

Observation 61ce2295-ac49-4efe-a6f4-b1a2f38a06a2 · outbound

This paper cites going on a vacation.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents going on a vacation

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.632527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:18.941064Z digest=sha256:bcb756823f346fc2604b8391d7ca1e1b6a1136494781628c9d217773e5136999

Observation 8b91d25d-f6ec-49d2-9dac-c2c543ed9e90 · outbound

This paper cites Large Language Model for Participatory Urban Planning.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Large Language Model for Participatory Urban Planning

Reference 75

Resolution
malformed identifier
no resolver link, observed 2026-08-07T14:47:19.072276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:19.072276Z digest=sha256:c883d0ebc57f3805afff711957856736f32b2de5759d3825ad24e831d270ad94

Observation 25e769a7-7f59-4d94-8167-be3f432879db · outbound

This paper cites answer":.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents answer":

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.417259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:19.242745Z digest=sha256:9aed01840ff932e32abb92a1955ac4c202cb8e478ce2f282eb5132967baa8130

Observation 0a1940f9-9f5b-4dcf-a6ce-05b98748b6fb · outbound

This paper cites Miscellaneous Shop.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents Miscellaneous Shop

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:21.155850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:19.378369Z digest=sha256:7f420fb2a91e5cb313714e52501ae45a5193763824c4bed006f66b4454fa998d

Observation 682dc218-38ba-4888-9e43-f4757dff36fb · outbound

This paper cites answer":.

USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents answer":

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:20.945046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:19.524779Z digest=sha256:25806fa9c925afa12888cf8dd9e11fe4f9a27f4801caa489cea925a918b615ff

Pith citing papers

Observation 12238777-3eda-4da9-85cd-1e1a49d2e174 · inbound

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation cites this paper.

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation USTBench: Benchmarking and Dissecting Spatiotemporal Reasoning of LLMs as Urban Agents

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:51:08.882918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T05:47:29.031123Z digest=sha256:186383962d694f25e6509665e7ef3c2184eb2472e864400c838c14e0a5468b46