Pith. sign in

Paper Citation Record · LEDGER

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding

As of 7 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 2 inbound Pith citation observations for arXiv:2506.23219.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.23219 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:52:25.446679Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T06:58:39.117228Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T07:26:45.809288Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact1
  • verified fuzzy35
  • unresolved28
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a66870f0-4af1-4c6e-b323-672ffb45e0fd · outbound

This paper cites LAMP: A Language Model on the Map.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding LAMP: A Language Model on the Map

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:19.753815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:19.753815Z digest=sha256:98933035e4ae6602d5c123880557980601f82931a5e86eee03edf680c2c3a661

Observation 69493768-4882-459e-b673-245ff3654219 · outbound

This paper cites City foundation models for learning general purpose rep- resentations from openstreetmap.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding City foundation models for learning general purpose rep- resentations from openstreetmap

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.883775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:19.839179Z digest=sha256:7840ad93d14c4f4fe69135a0c6fb71a32fccaefe46e57e8b1da4013e4f9b8ca0

Observation 983a5a70-1d8b-48e3-99a8-3f7c8161bcbc · outbound

This paper cites Street view imagery in urban analytics and gis: A review.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Street view imagery in urban analytics and gis: A review

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.689303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:19.924010Z digest=sha256:63d739318522f9a7b20506b829e5d8cfe79169787261ae87d3d8c69eb5f9c41b

Observation 10e1defb-e8da-4ac3-88bc-9edc797d602c · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.060217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.060217Z digest=sha256:032f81b04bd78e7bdad43e69f93fb609f0df726eee98065061278678fe665012

Observation bfce9a89-df1e-437e-b8b9-a9c1ed8543ae · outbound

This paper cites Touchdown: Natural language naviga- tion and spatial reasoning in visual street environments.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Touchdown: Natural language naviga- tion and spatial reasoning in visual street environments

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.487883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:20.178492Z digest=sha256:619c0eaa444b572db01b38fb767ca05500919d6f20c6811e5b9b6e9dfe3b6aab

Observation c0282037-4302-4a00-acf3-f66c822daa7d · outbound

This paper cites ShareGPT4V: Improving Large Multi-Modal Models with Better Captions.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.279386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.279386Z digest=sha256:1dbba8c36b0c3cc3ddc89b8c462d3af996c1c186e16d2740cdc59e60dbf9d844

Observation f5b2046a-852f-4c85-b199-852e4af903d3 · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.373110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.373110Z digest=sha256:1db25ece34c111c46418a1c99be01e3732497d9143e64ebfd45a6ab5a7e781f3

Observation 44fc0a9b-b231-4285-8142-b90368f7e855 · outbound

This paper cites Internvl: Scaling up vision founda- tion models and aligning for generic visual-linguistic tasks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Internvl: Scaling up vision founda- tion models and aligning for generic visual-linguistic tasks

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.338270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:20.462260Z digest=sha256:3591c4cc97d8aa454fbfe3ddc74cda58e3c1d673bb26f20eaeb87fb2e702b673

Observation ca1bd479-40fe-484f-a892-1beea35e5999 · outbound

This paper cites SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.579394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.579394Z digest=sha256:6076a8e54df6fa8b0cb7ec97eb061c3cf039fe2a4290f5977a56c52f11456958

Observation d5635dfc-a5bd-429a-a4ca-67bd4b671ad8 · outbound

This paper cites Understanding world or predict- ing future? a comprehensive survey of world models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Understanding world or predict- ing future? a comprehensive survey of world models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.691552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.691552Z digest=sha256:d2958aaa0a33cc60029f9c4aa2bcc40bf716c98c52e99da56b9334b165c303c5

Observation f8f69a58-ccc7-48f3-94c7-4d997701a538 · outbound

This paper cites Enhancing Chat Language Models by Scaling High-quality Instructional Conversations.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Enhancing Chat Language Models by Scaling High-quality Instructional Conversations

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.776159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.776159Z digest=sha256:48bfb521d4803bc775f799c2a139836d3fae4b0dd709cfa359957590991311a0

Observation b09e2d68-1545-4e28-b366-9000934f1574 · outbound

This paper cites How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.860831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.860831Z digest=sha256:18c5ec14a87f31f2a5fef623080535cb8f910d3128c01b96dc74f7fd5cb728eb

Observation 33b27999-3ad1-4f5a-ad15-fbb09880377a · outbound

This paper cites Vlmevalkit: An open- source toolkit for evaluating large multi-modality models,.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Vlmevalkit: An open- source toolkit for evaluating large multi-modality models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.969557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.969557Z digest=sha256:a203f7ea85709cc58a2ce853f6df6b08f661710ea6210585013dff3c30b3393d

Observation 1b813eca-e9b5-4467-99db-6e324c9003b4 · outbound

This paper cites Urban visual intelligence: Uncovering hidden city pro- files with street view images.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urban visual intelligence: Uncovering hidden city pro- files with street view images

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.158394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.035326Z digest=sha256:8f16a09672fba445d7a5ad6ea2b2e94fbadcffd98a91a7c3652757fb116e6e93

Observation d719cf63-a281-4dc5-b9a6-fbf82669f95c · outbound

This paper cites Agent- move: A large language model based agentic framework for zero-shot next location prediction.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Agent- move: A large language model based agentic framework for zero-shot next location prediction

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.016597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.152105Z digest=sha256:3535502d1a73560055a74501a500ad95629542040f00db437d0295c954dfa63b

Observation 5ab581ac-0389-4ed1-a2ef-94ca7df1bca1 · outbound

This paper cites Citygpt: Empowering urban spatial cognition of large language models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Citygpt: Empowering urban spatial cognition of large language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.818015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.237329Z digest=sha256:8629da7a101aede60c5bbf3c4b42f5303d67813c9928d75394319d15e203ab50

Observation a030dcfb-3b62-4f79-8cd5-b941de0a9c60 · outbound

This paper cites A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.335105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.335105Z digest=sha256:4e52635c1c69d3fc2b29ec2a91fc0f73b79d12ee62a3eaaed0125567b7519f5c

Observation 50e01ba8-4998-46e5-b4d7-22fb6e6410dc · outbound

This paper cites City- bench: Evaluating the capabilities of large language models for urban tasks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding City- bench: Evaluating the capabilities of large language models for urban tasks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.556275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.403131Z digest=sha256:d62701a5a2996f4474769d6bcd08a0eef8c6444050d1e510bd8ee1df6082ef59

Observation dc817789-0086-487f-a05f-52dacc81f7ab · outbound

This paper cites Imagebind: One embedding space to bind them all.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Imagebind: One embedding space to bind them all

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.502930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.502930Z digest=sha256:ee0273cc7e5512687dffdb00e11f74ab1effb5df2bf379404dd5c258ae54a7c9

Observation 4f1933ac-cb85-483b-b11f-7834f2a98765 · outbound

This paper cites Mobility-LLM: Learning Visiting Intentions and Travel Preferences from Human Mobility Data with Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Mobility-LLM: Learning Visiting Intentions and Travel Preferences from Human Mobility Data with Large Language Models

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:52:26.087467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.581742Z digest=sha256:9d25a76a4fdee577a2a0d1ca378e2486e4a8a531a2befc39b14c9b54ee291eab

Observation f61311cd-f3e4-4df6-b205-1b23b4838077 · outbound

This paper cites Regiongpt: Towards region understanding vision lan- guage model.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Regiongpt: Towards region understanding vision lan- guage model

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.357391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.665075Z digest=sha256:50d92cd385d86204a395a1ea3e6d02d820d8e4dce3a8f007ac5993ed04f05dff

Observation 57e99a67-4c0a-4127-8b1c-d71f82485cd5 · outbound

This paper cites UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.758851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.758851Z digest=sha256:2cd20862b86c6e0b93f30bbd3140b46bd718160d3acbb2920afd1349826ec106

Observation 8e59aa17-e176-46fe-b9ea-ac00f6d93b75 · outbound

This paper cites Vision-language models for medical report generation and visual question answering: A review, 2024.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Vision-language models for medical report generation and visual question answering: A review, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.222258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.863618Z digest=sha256:f5c1d570893fece65e2044e59dd6a6cf3d28fcdf76417a38c9a5c249ac505ed9

Observation 89386da8-4e18-4d7c-801e-a4d5714a0262 · outbound

This paper cites RSGPT: A Remote Sensing Vision Language Model and Benchmark.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.928081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.928081Z digest=sha256:54aeb9d3dc2008992deb3ea67af0bf74d64644b6e0772a6735776294737437a1

Observation 4c330697-6e82-442f-9a38-205e3a50ed8c · outbound

This paper cites Time-LLM: Time Series Forecasting by Reprogramming Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Time-LLM: Time Series Forecasting by Reprogramming Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:22.043641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:22.043641Z digest=sha256:b7f025cf84e3723c653477d1b6abbf59be630f411cc09e94ea5a802a60cc4889

Observation 969e79a5-5099-4815-b454-9f78d05a7ada · outbound

This paper cites Geochat: Grounded large vision-language model for remote sensing.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Geochat: Grounded large vision-language model for remote sensing

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.998699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:22.119635Z digest=sha256:ff7d3bd8491af549b6190005deb67081067bdfcb5af1875b562b7f663e89773e

Observation 1899628e-54f6-4046-8779-ac2b441d92f0 · outbound

This paper cites Llava-med: Training a large language- and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36, 2024.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Llava-med: Training a large language- and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36, 2024

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.803628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:22.243892Z digest=sha256:44bd0b7b269d7ef5f22d79a919b0d16bd6313a99b76e98aa56a9ad043d9356a2

Observation 0a1756f7-099d-4bc0-81f1-af3980b310af · outbound

This paper cites Urbangpt: Spatio- temporal large language models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urbangpt: Spatio- temporal large language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.628393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:22.380917Z digest=sha256:d8d1305cfcc5144700948b2035f9dd433cd2d7c1e90425bc0f8a22012d1aab07

Observation 7fcd0e7e-db48-4ec4-8f6f-70a57a09de3e · outbound

This paper cites Vila: On pre-training for visual language models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Vila: On pre-training for visual language models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.420034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:22.486258Z digest=sha256:5d7827fd9cc1c93dd311ca94a7a966d3cb7daacf75f623320bf20ef95fc915dd

Observation e414ca83-90fd-4e0b-a0ae-fe4fa6fc5785 · outbound

This paper cites Improved baselines with visual instruction tuning.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Improved baselines with visual instruction tuning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.256034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:22.603566Z digest=sha256:41dee42bacb7f5713a927e9167431982d8552e47397713bd328bd2272d800862

Observation 774b3427-2c81-4631-9e66-edd2016f0633 · outbound

This paper cites Visual instruction tuning.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Visual instruction tuning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.082401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:22.687759Z digest=sha256:f7c8bc9603f692167f23038f8ea9dc41584e57d2901ac624e47fa7395542175b

Observation e39f3b66-cf9c-4426-963e-f6fafe0362c4 · outbound

This paper cites Citylens: Bench- 10 marking large language-vision models for urban socioeco- nomic sensing.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Citylens: Bench- 10 marking large language-vision models for urban socioeco- nomic sensing

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:22.798308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:22.798308Z digest=sha256:a891aff9c70a0b0cc034d3836a5681558b74f488b095db7814208c8c32e4daaf

Observation cf9c645e-b765-4ad7-85d1-1def9183b5d5 · outbound

This paper cites SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:22.909176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:22.909176Z digest=sha256:7ada48c21b6c5b36ff8e4166a8c51ea769a90069a74ec496ca762c9df24403b8

Observation 6dd443bf-64e6-47c3-be1f-8389e6fb831e · outbound

This paper cites Dolphins: Multimodal Language Model for Driving.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Dolphins: Multimodal Language Model for Driving

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.026538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.026538Z digest=sha256:ec5c6ae7e7c24e8d0cc5e165ac185e357b2e900c13fb399c48a98a71b5059e62

Observation 335ccf75-0269-4d48-831b-104bd6031418 · outbound

This paper cites On the opportunities and chal- lenges of foundation models for geoai (vision paper).

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding On the opportunities and chal- lenges of foundation models for geoai (vision paper)

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.920846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.139438Z digest=sha256:f7ffb96b71fc6d6e8fe5c683737602f998ec0c79a90dd86608a7077986fa7eb7

Observation 188d55e0-3b17-4ea4-bc96-86149bdc97ce · outbound

This paper cites LLaMA 3.2: Advancing Vision, Edge, and Mo- bile Devices.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding LLaMA 3.2: Advancing Vision, Edge, and Mo- bile Devices

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.727923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.201817Z digest=sha256:9c143da1ec25c90e84fb6aa0452c8e4fe4fbc4dae7f38878d9f3a59bd12c7dea

Observation 2bc776dd-464a-45e3-b098-c2300c8c73d3 · outbound

This paper cites LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.293073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.293073Z digest=sha256:9f756dd1d4602156731f1795a20a012be5a3754f626127ddb993fe0f6ee5646d

Observation bccbedbe-3ea0-47a8-a30b-31564a8b6581 · outbound

This paper cites Introducing chatgpt.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Introducing chatgpt

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.520569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.383898Z digest=sha256:4cc37a589b7ed24d90fd162783e69c0a76f5979dcaacc691325358d0e9f25519

Observation 8013e578-b61f-4c03-8585-db60f194e6e8 · outbound

This paper cites Gpt-4v(ision) system card.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Gpt-4v(ision) system card

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.306992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.464642Z digest=sha256:3069b7c0dc555739da6449106c83303ae9e6be97ea72f7318bb9d65ba6200aa4

Observation 35a0cb28-947f-4e9c-842a-a0b494a6e2ea · outbound

This paper cites Hello GPT-4.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Hello GPT-4

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.119217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.579833Z digest=sha256:f9eb489b32ee98537ffd68ff8b3a69fa1ceb903810990d563ab924401c3cd0a4

Observation 985fb899-f005-463c-b1e5-e3499e64c72f · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.674340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.674340Z digest=sha256:645b4d729fc04f3bad9f536f433979f430717557690cca9884ad3f47ac5166e1

Observation bf746aaa-b2a3-428e-94c4-d7883a7ee9c8 · outbound

This paper cites VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language Tasks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language Tasks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.743835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.743835Z digest=sha256:983add431b2fc4ce2c763300698b872710b6aa2f226d3f490eaca79113507a21

Observation ed1fafd1-c4d0-420a-9d24-a1223f0adc7b · outbound

This paper cites V*: Guided visual search as a core mechanism in multimodal llms.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding V*: Guided visual search as a core mechanism in multimodal llms

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.950965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.852374Z digest=sha256:a36e91d1cd50d22a8d45fd9e938e764cca925219eb05464f41da3668b91aeb19

Observation 3f187676-0180-47b0-941b-b631ccba57d8 · outbound

This paper cites RealworldQA Dataset.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RealworldQA Dataset

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.803340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.943108Z digest=sha256:c56d17900c01575ac762c00637dbed8d5b642c6695d16d8c0caa86ee2d004bf2

Observation 83b5e54f-4f66-4e79-a575-7d0fce68ee2b · outbound

This paper cites Analyz- ing large language models’ capability in location prediction.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Analyz- ing large language models’ capability in location prediction

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.663944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.066125Z digest=sha256:e2952031d71f4174e9ecf9d88c98f7c3f76aa35ffcaa5d451c25e2ad74f223d5

Observation 8dc2dfb8-1c7c-4968-8c4e-506a1669c93f · outbound

This paper cites Urban Generative Intelligence (UGI): A Foundational Platform for Agents in Embodied City Environment.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urban Generative Intelligence (UGI): A Foundational Platform for Agents in Embodied City Environment

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.182301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.182301Z digest=sha256:547c50cea73f4a860dbb6d0e38f0e650df5d8aafdc1724f45e8245e50d7ebd0f

Observation 9df47446-7336-44b2-a638-89adaf931ff2 · outbound

This paper cites Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.272058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.272058Z digest=sha256:8af4d25cc8e5599aa7fbee8f12b12af5b494d00a9ec7d5c4b409507f7045b652

Observation a42cc69e-bb69-44fa-b228-7a09bb0316a3 · outbound

This paper cites Par- ticipatory cultural mapping based on collective behavior data in location-based social networks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Par- ticipatory cultural mapping based on collective behavior data in location-based social networks

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.479881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.356229Z digest=sha256:5df87577cb6a25a732d47d38d326b14b5d62e08e81067f9769dc3fcf76e4bf0e

Observation af2619d3-6b9d-4e69-a7fa-59a751aa8b8f · outbound

This paper cites A Survey on Multimodal Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding A Survey on Multimodal Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.419530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.419530Z digest=sha256:1386368a478e513a196575c951880ab60ade1d20e70ae20ebbd423a1e7f9abef

Observation cd0b7066-f87c-4c05-adf0-3d98eb6d0e1d · outbound

This paper cites Mm-vet: Evaluating large multimodal models for inte- grated capabilities.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Mm-vet: Evaluating large multimodal models for inte- grated capabilities

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.296895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.484165Z digest=sha256:86e66ceb4d6c25452ad8314c4788d994733814bf167a81010c28a72f3c349224

Observation 4ea88e8a-8db0-4ce1-8d73-77ef38fac776 · outbound

This paper cites SkyEyeGPT: Unifying Remote Sensing Vision-Language Tasks via Instruction Tuning with Large Language Model.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding SkyEyeGPT: Unifying Remote Sensing Vision-Language Tasks via Instruction Tuning with Large Language Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.557582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.557582Z digest=sha256:f9ef870511a0db06f4e2302abca6b127a06d07fb4fbb2ec493d33df1d807ca6a

Observation fb0d3055-79a3-4d39-af6d-94273c5476cc · outbound

This paper cites Earthgpt: A universal multi-modal large lan- guage model for multi-sensor image comprehension in re- mote sensing domain.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Earthgpt: A universal multi-modal large lan- guage model for multi-sensor image comprehension in re- mote sensing domain

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.120055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.628535Z digest=sha256:c33882def4350fbd5ec4b9c2f9c39efec4259b4dd24a862faf07fe49c7e628e5

Observation 88ad8c4c-0f1e-420f-8a82-ee386e1cc1c4 · outbound

This paper cites Urban foundation models: A survey.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urban foundation models: A survey

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.951040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.672320Z digest=sha256:d140c86bf8f6e40f58954a2e227a7d5cfef4ab7d0314bd1643d6212cb8db619b

Observation f16343f6-e71c-4829-b6df-6ff60778e59a · outbound

This paper cites UrbanMLLM: Joint learning of cross-view imagery for urban understanding, 2025.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding UrbanMLLM: Joint learning of cross-view imagery for urban understanding, 2025

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.815763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.721919Z digest=sha256:59ecf882dfd54a6f395631bb6a90078353412dfb33231962d72fe4b846cc0a78

Observation 5f8b831f-e89b-4263-a0c4-c25564bd2048 · outbound

This paper cites Per- ceiving urban inequality from imagery using visual language models with chain-of-thought reasoning.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Per- ceiving urban inequality from imagery using visual language models with chain-of-thought reasoning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.618477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.801393Z digest=sha256:05ef4de33497624e65d36e783e59e8ef9c80beb9de30587b6f6687807ec59bfc

Observation d2bfc31c-65af-4250-886b-738b385c421d · outbound

This paper cites Urbench: A comprehensive bench- mark for evaluating large multimodal models in multi-view urban scenarios.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urbench: A comprehensive bench- mark for evaluating large multimodal models in multi-view urban scenarios

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.440727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.854721Z digest=sha256:c325661a8cb36c44c02490d68daa4b33758d5794a6a49c9c59b5614d7172c465

Observation 29ccec06-f7c7-4493-9da2-98a5f2a90a87 · outbound

This paper cites Deep learning for cross-domain data fu- sion in urban computing: Taxonomy, advances, and outlook.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Deep learning for cross-domain data fu- sion in urban computing: Taxonomy, advances, and outlook

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.257763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.894192Z digest=sha256:43984d97d7263598ed40c9868f732f03deb34937dd166c3467b53e34947ae287

Observation ef3ed754-043d-4ed6-bdfa-550a9d2985bb · outbound

This paper cites Figure 9.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Figure 9

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.008529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.963116Z digest=sha256:bbf5e3168f8cc3315117d03dac4870b514231491ac4642a9a5076923f1728d60

Observation f83a9c12-7d2c-48a3-a27f-632c538467da · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:28.721388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.019923Z digest=sha256:6edec4cf9b25d80a9333c4d523a96f375c83dda841299cabf05035fdda8c4867

Observation 69414709-d8dc-4be6-ad26-cff32443ebda · outbound

This paper cites Table 2 in Section 3.2 is the aggregated results of these three tables.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Table 2 in Section 3.2 is the aggregated results of these three tables

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:28.498738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.067667Z digest=sha256:6aa7f910ad1fd655b2fb8d9b58fb3b3f44e98bab6fe6c73e6856464ad76dc2dd

Observation 422b2167-4930-4227-b49b-cd6c734d3fad · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:28.278570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.119209Z digest=sha256:378fbf436a15f664c3bcf22942f7f63b363cd87c531de76d2b10a7b81256fad0

Observation fd295feb-c4b9-4d64-964b-22449050b58c · outbound

This paper cites 11 presents training results with different amounts, ex- hibiting the high quality of UData.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding 11 presents training results with different amounts, ex- hibiting the high quality of UData

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:27.972132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.203938Z digest=sha256:aca8a9b9464bb678b9002a53b501f29f85e63952c1115087d0f475a597d4d5ef

Observation 8344fae6-6c4b-4b8f-a970-3f0b5373f535 · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:27.625332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.253531Z digest=sha256:88d6093d8213ec692386a8a91222721f8941c0114f25fbe376df0c44c9c5a3a4

Observation 7385252b-4d0d-42e6-8b30-6911e05c794a · outbound

This paper cites However, for certain tasks, models of different sizes exhibit similar capabilities.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding However, for certain tasks, models of different sizes exhibit similar capabilities

Reference 64

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T21:52:27.351086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.315959Z digest=sha256:7b6fe73f894009f693e70c094f25c504d8e0ff53b52bcb51e6f844202d0dbe6a

Observation bd8fd18b-f630-4512-a500-92c471c491ba · outbound

This paper cites This task needs a model to speculate the land use type (commercial, residential, agricultural, etc.) based on a satellite image.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding This task needs a model to speculate the land use type (commercial, residential, agricultural, etc.) based on a satellite image

Reference 65

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T21:52:27.097941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.382008Z digest=sha256:805dfa3d94272c794b5be5dd8790251c75d9d50d72c3f360dd5d70ef6077eabb

Observation 272f3bd7-3e1a-488c-b297-f5dd07e3e998 · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:26.762359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.446679Z digest=sha256:8ecdd8dff6df2ca76620d00fee1503e8f2a99220b7be5587086faaa3a041c60c

Pith citing papers

Observation eba8392f-a4ca-4c0c-b565-bdf2db73376e · inbound

IoT-Brain: Grounding LLMs for Semantic-Spatial Sensor Scheduling cites this paper.

IoT-Brain: Grounding LLMs for Semantic-Spatial Sensor Scheduling UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:05:56.875660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T17:49:16.409834Z digest=sha256:19bd9862778b46c032cc0f41043f9b5c6be890d2bf8d2207cedc7b5187d3d83d

Observation 4a8bd704-34ac-4f67-ab3b-97c1b8c68f6c · inbound

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models cites this paper.

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:26:45.810959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T06:58:39.117228Z digest=sha256:f064f1f8fa57538bf0cec9ddd755762a21bc3c049c097ed70da8875d6c398f91