Pith. sign in

Paper Citation Record · LEDGER

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding

As of 7 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 2 inbound Pith citation observations for arXiv:2506.23219.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.23219 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:52:25.446679Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T06:58:39.117228Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T07:26:45.809288Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact1
  • verified fuzzy35
  • unresolved28
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a66870f0-4af1-4c6e-b323-672ffb45e0fd · outbound

This paper cites LAMP: A Language Model on the Map.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding LAMP: A Language Model on the Map

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:19.753815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:19.753815Z digest=sha256:7a22b5c14f8a3195620c775bc2578ddb06f4f54d870b5179a5c336ce0791fb7e

Observation 69493768-4882-459e-b673-245ff3654219 · outbound

This paper cites City foundation models for learning general purpose rep- resentations from openstreetmap.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding City foundation models for learning general purpose rep- resentations from openstreetmap

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.883775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:19.839179Z digest=sha256:4a93b4b048aed3313b8f47223ddca5d804dd0ab4f79a0a6c32c11e025cf55567

Observation 983a5a70-1d8b-48e3-99a8-3f7c8161bcbc · outbound

This paper cites Street view imagery in urban analytics and gis: A review.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Street view imagery in urban analytics and gis: A review

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.689303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:19.924010Z digest=sha256:627d82556f3838939119e968dd7b07c8da06916147fdb8f4676546d6e90dee07

Observation 10e1defb-e8da-4ac3-88bc-9edc797d602c · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.060217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.060217Z digest=sha256:282b26fa95324a8e3797d2a099a6fbd01083387091f2dc673e2534848b4416be

Observation bfce9a89-df1e-437e-b8b9-a9c1ed8543ae · outbound

This paper cites Touchdown: Natural language naviga- tion and spatial reasoning in visual street environments.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Touchdown: Natural language naviga- tion and spatial reasoning in visual street environments

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.487883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:20.178492Z digest=sha256:37277293dca550e6d7e29847e995c1092a2f95da56d6c09d427236ce073e0a3c

Observation c0282037-4302-4a00-acf3-f66c822daa7d · outbound

This paper cites ShareGPT4V: Improving Large Multi-Modal Models with Better Captions.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.279386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.279386Z digest=sha256:eeb3a72a54ec0c1148656fdd9b1e165fcc976aeae6ad74b4dcb7aa1f9f724ea8

Observation f5b2046a-852f-4c85-b199-852e4af903d3 · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.373110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.373110Z digest=sha256:c7de9ff480db6f6bb744ca469fbe73c4ff4acdd185860f4433a27258a12c32c3

Observation 44fc0a9b-b231-4285-8142-b90368f7e855 · outbound

This paper cites Internvl: Scaling up vision founda- tion models and aligning for generic visual-linguistic tasks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Internvl: Scaling up vision founda- tion models and aligning for generic visual-linguistic tasks

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.338270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:20.462260Z digest=sha256:33025c6950da2c64bc093aa1f04ae897ac1239b4519254fbded7857125cfa0e8

Observation ca1bd479-40fe-484f-a892-1beea35e5999 · outbound

This paper cites SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.579394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.579394Z digest=sha256:be9495e73c5b9dbc2167f2ba4177014a99ef00e98affe08bfd033d2b36ee2728

Observation d5635dfc-a5bd-429a-a4ca-67bd4b671ad8 · outbound

This paper cites Understanding world or predict- ing future? a comprehensive survey of world models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Understanding world or predict- ing future? a comprehensive survey of world models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.691552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.691552Z digest=sha256:cf7f19dacb500db34a3025a0f94e318dcfdd9866fb6e298d0461fb57c00e6a96

Observation f8f69a58-ccc7-48f3-94c7-4d997701a538 · outbound

This paper cites Enhancing Chat Language Models by Scaling High-quality Instructional Conversations.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Enhancing Chat Language Models by Scaling High-quality Instructional Conversations

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.776159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.776159Z digest=sha256:9780daec2e47f8944214b7cb7ba00a2ce4f26f9704e25dc4bffdaad598ea19b5

Observation b09e2d68-1545-4e28-b366-9000934f1574 · outbound

This paper cites How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.860831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.860831Z digest=sha256:99735908b27919e54c59d7a0581d2511293fbccde751d6e725847202c2324337

Observation 33b27999-3ad1-4f5a-ad15-fbb09880377a · outbound

This paper cites Vlmevalkit: An open- source toolkit for evaluating large multi-modality models,.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Vlmevalkit: An open- source toolkit for evaluating large multi-modality models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.969557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.969557Z digest=sha256:87d6e9b4a20a8b96f6cd5636fe6fff44f8d5bd8d5436ee3e6692a8c90e3ba3e1

Observation 1b813eca-e9b5-4467-99db-6e324c9003b4 · outbound

This paper cites Urban visual intelligence: Uncovering hidden city pro- files with street view images.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urban visual intelligence: Uncovering hidden city pro- files with street view images

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.158394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.035326Z digest=sha256:8d542a75c1936372bb0790827f28b8accb1adc163e1f6ac2d9670fc4cae16d83

Observation d719cf63-a281-4dc5-b9a6-fbf82669f95c · outbound

This paper cites Agent- move: A large language model based agentic framework for zero-shot next location prediction.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Agent- move: A large language model based agentic framework for zero-shot next location prediction

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.016597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.152105Z digest=sha256:3c2f46042983bd4913cf65e269e941795ad7a83c5d951681b40a8f0af3298a5d

Observation 5ab581ac-0389-4ed1-a2ef-94ca7df1bca1 · outbound

This paper cites Citygpt: Empowering urban spatial cognition of large language models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Citygpt: Empowering urban spatial cognition of large language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.818015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.237329Z digest=sha256:efee302f55c3462cf572c4ad55730a0cd79e3f5838f983c59da9f53cc68249a4

Observation a030dcfb-3b62-4f79-8cd5-b941de0a9c60 · outbound

This paper cites A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.335105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.335105Z digest=sha256:6c605de910ec6437affc4c59f0e39a0257e7061966c0693c27cf4ebebce99352

Observation 50e01ba8-4998-46e5-b4d7-22fb6e6410dc · outbound

This paper cites City- bench: Evaluating the capabilities of large language models for urban tasks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding City- bench: Evaluating the capabilities of large language models for urban tasks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.556275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.403131Z digest=sha256:34045d5c14122ffd8bc3b208fc6bb11f59f7f021a3893f529dfd67412ae90fb7

Observation dc817789-0086-487f-a05f-52dacc81f7ab · outbound

This paper cites Imagebind: One embedding space to bind them all.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Imagebind: One embedding space to bind them all

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.502930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.502930Z digest=sha256:8484c464b871230a8bb4c43d2ba17ef9ae5180c43dd73f65799eadcc91ecf9fb

Observation 4f1933ac-cb85-483b-b11f-7834f2a98765 · outbound

This paper cites Mobility-LLM: Learning Visiting Intentions and Travel Preferences from Human Mobility Data with Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Mobility-LLM: Learning Visiting Intentions and Travel Preferences from Human Mobility Data with Large Language Models

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:52:26.087467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.581742Z digest=sha256:6006d4c7c46ccd667843fe2d7d929b1f6d1d59409445e107aef3e9246884b8bc

Observation f61311cd-f3e4-4df6-b205-1b23b4838077 · outbound

This paper cites Regiongpt: Towards region understanding vision lan- guage model.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Regiongpt: Towards region understanding vision lan- guage model

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.357391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.665075Z digest=sha256:6ef5dc1575138aefae194a4efbbaa4113a6f10ac0325c59d6d51d032a6608d66

Observation 57e99a67-4c0a-4127-8b1c-d71f82485cd5 · outbound

This paper cites UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.758851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.758851Z digest=sha256:03afff7eaa7c87fb48907e30dc0ab11ee15ad00880269d57464f36366cbf1f8f

Observation 8e59aa17-e176-46fe-b9ea-ac00f6d93b75 · outbound

This paper cites Vision-language models for medical report generation and visual question answering: A review, 2024.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Vision-language models for medical report generation and visual question answering: A review, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.222258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:21.863618Z digest=sha256:57181c940a888fcee9fca026b76170fe00adf291619d9888a887d16771004d04

Observation 89386da8-4e18-4d7c-801e-a4d5714a0262 · outbound

This paper cites RSGPT: A Remote Sensing Vision Language Model and Benchmark.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.928081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.928081Z digest=sha256:26da1939e8d937f75f8d263ac01f6725815cd9e11fd4442365a6374411c1c3e8

Observation 4c330697-6e82-442f-9a38-205e3a50ed8c · outbound

This paper cites Time-LLM: Time Series Forecasting by Reprogramming Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Time-LLM: Time Series Forecasting by Reprogramming Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:22.043641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:22.043641Z digest=sha256:edd103745bc4fa77e102400aedb49a7eaf0f9ea260666fbcd183e82c10f906b0

Observation 969e79a5-5099-4815-b454-9f78d05a7ada · outbound

This paper cites Geochat: Grounded large vision-language model for remote sensing.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Geochat: Grounded large vision-language model for remote sensing

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.998699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:22.119635Z digest=sha256:240fce5964a8f5faf8358031de82cff0b2a6991388592e1140a2773ff93b5305

Observation 1899628e-54f6-4046-8779-ac2b441d92f0 · outbound

This paper cites Llava-med: Training a large language- and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36, 2024.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Llava-med: Training a large language- and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36, 2024

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.803628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:22.243892Z digest=sha256:7711f74415b48d00f9e9712d1360cdcb162f083ec2ee4e59af2f1483e16c63d9

Observation 0a1756f7-099d-4bc0-81f1-af3980b310af · outbound

This paper cites Urbangpt: Spatio- temporal large language models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urbangpt: Spatio- temporal large language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.628393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:22.380917Z digest=sha256:dd9c2165a2df4ed5c8af738f4f13c483c4096d2f0c489532ec2834073b49e8a9

Observation 7fcd0e7e-db48-4ec4-8f6f-70a57a09de3e · outbound

This paper cites Vila: On pre-training for visual language models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Vila: On pre-training for visual language models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.420034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:22.486258Z digest=sha256:d894693868f4ed8d6c7a0502726cbda4552a586ef6641a970fe7f366027edfaa

Observation e414ca83-90fd-4e0b-a0ae-fe4fa6fc5785 · outbound

This paper cites Improved baselines with visual instruction tuning.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Improved baselines with visual instruction tuning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.256034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:22.603566Z digest=sha256:efb6aacf72756206197040610029a572bfe4bd52dc6bb0a1a8f9e4a44c1c9e46

Observation 774b3427-2c81-4631-9e66-edd2016f0633 · outbound

This paper cites Visual instruction tuning.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Visual instruction tuning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.082401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:22.687759Z digest=sha256:3bec55dec1e05ebb06704b6ddaf44bc89851d36c1b8f1ece810916d6b026e9cc

Observation e39f3b66-cf9c-4426-963e-f6fafe0362c4 · outbound

This paper cites Citylens: Bench- 10 marking large language-vision models for urban socioeco- nomic sensing.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Citylens: Bench- 10 marking large language-vision models for urban socioeco- nomic sensing

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:22.798308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:22.798308Z digest=sha256:81d4140891083b9f11269341c09c9ea2f90ce6180c6bcc11bfeb9d32deb4852d

Observation cf9c645e-b765-4ad7-85d1-1def9183b5d5 · outbound

This paper cites SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:22.909176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:22.909176Z digest=sha256:2fa4daf0fda78679fbb80388a91ea4f4528c27057b108ce4eb1084846c936b46

Observation 6dd443bf-64e6-47c3-be1f-8389e6fb831e · outbound

This paper cites Dolphins: Multimodal Language Model for Driving.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Dolphins: Multimodal Language Model for Driving

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.026538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.026538Z digest=sha256:94767b8eea22702171b0cd5ada7fa7e3a8cf9d65e882a593750e061783908db7

Observation 335ccf75-0269-4d48-831b-104bd6031418 · outbound

This paper cites On the opportunities and chal- lenges of foundation models for geoai (vision paper).

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding On the opportunities and chal- lenges of foundation models for geoai (vision paper)

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.920846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.139438Z digest=sha256:fa17930a2c35e301c13a2404d7079ae7d9ecebc09cefffc7c81523b22d10ae11

Observation 188d55e0-3b17-4ea4-bc96-86149bdc97ce · outbound

This paper cites LLaMA 3.2: Advancing Vision, Edge, and Mo- bile Devices.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding LLaMA 3.2: Advancing Vision, Edge, and Mo- bile Devices

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.727923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.201817Z digest=sha256:22da062a5822a703be7bcb78bb7d62d81c1d34182de188a5c5bb45dde4f8ee9b

Observation 2bc776dd-464a-45e3-b098-c2300c8c73d3 · outbound

This paper cites LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.293073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.293073Z digest=sha256:cb4a5a73ef95c302f2966a3e0ea04a7f4d42472c96df216c1e166c4b4571dab5

Observation bccbedbe-3ea0-47a8-a30b-31564a8b6581 · outbound

This paper cites Introducing chatgpt.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Introducing chatgpt

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.520569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.383898Z digest=sha256:6509ea09eb586be373d28085e48f20c319c3df08a0153a4f4194e7032bd508ee

Observation 8013e578-b61f-4c03-8585-db60f194e6e8 · outbound

This paper cites Gpt-4v(ision) system card.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Gpt-4v(ision) system card

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.306992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.464642Z digest=sha256:6229bf4848c6526a2c8b7fe600fff03e0e3086b9b3110f9f0e19f8815e47c5d8

Observation 35a0cb28-947f-4e9c-842a-a0b494a6e2ea · outbound

This paper cites Hello GPT-4.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Hello GPT-4

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.119217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.579833Z digest=sha256:cd50ec60808f37975dcbdb53f01e63645a7f2da0b8a05f2b22c751c9aa21a341

Observation 985fb899-f005-463c-b1e5-e3499e64c72f · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.674340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.674340Z digest=sha256:f1cd8d9d6fcede976dce0a6aa6f474d22525fc2703f4b2590e3702b66d54946b

Observation bf746aaa-b2a3-428e-94c4-d7883a7ee9c8 · outbound

This paper cites VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language Tasks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language Tasks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.743835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.743835Z digest=sha256:f161d42f92e554b08ef52430406db2ef7615d8fbc83c738a40574f869ecbe061

Observation ed1fafd1-c4d0-420a-9d24-a1223f0adc7b · outbound

This paper cites V*: Guided visual search as a core mechanism in multimodal llms.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding V*: Guided visual search as a core mechanism in multimodal llms

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.950965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.852374Z digest=sha256:3386f0ffa927ac7b33d0f4e676a65e09a08602b446f356ef523046b518f1edeb

Observation 3f187676-0180-47b0-941b-b631ccba57d8 · outbound

This paper cites RealworldQA Dataset.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RealworldQA Dataset

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.803340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:23.943108Z digest=sha256:fbfb2e21123decbb0ec9f0ae602f3c9593d43886cd175e5f7d7ecd9437c2e66b

Observation 83b5e54f-4f66-4e79-a575-7d0fce68ee2b · outbound

This paper cites Analyz- ing large language models’ capability in location prediction.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Analyz- ing large language models’ capability in location prediction

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.663944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.066125Z digest=sha256:3da4748645e4a9cbc831c3e92edfc1f8cdc6a3fc90f288d21aba638087771cd9

Observation 8dc2dfb8-1c7c-4968-8c4e-506a1669c93f · outbound

This paper cites Urban Generative Intelligence (UGI): A Foundational Platform for Agents in Embodied City Environment.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urban Generative Intelligence (UGI): A Foundational Platform for Agents in Embodied City Environment

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.182301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.182301Z digest=sha256:c32f83c548009a8ec07f0aede265a9d36ee3ac13c47c9b5483083925b977a9ec

Observation 9df47446-7336-44b2-a638-89adaf931ff2 · outbound

This paper cites Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.272058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.272058Z digest=sha256:7fd9a4a9e5fbada4c877b802f778cea1c553bb3fa527d67238fb045c5b854f4e

Observation a42cc69e-bb69-44fa-b228-7a09bb0316a3 · outbound

This paper cites Par- ticipatory cultural mapping based on collective behavior data in location-based social networks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Par- ticipatory cultural mapping based on collective behavior data in location-based social networks

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.479881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.356229Z digest=sha256:298aedb0ae049228169b68b27111706e47342c6d00cbee7987ad1799effe5f2d

Observation af2619d3-6b9d-4e69-a7fa-59a751aa8b8f · outbound

This paper cites A Survey on Multimodal Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding A Survey on Multimodal Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.419530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.419530Z digest=sha256:1535508f3a60e197981799b241cbd61aea6e824c53213e30692bbe29792542eb

Observation cd0b7066-f87c-4c05-adf0-3d98eb6d0e1d · outbound

This paper cites Mm-vet: Evaluating large multimodal models for inte- grated capabilities.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Mm-vet: Evaluating large multimodal models for inte- grated capabilities

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.296895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.484165Z digest=sha256:d058b48c1f48fcdee543d0be1965b7d5534c7fbe01191d446b13e6f3027b799d

Observation 4ea88e8a-8db0-4ce1-8d73-77ef38fac776 · outbound

This paper cites SkyEyeGPT: Unifying Remote Sensing Vision-Language Tasks via Instruction Tuning with Large Language Model.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding SkyEyeGPT: Unifying Remote Sensing Vision-Language Tasks via Instruction Tuning with Large Language Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.557582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.557582Z digest=sha256:aa5c9c1275a1e2e057ce98f5bd309af5fd8fa13dbe371db38f2b3ae70eaff0c8

Observation fb0d3055-79a3-4d39-af6d-94273c5476cc · outbound

This paper cites Earthgpt: A universal multi-modal large lan- guage model for multi-sensor image comprehension in re- mote sensing domain.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Earthgpt: A universal multi-modal large lan- guage model for multi-sensor image comprehension in re- mote sensing domain

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.120055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.628535Z digest=sha256:8a66abba1b6e7f44f950c51e3f24a0a51ddb1526c28f7b8c24b3aa7955b88eb5

Observation 88ad8c4c-0f1e-420f-8a82-ee386e1cc1c4 · outbound

This paper cites Urban foundation models: A survey.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urban foundation models: A survey

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.951040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.672320Z digest=sha256:3accfcefebc5bb5d1dafc5b02cecce03539430225607095475e7e9cc392f8afd

Observation f16343f6-e71c-4829-b6df-6ff60778e59a · outbound

This paper cites UrbanMLLM: Joint learning of cross-view imagery for urban understanding, 2025.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding UrbanMLLM: Joint learning of cross-view imagery for urban understanding, 2025

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.815763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.721919Z digest=sha256:db270879564d06d2a3b49383fb4b3854a638d42ab140435cc3a683700954b1eb

Observation 5f8b831f-e89b-4263-a0c4-c25564bd2048 · outbound

This paper cites Per- ceiving urban inequality from imagery using visual language models with chain-of-thought reasoning.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Per- ceiving urban inequality from imagery using visual language models with chain-of-thought reasoning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.618477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.801393Z digest=sha256:57085210e76a75f720d69fcdb1556d0402624f49f85f6f3f6a3129c7fea8da93

Observation d2bfc31c-65af-4250-886b-738b385c421d · outbound

This paper cites Urbench: A comprehensive bench- mark for evaluating large multimodal models in multi-view urban scenarios.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urbench: A comprehensive bench- mark for evaluating large multimodal models in multi-view urban scenarios

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.440727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.854721Z digest=sha256:e43e437acccbfdcc9ece34f6ee43b756c632c550a1abfe8251edbb8db97e32a7

Observation 29ccec06-f7c7-4493-9da2-98a5f2a90a87 · outbound

This paper cites Deep learning for cross-domain data fu- sion in urban computing: Taxonomy, advances, and outlook.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Deep learning for cross-domain data fu- sion in urban computing: Taxonomy, advances, and outlook

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.257763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.894192Z digest=sha256:e20bdcbe5a8a47715384194b1523bd4a8c071628895065e5b4c88e3f194a6d8d

Observation ef3ed754-043d-4ed6-bdfa-550a9d2985bb · outbound

This paper cites Figure 9.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Figure 9

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.008529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:24.963116Z digest=sha256:a2505205e27b053f292629c997ca7a0106fe1d2c3bc7f1363fb08e5264c0d1bc

Observation f83a9c12-7d2c-48a3-a27f-632c538467da · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:28.721388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.019923Z digest=sha256:67794d92ed2e1514040be286e0f4d62255f58c2acf8777b1ecf8570505988880

Observation 69414709-d8dc-4be6-ad26-cff32443ebda · outbound

This paper cites Table 2 in Section 3.2 is the aggregated results of these three tables.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Table 2 in Section 3.2 is the aggregated results of these three tables

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:28.498738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.067667Z digest=sha256:c43424e942e4ba7ee82a31b38dd1d1b4e7e4d56f5b869abe724ed65f3f7ee1d7

Observation 422b2167-4930-4227-b49b-cd6c734d3fad · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:28.278570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.119209Z digest=sha256:441b7cf552661e4e8f8bfc12357d1cb5c2ba7ac6cf213e2c4e97be806cfdc2c7

Observation fd295feb-c4b9-4d64-964b-22449050b58c · outbound

This paper cites 11 presents training results with different amounts, ex- hibiting the high quality of UData.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding 11 presents training results with different amounts, ex- hibiting the high quality of UData

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:27.972132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.203938Z digest=sha256:5c7ec227be8badf697baf75bb47f0abf46a59c373d5720793daa4b57e3663c1e

Observation 8344fae6-6c4b-4b8f-a970-3f0b5373f535 · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:27.625332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.253531Z digest=sha256:2ebc54e7d3091d698d900df4048d4aa8ea1bc20bbfe0e02dcbe18205258844b1

Observation 7385252b-4d0d-42e6-8b30-6911e05c794a · outbound

This paper cites However, for certain tasks, models of different sizes exhibit similar capabilities.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding However, for certain tasks, models of different sizes exhibit similar capabilities

Reference 64

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T21:52:27.351086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.315959Z digest=sha256:4dc271659d9b14a000330f9bf617780aeae8de6a44dbb71169231a53f1144225

Observation bd8fd18b-f630-4512-a500-92c471c491ba · outbound

This paper cites This task needs a model to speculate the land use type (commercial, residential, agricultural, etc.) based on a satellite image.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding This task needs a model to speculate the land use type (commercial, residential, agricultural, etc.) based on a satellite image

Reference 65

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T21:52:27.097941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.382008Z digest=sha256:062279b6ec2fd041ffedea51ea25113623fdcf0238fa574c7ae239063eac5cbf

Observation 272f3bd7-3e1a-488c-b297-f5dd07e3e998 · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:26.762359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:52:25.446679Z digest=sha256:c68ac65e922f9b7399e95a5179109046969222d18cd27ce9b2071440477540db

Pith citing papers

Observation eba8392f-a4ca-4c0c-b565-bdf2db73376e · inbound

IoT-Brain: Grounding LLMs for Semantic-Spatial Sensor Scheduling cites this paper.

IoT-Brain: Grounding LLMs for Semantic-Spatial Sensor Scheduling UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:05:56.875660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T17:49:16.409834Z digest=sha256:24b43a9c8dfd2c2ffd19a62dcdda98d3612c5e8d78b7ceacb017181b3873cbde

Observation 4a8bd704-34ac-4f67-ab3b-97c1b8c68f6c · inbound

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models cites this paper.

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:26:45.810959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T06:58:39.117228Z digest=sha256:3aba2a8b01630f3feb034585e275e10572c26639e8691515b553578025af4157