Pith. sign in

Paper Citation Record · LEDGER

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding

As of 15 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 2 inbound Pith citation observations for arXiv:2506.23219.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.23219 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:52:25.446679Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T06:58:39.117228Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T07:26:45.809288Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact1
  • verified fuzzy35
  • unresolved28
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a66870f0-4af1-4c6e-b323-672ffb45e0fd · outbound

This paper cites LAMP: A Language Model on the Map.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding LAMP: A Language Model on the Map

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:19.753815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:19.753815Z digest=sha256:783122e415791c68ef6038c89fd91dd414e7cb966de00b87d4a2bbc7af2cd9fb

Observation 69493768-4882-459e-b673-245ff3654219 · outbound

This paper cites City foundation models for learning general purpose rep- resentations from openstreetmap.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding City foundation models for learning general purpose rep- resentations from openstreetmap

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.883775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:19.839179Z digest=sha256:6d1dfbf8073b79187d49319eaf479ac4e6b3cf1e18417018498608274a9480a5

Observation 983a5a70-1d8b-48e3-99a8-3f7c8161bcbc · outbound

This paper cites Street view imagery in urban analytics and gis: A review.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Street view imagery in urban analytics and gis: A review

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.689303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:19.924010Z digest=sha256:d03bb456c2c5a9e809b0de5ef40fffe6b3892a38408c19d5151f3b581eda9968

Observation 10e1defb-e8da-4ac3-88bc-9edc797d602c · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.060217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.060217Z digest=sha256:d63055eb7fa862727d32269d5339882db7e43a807a2d4c5fe4898b695b4c5530

Observation bfce9a89-df1e-437e-b8b9-a9c1ed8543ae · outbound

This paper cites Touchdown: Natural language naviga- tion and spatial reasoning in visual street environments.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Touchdown: Natural language naviga- tion and spatial reasoning in visual street environments

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.487883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:20.178492Z digest=sha256:f28f6f2653508f2c727ef7e32b9f495174b7ffc80300c1059062dc8fb45e10a2

Observation c0282037-4302-4a00-acf3-f66c822daa7d · outbound

This paper cites ShareGPT4V: Improving Large Multi-Modal Models with Better Captions.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.279386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.279386Z digest=sha256:9a65c29f1cb36c64d98a30952c8c5cf58349940df3ae7fae657e1dc66b5d6596

Observation f5b2046a-852f-4c85-b199-852e4af903d3 · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.373110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.373110Z digest=sha256:76c2e75260aee7eb315a1dfc054475fc5b53ee44506a61fd5237949cac3e066b

Observation 44fc0a9b-b231-4285-8142-b90368f7e855 · outbound

This paper cites Internvl: Scaling up vision founda- tion models and aligning for generic visual-linguistic tasks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Internvl: Scaling up vision founda- tion models and aligning for generic visual-linguistic tasks

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.338270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:20.462260Z digest=sha256:f9bd97306697d377907631f7c805eb5b583a887945c077da84673c57a547d869

Observation ca1bd479-40fe-484f-a892-1beea35e5999 · outbound

This paper cites SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.579394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.579394Z digest=sha256:3460e52df0fa887bcbb58cfc944d69514919f31f8f8bbd2e90ecdcb2b8e1a4a3

Observation d5635dfc-a5bd-429a-a4ca-67bd4b671ad8 · outbound

This paper cites Understanding world or predict- ing future? a comprehensive survey of world models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Understanding world or predict- ing future? a comprehensive survey of world models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.691552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.691552Z digest=sha256:3cbf055aa4e404cce5eb26f33e2e645809a124940e18cbad9dd356d53a1e4fc2

Observation f8f69a58-ccc7-48f3-94c7-4d997701a538 · outbound

This paper cites Enhancing Chat Language Models by Scaling High-quality Instructional Conversations.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Enhancing Chat Language Models by Scaling High-quality Instructional Conversations

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.776159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.776159Z digest=sha256:29784efe30e85036aa26719d13a018d2652d79b637d4aee8bc21e9164d2cdd9e

Observation b09e2d68-1545-4e28-b366-9000934f1574 · outbound

This paper cites How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.860831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.860831Z digest=sha256:0994fca7c8a89decc2d17dabe96b32e014f5817b81e8b87da38823b02e3deb2c

Observation 33b27999-3ad1-4f5a-ad15-fbb09880377a · outbound

This paper cites Vlmevalkit: An open- source toolkit for evaluating large multi-modality models,.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Vlmevalkit: An open- source toolkit for evaluating large multi-modality models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.969557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.969557Z digest=sha256:a776823d5839b1f1b1351b03a200d81c9a50bd893091e50adb8a17b8bdb318e0

Observation 1b813eca-e9b5-4467-99db-6e324c9003b4 · outbound

This paper cites Urban visual intelligence: Uncovering hidden city pro- files with street view images.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urban visual intelligence: Uncovering hidden city pro- files with street view images

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.158394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.035326Z digest=sha256:b7b2ccbb624504535669d3ff3c7b66ceabb398d2ff1bda9b2ac75c3a3767241d

Observation d719cf63-a281-4dc5-b9a6-fbf82669f95c · outbound

This paper cites Agent- move: A large language model based agentic framework for zero-shot next location prediction.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Agent- move: A large language model based agentic framework for zero-shot next location prediction

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.016597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.152105Z digest=sha256:0f996d94fe26a03b60eb457032a9de13b0ac78124c11a0f0be2e72320d8fa559

Observation 5ab581ac-0389-4ed1-a2ef-94ca7df1bca1 · outbound

This paper cites Citygpt: Empowering urban spatial cognition of large language models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Citygpt: Empowering urban spatial cognition of large language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.818015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.237329Z digest=sha256:109c58a9d82a26beae9705e260c5761c5f6a3eda34abf271db3f1cd922b528b6

Observation a030dcfb-3b62-4f79-8cd5-b941de0a9c60 · outbound

This paper cites A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.335105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.335105Z digest=sha256:0f4aa661283c83deb66a9eba784d0b1ffcc0ddbd84bd591213d6ba16024b216c

Observation 50e01ba8-4998-46e5-b4d7-22fb6e6410dc · outbound

This paper cites City- bench: Evaluating the capabilities of large language models for urban tasks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding City- bench: Evaluating the capabilities of large language models for urban tasks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.556275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.403131Z digest=sha256:5f847d868cf18c36766b06771c94526a50cb237bb6f135817dcd56a685d9bc35

Observation dc817789-0086-487f-a05f-52dacc81f7ab · outbound

This paper cites Imagebind: One embedding space to bind them all.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Imagebind: One embedding space to bind them all

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.502930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.502930Z digest=sha256:5ca88994028d405797cd3e5cf05be4015d54d7805aa9346fbd58536e7ed6d2a2

Observation 4f1933ac-cb85-483b-b11f-7834f2a98765 · outbound

This paper cites Mobility-LLM: Learning Visiting Intentions and Travel Preferences from Human Mobility Data with Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Mobility-LLM: Learning Visiting Intentions and Travel Preferences from Human Mobility Data with Large Language Models

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:52:26.087467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.581742Z digest=sha256:fc13593573543834763f4422fc178fd228f45098759572c01d71d32839b8fa17

Observation f61311cd-f3e4-4df6-b205-1b23b4838077 · outbound

This paper cites Regiongpt: Towards region understanding vision lan- guage model.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Regiongpt: Towards region understanding vision lan- guage model

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.357391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.665075Z digest=sha256:eb54f49176243378a800a73548231e9bcf6628a130878eb2d2b4d0a832349e4a

Observation 57e99a67-4c0a-4127-8b1c-d71f82485cd5 · outbound

This paper cites UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.758851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.758851Z digest=sha256:eb03f5678e9b92d9501ed97bd9385fdb75c5d8e6b3239facf0639e1f2876d7c1

Observation 8e59aa17-e176-46fe-b9ea-ac00f6d93b75 · outbound

This paper cites Vision-language models for medical report generation and visual question answering: A review, 2024.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Vision-language models for medical report generation and visual question answering: A review, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.222258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.863618Z digest=sha256:2c1d32fbb6cefad13a48c2691030a1896a065d5fc943b7344746d7bc43bfa3b3

Observation 89386da8-4e18-4d7c-801e-a4d5714a0262 · outbound

This paper cites RSGPT: A Remote Sensing Vision Language Model and Benchmark.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.928081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.928081Z digest=sha256:0a89eec55617c3d9af3c8cf8d24fbd7af880f849027f7dea9c02488c8c0b9c80

Observation 4c330697-6e82-442f-9a38-205e3a50ed8c · outbound

This paper cites Time-LLM: Time Series Forecasting by Reprogramming Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Time-LLM: Time Series Forecasting by Reprogramming Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:22.043641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:22.043641Z digest=sha256:e44e206b4a50d5f49a20460d571cced31f31d311ef7377c0936a449bb1ff159c

Observation 969e79a5-5099-4815-b454-9f78d05a7ada · outbound

This paper cites Geochat: Grounded large vision-language model for remote sensing.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Geochat: Grounded large vision-language model for remote sensing

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.998699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:22.119635Z digest=sha256:b9cca552a88b9fba2ae8423f1f3698622f48a8674a1e44c52cb0fe46551f298b

Observation 1899628e-54f6-4046-8779-ac2b441d92f0 · outbound

This paper cites Llava-med: Training a large language- and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36, 2024.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Llava-med: Training a large language- and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36, 2024

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.803628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:22.243892Z digest=sha256:ae05af29f7d8d9f5f320156b6047d57d7de2abb1bafa2f5c1153648501f8437c

Observation 0a1756f7-099d-4bc0-81f1-af3980b310af · outbound

This paper cites Urbangpt: Spatio- temporal large language models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urbangpt: Spatio- temporal large language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.628393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:22.380917Z digest=sha256:5326d7c10f7a5b3825855ecb59e2cc9f9068c472d4b7be3c21e1eecc1c765a20

Observation 7fcd0e7e-db48-4ec4-8f6f-70a57a09de3e · outbound

This paper cites Vila: On pre-training for visual language models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Vila: On pre-training for visual language models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.420034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:22.486258Z digest=sha256:268c0a95a9eb88039d2824fa80d7fea8f584d5e5e176624962c1042baac06592

Observation e414ca83-90fd-4e0b-a0ae-fe4fa6fc5785 · outbound

This paper cites Improved baselines with visual instruction tuning.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Improved baselines with visual instruction tuning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.256034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:22.603566Z digest=sha256:d14532b8ac35bd9386acb7336b82f7b81346400973c03b440f3a1ffb3bb815b0

Observation 774b3427-2c81-4631-9e66-edd2016f0633 · outbound

This paper cites Visual instruction tuning.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Visual instruction tuning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.082401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:22.687759Z digest=sha256:46e0c1077a3399b9feb407be25ca479fa7648a8689854d1249cc7a47e932177b

Observation e39f3b66-cf9c-4426-963e-f6fafe0362c4 · outbound

This paper cites Citylens: Bench- 10 marking large language-vision models for urban socioeco- nomic sensing.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Citylens: Bench- 10 marking large language-vision models for urban socioeco- nomic sensing

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:22.798308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:22.798308Z digest=sha256:6eef582b0e96dabf27d0e1dddf84abe91abf0bc2c2630d3f8aea95a76edf29e4

Observation cf9c645e-b765-4ad7-85d1-1def9183b5d5 · outbound

This paper cites SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:22.909176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:22.909176Z digest=sha256:141832ec622e876d96f6eab3e87de4c58c17feddb50330b9bf6dd1a146b26147

Observation 6dd443bf-64e6-47c3-be1f-8389e6fb831e · outbound

This paper cites Dolphins: Multimodal Language Model for Driving.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Dolphins: Multimodal Language Model for Driving

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.026538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.026538Z digest=sha256:9647c5a11585de4d535fed1426069697e260bcf0cbcc1513d535e99c1a54c13e

Observation 335ccf75-0269-4d48-831b-104bd6031418 · outbound

This paper cites On the opportunities and chal- lenges of foundation models for geoai (vision paper).

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding On the opportunities and chal- lenges of foundation models for geoai (vision paper)

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.920846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.139438Z digest=sha256:b3489009275f8fdb8b954d47146496f0e090cc0e1a761b4d7102fa38a15f7e2d

Observation 188d55e0-3b17-4ea4-bc96-86149bdc97ce · outbound

This paper cites LLaMA 3.2: Advancing Vision, Edge, and Mo- bile Devices.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding LLaMA 3.2: Advancing Vision, Edge, and Mo- bile Devices

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.727923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.201817Z digest=sha256:104e981c8d155cdfa397ae1c5f1e001b287d3f3621b00d304514412794234e66

Observation 2bc776dd-464a-45e3-b098-c2300c8c73d3 · outbound

This paper cites LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.293073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.293073Z digest=sha256:1ab32530c744e045c6fc2e5dd182505f600b37def15afbeb98df9a1fb224943a

Observation bccbedbe-3ea0-47a8-a30b-31564a8b6581 · outbound

This paper cites Introducing chatgpt.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Introducing chatgpt

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.520569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.383898Z digest=sha256:4808ee051173ba069508a34ffeabb560ff0f45b9e9fc39a15e84fd96bb3d9fc8

Observation 8013e578-b61f-4c03-8585-db60f194e6e8 · outbound

This paper cites Gpt-4v(ision) system card.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Gpt-4v(ision) system card

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.306992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.464642Z digest=sha256:ecdfbf2e83cfb5e32642c9824fcf8f907b0827ef1425d428509ea23d1ee126d9

Observation 35a0cb28-947f-4e9c-842a-a0b494a6e2ea · outbound

This paper cites Hello GPT-4.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Hello GPT-4

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.119217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.579833Z digest=sha256:1024e4f064aab980482ea18c01699155a69dc784a8cfa5df34babb35afacdc56

Observation 985fb899-f005-463c-b1e5-e3499e64c72f · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.674340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.674340Z digest=sha256:29a871ea0639d712969b58a7a8515d43138dc89c362fb97e4a0a5471d29bb253

Observation bf746aaa-b2a3-428e-94c4-d7883a7ee9c8 · outbound

This paper cites VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language Tasks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language Tasks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.743835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.743835Z digest=sha256:b25f1a2b37e48cf007268603224a588e594668ad42dc815bc2c71940133f6874

Observation ed1fafd1-c4d0-420a-9d24-a1223f0adc7b · outbound

This paper cites V*: Guided visual search as a core mechanism in multimodal llms.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding V*: Guided visual search as a core mechanism in multimodal llms

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.950965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.852374Z digest=sha256:a3178a93dd2ee4add4e4a6a288f7e8ea00ae71e4faab39c6821056934bf87497

Observation 3f187676-0180-47b0-941b-b631ccba57d8 · outbound

This paper cites RealworldQA Dataset.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RealworldQA Dataset

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.803340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.943108Z digest=sha256:c3aedd097277f9268edd875839301c9a236e9df9adebb5e853dbf7d392c4a26c

Observation 83b5e54f-4f66-4e79-a575-7d0fce68ee2b · outbound

This paper cites Analyz- ing large language models’ capability in location prediction.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Analyz- ing large language models’ capability in location prediction

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.663944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.066125Z digest=sha256:0eccf6f7f1299821efab6570ac9f0d3466f1e5c471795bd8d6301c805623285f

Observation 8dc2dfb8-1c7c-4968-8c4e-506a1669c93f · outbound

This paper cites Urban Generative Intelligence (UGI): A Foundational Platform for Agents in Embodied City Environment.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urban Generative Intelligence (UGI): A Foundational Platform for Agents in Embodied City Environment

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.182301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.182301Z digest=sha256:630a14eae9152592f90e62e3e84f1ce89af74b7432178541e9745e389c970ca0

Observation 9df47446-7336-44b2-a638-89adaf931ff2 · outbound

This paper cites Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.272058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.272058Z digest=sha256:d4ad1b0e243ae1476bb2899b4f89914d1caeb2f7cab9961d3bcf4088b40e559a

Observation a42cc69e-bb69-44fa-b228-7a09bb0316a3 · outbound

This paper cites Par- ticipatory cultural mapping based on collective behavior data in location-based social networks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Par- ticipatory cultural mapping based on collective behavior data in location-based social networks

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.479881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.356229Z digest=sha256:ef5a5f7448da007d87ead79e266a6ac54610b69912deed1b5295bc24211e54b7

Observation af2619d3-6b9d-4e69-a7fa-59a751aa8b8f · outbound

This paper cites A Survey on Multimodal Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding A Survey on Multimodal Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.419530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.419530Z digest=sha256:933b721fe56a0e79158c40d4d8b5296bb9f4fb10dccdabe619b5d19770b21e26

Observation cd0b7066-f87c-4c05-adf0-3d98eb6d0e1d · outbound

This paper cites Mm-vet: Evaluating large multimodal models for inte- grated capabilities.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Mm-vet: Evaluating large multimodal models for inte- grated capabilities

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.296895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.484165Z digest=sha256:52641a99e138d8984d4dbe1ae9b8027c442dcc5749e5a0529705e62e3aed11b7

Observation 4ea88e8a-8db0-4ce1-8d73-77ef38fac776 · outbound

This paper cites SkyEyeGPT: Unifying Remote Sensing Vision-Language Tasks via Instruction Tuning with Large Language Model.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding SkyEyeGPT: Unifying Remote Sensing Vision-Language Tasks via Instruction Tuning with Large Language Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.557582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.557582Z digest=sha256:d9882aa29f005301bf75dd9fdc805fa39e074d32135933470984c51f03641a21

Observation fb0d3055-79a3-4d39-af6d-94273c5476cc · outbound

This paper cites Earthgpt: A universal multi-modal large lan- guage model for multi-sensor image comprehension in re- mote sensing domain.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Earthgpt: A universal multi-modal large lan- guage model for multi-sensor image comprehension in re- mote sensing domain

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.120055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.628535Z digest=sha256:1ee9394d750ea5ab8198d9c7bdacb275710a905db8b2ac82703c3909bac74bb0

Observation 88ad8c4c-0f1e-420f-8a82-ee386e1cc1c4 · outbound

This paper cites Urban foundation models: A survey.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urban foundation models: A survey

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.951040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.672320Z digest=sha256:0fe0ea3648ec4e962dde98833b1fdac3faa54ea9f3f83b466eab0a2ab59f2a14

Observation f16343f6-e71c-4829-b6df-6ff60778e59a · outbound

This paper cites UrbanMLLM: Joint learning of cross-view imagery for urban understanding, 2025.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding UrbanMLLM: Joint learning of cross-view imagery for urban understanding, 2025

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.815763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.721919Z digest=sha256:bb8e4e9c9e753dffcc224feee120dcc3f6fca2c664c4a82da10b88449b4e5822

Observation 5f8b831f-e89b-4263-a0c4-c25564bd2048 · outbound

This paper cites Per- ceiving urban inequality from imagery using visual language models with chain-of-thought reasoning.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Per- ceiving urban inequality from imagery using visual language models with chain-of-thought reasoning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.618477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.801393Z digest=sha256:0eecfb06dc8da060e9c8aabbf7bff5ffe2977a60d25960c04f161bd0361fa77f

Observation d2bfc31c-65af-4250-886b-738b385c421d · outbound

This paper cites Urbench: A comprehensive bench- mark for evaluating large multimodal models in multi-view urban scenarios.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urbench: A comprehensive bench- mark for evaluating large multimodal models in multi-view urban scenarios

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.440727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.854721Z digest=sha256:58b2f468d2c0f72a9b312e1b2ec17496dcf9cdc33dd335a630d7d4b20b4a5961

Observation 29ccec06-f7c7-4493-9da2-98a5f2a90a87 · outbound

This paper cites Deep learning for cross-domain data fu- sion in urban computing: Taxonomy, advances, and outlook.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Deep learning for cross-domain data fu- sion in urban computing: Taxonomy, advances, and outlook

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.257763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.894192Z digest=sha256:a08232004391a69c82e2a24505d3f8366a45a23e1deff5f64b42b289c508f154

Observation ef3ed754-043d-4ed6-bdfa-550a9d2985bb · outbound

This paper cites Figure 9.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Figure 9

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.008529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.963116Z digest=sha256:58412909f87497690d0c94e026974ff7b68e99ed38b19249ae1e10959cf369f9

Observation f83a9c12-7d2c-48a3-a27f-632c538467da · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:28.721388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.019923Z digest=sha256:4f4f669886793c500bc00be69f1f6be369bf9effccdd61047f1a3bfd57f8b22c

Observation 69414709-d8dc-4be6-ad26-cff32443ebda · outbound

This paper cites Table 2 in Section 3.2 is the aggregated results of these three tables.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Table 2 in Section 3.2 is the aggregated results of these three tables

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:28.498738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.067667Z digest=sha256:c7805f67e1e4f3f3a8206f79a0fd82b73834954503669a25ba371aa81d571fb9

Observation 422b2167-4930-4227-b49b-cd6c734d3fad · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:28.278570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.119209Z digest=sha256:130e025f2a28df5f3b3f3849fb1d0cc0b528170fcbe7af88bd87a40f85ac2064

Observation fd295feb-c4b9-4d64-964b-22449050b58c · outbound

This paper cites 11 presents training results with different amounts, ex- hibiting the high quality of UData.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding 11 presents training results with different amounts, ex- hibiting the high quality of UData

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:27.972132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.203938Z digest=sha256:22fb01f01a9a02df87f9619df631b28d07a3f032d2bbaca04ebffc2a49d2b527

Observation 8344fae6-6c4b-4b8f-a970-3f0b5373f535 · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:27.625332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.253531Z digest=sha256:dea8b2112b40de2a616f87ea7b6848302062d8ab13f972dabc3e593b8ec1fb91

Observation 7385252b-4d0d-42e6-8b30-6911e05c794a · outbound

This paper cites However, for certain tasks, models of different sizes exhibit similar capabilities.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding However, for certain tasks, models of different sizes exhibit similar capabilities

Reference 64

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T21:52:27.351086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.315959Z digest=sha256:f417e5c0d4567442d75fded9eaa1364e088c44aed45a1b002c111dd7ea1abb9b

Observation bd8fd18b-f630-4512-a500-92c471c491ba · outbound

This paper cites This task needs a model to speculate the land use type (commercial, residential, agricultural, etc.) based on a satellite image.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding This task needs a model to speculate the land use type (commercial, residential, agricultural, etc.) based on a satellite image

Reference 65

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T21:52:27.097941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.382008Z digest=sha256:c02b1913bdea8f53af30d8c789434c4db25ca4e7839e44cacdc5cade5d20f931

Observation 272f3bd7-3e1a-488c-b297-f5dd07e3e998 · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:26.762359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.446679Z digest=sha256:4910a210073bf765a1d66bb7966e573c2c744d4c863039d985c87798e627d5a6

Pith citing papers

Observation eba8392f-a4ca-4c0c-b565-bdf2db73376e · inbound

IoT-Brain: Grounding LLMs for Semantic-Spatial Sensor Scheduling cites this paper.

IoT-Brain: Grounding LLMs for Semantic-Spatial Sensor Scheduling UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:05:56.875660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T17:49:16.409834Z digest=sha256:17a1eeb439b4826dbe6a9913de7a4223e304ac68e1c4958b9d5bca0d220d0435

Observation 4a8bd704-34ac-4f67-ab3b-97c1b8c68f6c · inbound

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models cites this paper.

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:26:45.810959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T06:58:39.117228Z digest=sha256:58b80630ebc878114268268684ba316140bde36c61b021ec57e78dac66132c10