Pith. sign in

Paper Citation Record · LEDGER

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding

As of 16 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 2 inbound Pith citation observations for arXiv:2506.23219.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.23219 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:52:25.446679Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T06:58:39.117228Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T07:26:45.809288Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact1
  • verified fuzzy35
  • unresolved28
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a66870f0-4af1-4c6e-b323-672ffb45e0fd · outbound

This paper cites LAMP: A Language Model on the Map.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding LAMP: A Language Model on the Map

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:19.753815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:19.753815Z digest=sha256:4be030845d67d783290f938e75705647e5dcdbbf9d26c32fc76926f00497e0e5

Observation 69493768-4882-459e-b673-245ff3654219 · outbound

This paper cites City foundation models for learning general purpose rep- resentations from openstreetmap.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding City foundation models for learning general purpose rep- resentations from openstreetmap

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.883775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:19.839179Z digest=sha256:db432c3b5425293019c981737c5246fbabb793a9ce67a9fa1a91398649889670

Observation 983a5a70-1d8b-48e3-99a8-3f7c8161bcbc · outbound

This paper cites Street view imagery in urban analytics and gis: A review.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Street view imagery in urban analytics and gis: A review

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.689303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:19.924010Z digest=sha256:d840b5a88c724d0cc14e15524ff9cc592d182f0742b384371b24837cdfc78f9f

Observation 10e1defb-e8da-4ac3-88bc-9edc797d602c · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.060217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.060217Z digest=sha256:425cd9516f455f44f156545afa7c55aab03063ade2a6abdc5ae5d1f7b495cfc0

Observation bfce9a89-df1e-437e-b8b9-a9c1ed8543ae · outbound

This paper cites Touchdown: Natural language naviga- tion and spatial reasoning in visual street environments.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Touchdown: Natural language naviga- tion and spatial reasoning in visual street environments

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.487883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:20.178492Z digest=sha256:f7c3c4d67b4651fe76efef7753fa1f188522175ba4e54b96e431b59099655dd2

Observation c0282037-4302-4a00-acf3-f66c822daa7d · outbound

This paper cites ShareGPT4V: Improving Large Multi-Modal Models with Better Captions.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.279386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.279386Z digest=sha256:015c70123ee96421df74c535193be1fc3a01889f35f967b399d5ba22c4b74e78

Observation f5b2046a-852f-4c85-b199-852e4af903d3 · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.373110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.373110Z digest=sha256:645a272029d3b595cdb2be4f2b8d1a4c06c53ebf0191f81fe188e09a8ee6905c

Observation 44fc0a9b-b231-4285-8142-b90368f7e855 · outbound

This paper cites Internvl: Scaling up vision founda- tion models and aligning for generic visual-linguistic tasks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Internvl: Scaling up vision founda- tion models and aligning for generic visual-linguistic tasks

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.338270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:20.462260Z digest=sha256:a0e84b1f3017ed95fb361cd631cdfc75261363b2340229781e6c19eee83c65ab

Observation ca1bd479-40fe-484f-a892-1beea35e5999 · outbound

This paper cites SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.579394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.579394Z digest=sha256:ab6e6b23557a4f87a1fb8eb27b4ae7a32c24c345ac9cf9f5d77a9add45db69a4

Observation d5635dfc-a5bd-429a-a4ca-67bd4b671ad8 · outbound

This paper cites Understanding world or predict- ing future? a comprehensive survey of world models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Understanding world or predict- ing future? a comprehensive survey of world models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.691552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.691552Z digest=sha256:37d155c202093aaaba8b983b1aabe80fa702771ebd447c922e4e4133e3ff5b1d

Observation f8f69a58-ccc7-48f3-94c7-4d997701a538 · outbound

This paper cites Enhancing Chat Language Models by Scaling High-quality Instructional Conversations.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Enhancing Chat Language Models by Scaling High-quality Instructional Conversations

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.776159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.776159Z digest=sha256:49e6ff7d701f991bf5abf18f19d6bd52a7390e5167668d984a6b04f98c44e112

Observation b09e2d68-1545-4e28-b366-9000934f1574 · outbound

This paper cites How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.860831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.860831Z digest=sha256:f2d1df33d027e6e57d8c7df1b043c60212d65f73d5d3b0b991f4b9e52bfa939c

Observation 33b27999-3ad1-4f5a-ad15-fbb09880377a · outbound

This paper cites Vlmevalkit: An open- source toolkit for evaluating large multi-modality models,.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Vlmevalkit: An open- source toolkit for evaluating large multi-modality models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:20.969557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:20.969557Z digest=sha256:7fd3992b43b8cbe111d5677e7e01640f373f08195ce3a3d2f5a7e7f8995d7817

Observation 1b813eca-e9b5-4467-99db-6e324c9003b4 · outbound

This paper cites Urban visual intelligence: Uncovering hidden city pro- files with street view images.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urban visual intelligence: Uncovering hidden city pro- files with street view images

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.158394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.035326Z digest=sha256:d743ea0491977b3b8924e86c506cdbe3606eb089cb88c00e049fe6c23eadf23e

Observation d719cf63-a281-4dc5-b9a6-fbf82669f95c · outbound

This paper cites Agent- move: A large language model based agentic framework for zero-shot next location prediction.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Agent- move: A large language model based agentic framework for zero-shot next location prediction

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:34.016597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.152105Z digest=sha256:acc0229120346a26458ac813c6cd8cb04cfd9ce00e19f162cd9c161ba73f28fe

Observation 5ab581ac-0389-4ed1-a2ef-94ca7df1bca1 · outbound

This paper cites Citygpt: Empowering urban spatial cognition of large language models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Citygpt: Empowering urban spatial cognition of large language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.818015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.237329Z digest=sha256:1aa160f310c95ad34aeac1297d3c0046e8eb6564124a7f9199bf33a5000708fd

Observation a030dcfb-3b62-4f79-8cd5-b941de0a9c60 · outbound

This paper cites A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.335105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.335105Z digest=sha256:fa6cecb713923f3bb00f01db3499dff703344af0f9e0402753b428330e4be134

Observation 50e01ba8-4998-46e5-b4d7-22fb6e6410dc · outbound

This paper cites City- bench: Evaluating the capabilities of large language models for urban tasks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding City- bench: Evaluating the capabilities of large language models for urban tasks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.556275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.403131Z digest=sha256:c21f27c8e0e3d8e3669e7961071812873fa53568d2e2871e3264b4c27311df2e

Observation dc817789-0086-487f-a05f-52dacc81f7ab · outbound

This paper cites Imagebind: One embedding space to bind them all.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Imagebind: One embedding space to bind them all

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.502930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.502930Z digest=sha256:bcbb37984f64f830fb688b048125ce163c366a55cf568beefd1e620556f1c813

Observation 4f1933ac-cb85-483b-b11f-7834f2a98765 · outbound

This paper cites Mobility-LLM: Learning Visiting Intentions and Travel Preferences from Human Mobility Data with Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Mobility-LLM: Learning Visiting Intentions and Travel Preferences from Human Mobility Data with Large Language Models

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:52:26.087467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.581742Z digest=sha256:609285f853a8db4fb002df919aedb27d37e74ade2c0acbe4087ab13256cd7bfd

Observation f61311cd-f3e4-4df6-b205-1b23b4838077 · outbound

This paper cites Regiongpt: Towards region understanding vision lan- guage model.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Regiongpt: Towards region understanding vision lan- guage model

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.357391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.665075Z digest=sha256:d12cac192add959fd9ceea005e6a1a0aa7006fafaf466f9dfda01b2d7f44befe

Observation 57e99a67-4c0a-4127-8b1c-d71f82485cd5 · outbound

This paper cites UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.758851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.758851Z digest=sha256:c5c0f00cc569f39af466dfbedf92ba10127636d71781574f309a1c9915f4fae7

Observation 8e59aa17-e176-46fe-b9ea-ac00f6d93b75 · outbound

This paper cites Vision-language models for medical report generation and visual question answering: A review, 2024.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Vision-language models for medical report generation and visual question answering: A review, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:33.222258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:21.863618Z digest=sha256:888bae4ad296c4b6d73b108b7b65de0cecf6282aa38e919a74a2da5b2faf9793

Observation 89386da8-4e18-4d7c-801e-a4d5714a0262 · outbound

This paper cites RSGPT: A Remote Sensing Vision Language Model and Benchmark.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RSGPT: A Remote Sensing Vision Language Model and Benchmark

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:21.928081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:21.928081Z digest=sha256:c21f679885574220505e0a2f84a9f8a2a6873f695449daa0ef6055272821d063

Observation 4c330697-6e82-442f-9a38-205e3a50ed8c · outbound

This paper cites Time-LLM: Time Series Forecasting by Reprogramming Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Time-LLM: Time Series Forecasting by Reprogramming Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:22.043641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:22.043641Z digest=sha256:bcdd0e81ed964684792f23cd1f560a0de9f8a6934da5c43ae0f7a6def7b0a31d

Observation 969e79a5-5099-4815-b454-9f78d05a7ada · outbound

This paper cites Geochat: Grounded large vision-language model for remote sensing.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Geochat: Grounded large vision-language model for remote sensing

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.998699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:22.119635Z digest=sha256:3b31a6933e2936c4d0bae5be21298b70b85da8af84ef9d9ae455c5d2263d86e2

Observation 1899628e-54f6-4046-8779-ac2b441d92f0 · outbound

This paper cites Llava-med: Training a large language- and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36, 2024.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Llava-med: Training a large language- and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36, 2024

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.803628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:22.243892Z digest=sha256:f71c6967e2f22767e0f7cfe60830365cfb32957f113adf78117de83aa8d85657

Observation 0a1756f7-099d-4bc0-81f1-af3980b310af · outbound

This paper cites Urbangpt: Spatio- temporal large language models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urbangpt: Spatio- temporal large language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.628393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:22.380917Z digest=sha256:7ccdc0db00a5d09f3cce4ca8831108e90831f77c629fbf50cfa56b4e6c8f6b3c

Observation 7fcd0e7e-db48-4ec4-8f6f-70a57a09de3e · outbound

This paper cites Vila: On pre-training for visual language models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Vila: On pre-training for visual language models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.420034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:22.486258Z digest=sha256:557d7ad13f8c82c2f06cb320ada1b4cf1964d685dedc9c3504fa3548c1fa4282

Observation e414ca83-90fd-4e0b-a0ae-fe4fa6fc5785 · outbound

This paper cites Improved baselines with visual instruction tuning.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Improved baselines with visual instruction tuning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.256034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:22.603566Z digest=sha256:2daaaf03a60dd80e26443572cf8a6ad836fca3cd94b3fcb824531c6cfbd03b42

Observation 774b3427-2c81-4631-9e66-edd2016f0633 · outbound

This paper cites Visual instruction tuning.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Visual instruction tuning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:32.082401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:22.687759Z digest=sha256:f617f578a4e840256d41948cb6cd38d2799ec8add9b0a6386a2b1c7ac135258f

Observation e39f3b66-cf9c-4426-963e-f6fafe0362c4 · outbound

This paper cites Citylens: Bench- 10 marking large language-vision models for urban socioeco- nomic sensing.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Citylens: Bench- 10 marking large language-vision models for urban socioeco- nomic sensing

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:22.798308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:22.798308Z digest=sha256:751dc9d0a92c6769a3f736cecd250118f299a43ce498d359d43cdfbfbe20828f

Observation cf9c645e-b765-4ad7-85d1-1def9183b5d5 · outbound

This paper cites SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:22.909176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:22.909176Z digest=sha256:183d6051b3b09f50f52647c140ec2adf7f7837f0fad878feeecf4c3783ac8033

Observation 6dd443bf-64e6-47c3-be1f-8389e6fb831e · outbound

This paper cites Dolphins: Multimodal Language Model for Driving.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Dolphins: Multimodal Language Model for Driving

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.026538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.026538Z digest=sha256:5f5c3505011b5f414beba0e235f015b29ad80365c7c94c3d77ea07f1c8704ad3

Observation 335ccf75-0269-4d48-831b-104bd6031418 · outbound

This paper cites On the opportunities and chal- lenges of foundation models for geoai (vision paper).

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding On the opportunities and chal- lenges of foundation models for geoai (vision paper)

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.920846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.139438Z digest=sha256:da751bdf1a5364f5da819b2b5bdefbf469ccd2fce4e6d57274693633d7e74a31

Observation 188d55e0-3b17-4ea4-bc96-86149bdc97ce · outbound

This paper cites LLaMA 3.2: Advancing Vision, Edge, and Mo- bile Devices.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding LLaMA 3.2: Advancing Vision, Edge, and Mo- bile Devices

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.727923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.201817Z digest=sha256:60f186fd571367d9b154b795feec4a8498add348ac296ea91a0cfdf8fb753b8d

Observation 2bc776dd-464a-45e3-b098-c2300c8c73d3 · outbound

This paper cites LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.293073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.293073Z digest=sha256:5887a94a706f26e7014157654eae3b998243f10e8a568ed8402942e6c946d813

Observation bccbedbe-3ea0-47a8-a30b-31564a8b6581 · outbound

This paper cites Introducing chatgpt.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Introducing chatgpt

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.520569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.383898Z digest=sha256:aaf6a39440b105f8950adb40c723273febe5c8bdfe881f5140549959db636fd0

Observation 8013e578-b61f-4c03-8585-db60f194e6e8 · outbound

This paper cites Gpt-4v(ision) system card.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Gpt-4v(ision) system card

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.306992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.464642Z digest=sha256:015f9ab228bfc960fc2dba72face37af265625dc9aa616bd0bb5b38c8582cd47

Observation 35a0cb28-947f-4e9c-842a-a0b494a6e2ea · outbound

This paper cites Hello GPT-4.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Hello GPT-4

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:31.119217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.579833Z digest=sha256:94e1731b5e1249b22f78424367e4f3161a6e8442a0235948b96daff05adceb6e

Observation 985fb899-f005-463c-b1e5-e3499e64c72f · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.674340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.674340Z digest=sha256:129d2bfe006fd9d546085a048d2037df3e8dceebacf5b6645eee9a071906fb17

Observation bf746aaa-b2a3-428e-94c4-d7883a7ee9c8 · outbound

This paper cites VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language Tasks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language Tasks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:23.743835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:23.743835Z digest=sha256:74db8563a6ba7a3f079a76d66a381ff47308ab299f0c2f9c62d5ef917d8e2950

Observation ed1fafd1-c4d0-420a-9d24-a1223f0adc7b · outbound

This paper cites V*: Guided visual search as a core mechanism in multimodal llms.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding V*: Guided visual search as a core mechanism in multimodal llms

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.950965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.852374Z digest=sha256:206716b0be4f0721636950ef96a8d0d3b1125af2c43dbf564343db2142cbd435

Observation 3f187676-0180-47b0-941b-b631ccba57d8 · outbound

This paper cites RealworldQA Dataset.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding RealworldQA Dataset

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.803340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:23.943108Z digest=sha256:882c89d8e832f4f19519bdf7a4cd56ea0d8db7fcf280d1530363cbf6818387b4

Observation 83b5e54f-4f66-4e79-a575-7d0fce68ee2b · outbound

This paper cites Analyz- ing large language models’ capability in location prediction.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Analyz- ing large language models’ capability in location prediction

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.663944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.066125Z digest=sha256:60e5d985f1401379d9fc58a4e7947b5ca15726fc3455897e607262b1c5254203

Observation 8dc2dfb8-1c7c-4968-8c4e-506a1669c93f · outbound

This paper cites Urban Generative Intelligence (UGI): A Foundational Platform for Agents in Embodied City Environment.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urban Generative Intelligence (UGI): A Foundational Platform for Agents in Embodied City Environment

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.182301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.182301Z digest=sha256:24be322629fd0a41feb741486b6f1f2d0d86acb8611a6a75b8b8bb7abae55c70

Observation 9df47446-7336-44b2-a638-89adaf931ff2 · outbound

This paper cites Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.272058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.272058Z digest=sha256:8ad34744b145c9c01facf601d2b39156f08a48b459a5851a8c18ed70edae2d63

Observation a42cc69e-bb69-44fa-b228-7a09bb0316a3 · outbound

This paper cites Par- ticipatory cultural mapping based on collective behavior data in location-based social networks.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Par- ticipatory cultural mapping based on collective behavior data in location-based social networks

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.479881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.356229Z digest=sha256:32934b2fc9f3e62f19ececc5ebb3c5019cd6e0be00d90978f1c59f28c1111bd4

Observation af2619d3-6b9d-4e69-a7fa-59a751aa8b8f · outbound

This paper cites A Survey on Multimodal Large Language Models.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding A Survey on Multimodal Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.419530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.419530Z digest=sha256:cad8c27c30da604835ca1c8932a6597f3734ed531822f82507f125e626995267

Observation cd0b7066-f87c-4c05-adf0-3d98eb6d0e1d · outbound

This paper cites Mm-vet: Evaluating large multimodal models for inte- grated capabilities.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Mm-vet: Evaluating large multimodal models for inte- grated capabilities

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.296895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.484165Z digest=sha256:52dfa33028d703210c7d940bee1277c010aa98f24bce464a88dee282e8a6b076

Observation 4ea88e8a-8db0-4ce1-8d73-77ef38fac776 · outbound

This paper cites SkyEyeGPT: Unifying Remote Sensing Vision-Language Tasks via Instruction Tuning with Large Language Model.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding SkyEyeGPT: Unifying Remote Sensing Vision-Language Tasks via Instruction Tuning with Large Language Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.557582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.557582Z digest=sha256:66499e191251b5043d8f356d9d86c59b1d335c5002d8b2f436b19b9f487bdf3f

Observation fb0d3055-79a3-4d39-af6d-94273c5476cc · outbound

This paper cites Earthgpt: A universal multi-modal large lan- guage model for multi-sensor image comprehension in re- mote sensing domain.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Earthgpt: A universal multi-modal large lan- guage model for multi-sensor image comprehension in re- mote sensing domain

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:30.120055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.628535Z digest=sha256:a52c5480e8a7b236138b3673b3bccbb48f9a2531d6df3b57d66fb54336176fe3

Observation 88ad8c4c-0f1e-420f-8a82-ee386e1cc1c4 · outbound

This paper cites Urban foundation models: A survey.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urban foundation models: A survey

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.951040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.672320Z digest=sha256:f86428cd41580a22f74008b4a55b296bc19bd9857019f34bcf1d18baabc570df

Observation f16343f6-e71c-4829-b6df-6ff60778e59a · outbound

This paper cites UrbanMLLM: Joint learning of cross-view imagery for urban understanding, 2025.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding UrbanMLLM: Joint learning of cross-view imagery for urban understanding, 2025

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.815763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.721919Z digest=sha256:cc5f785c287c460bad2beb62f003ecbe50821dba289f32c9692d2ad978d8d5f7

Observation 5f8b831f-e89b-4263-a0c4-c25564bd2048 · outbound

This paper cites Per- ceiving urban inequality from imagery using visual language models with chain-of-thought reasoning.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Per- ceiving urban inequality from imagery using visual language models with chain-of-thought reasoning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.618477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.801393Z digest=sha256:af4d27e043a6d3fa5c50eab9c6482c7dc64b7b43b3fcd73ee7f40f45144f0a05

Observation d2bfc31c-65af-4250-886b-738b385c421d · outbound

This paper cites Urbench: A comprehensive bench- mark for evaluating large multimodal models in multi-view urban scenarios.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Urbench: A comprehensive bench- mark for evaluating large multimodal models in multi-view urban scenarios

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.440727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.854721Z digest=sha256:cdc1453b596ee59efad72b07a015ba963e1b067930903e63793536058739cdec

Observation 29ccec06-f7c7-4493-9da2-98a5f2a90a87 · outbound

This paper cites Deep learning for cross-domain data fu- sion in urban computing: Taxonomy, advances, and outlook.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Deep learning for cross-domain data fu- sion in urban computing: Taxonomy, advances, and outlook

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.257763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.894192Z digest=sha256:0d58c434cadb964c77002848e76f2912425b7c9a8795ff8d7cb3688c9ec3c9cf

Observation ef3ed754-043d-4ed6-bdfa-550a9d2985bb · outbound

This paper cites Figure 9.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Figure 9

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:29.008529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:24.963116Z digest=sha256:d25a644f41f7d047b4f2f77596b6f0c6efc103e997c25f66f2f44b0b69e13f45

Observation f83a9c12-7d2c-48a3-a27f-632c538467da · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:28.721388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.019923Z digest=sha256:97bad68a1281754c9b65a596faab528ee27c39bef67ab098662daf5c93d89a1a

Observation 69414709-d8dc-4be6-ad26-cff32443ebda · outbound

This paper cites Table 2 in Section 3.2 is the aggregated results of these three tables.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Table 2 in Section 3.2 is the aggregated results of these three tables

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:28.498738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.067667Z digest=sha256:3c484a131fac702b5131a678708389077e2f4e1919e13b815cec658c61e208ed

Observation 422b2167-4930-4227-b49b-cd6c734d3fad · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:28.278570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.119209Z digest=sha256:fda89f1192a94b8ea23884a78f2b5e6aa09c7bcf94ab7a871e1302148202b48f

Observation fd295feb-c4b9-4d64-964b-22449050b58c · outbound

This paper cites 11 presents training results with different amounts, ex- hibiting the high quality of UData.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding 11 presents training results with different amounts, ex- hibiting the high quality of UData

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:52:27.972132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.203938Z digest=sha256:8c9c01d5ba6dc77de01cef77b31ca98d205ffaafefc4424532388f9997ca7ce5

Observation 8344fae6-6c4b-4b8f-a970-3f0b5373f535 · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:27.625332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.253531Z digest=sha256:c901ed3897044f43015a07990f0af96cf620e4c4d0279a050b2e8f05c4e7bb36

Observation 7385252b-4d0d-42e6-8b30-6911e05c794a · outbound

This paper cites However, for certain tasks, models of different sizes exhibit similar capabilities.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding However, for certain tasks, models of different sizes exhibit similar capabilities

Reference 64

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T21:52:27.351086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.315959Z digest=sha256:11d25606f423071fd2d54b9f3088b3ae3fa0ae3635d1ef5b420dd5baf039686a

Observation bd8fd18b-f630-4512-a500-92c471c491ba · outbound

This paper cites This task needs a model to speculate the land use type (commercial, residential, agricultural, etc.) based on a satellite image.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding This task needs a model to speculate the land use type (commercial, residential, agricultural, etc.) based on a satellite image

Reference 65

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T21:52:27.097941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.382008Z digest=sha256:1386dcba22b5b31bf5b35e07c9eca48c12cc4ecaa5a766dc858f568557711f4a

Observation 272f3bd7-3e1a-488c-b297-f5dd07e3e998 · outbound

This paper cites an unresolved cited work.

UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:52:26.762359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:52:25.446679Z digest=sha256:2a6fa4a47b6e4aa4abadea2d57971abdadecee219ee3d2f2277009f9a6051ffb

Pith citing papers

Observation eba8392f-a4ca-4c0c-b565-bdf2db73376e · inbound

IoT-Brain: Grounding LLMs for Semantic-Spatial Sensor Scheduling cites this paper.

IoT-Brain: Grounding LLMs for Semantic-Spatial Sensor Scheduling UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:05:56.875660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T17:49:16.409834Z digest=sha256:8fb27b2bc6729eefaeaabf909aa7a0802d1cc634b4ab80606f0a6597ef299c48

Observation 4a8bd704-34ac-4f67-ab3b-97c1b8c68f6c · inbound

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models cites this paper.

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:26:45.810959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T06:58:39.117228Z digest=sha256:35be8f1f2df4642a5a76a4908b28cfe0fe93419d314f0d2fed778cc3810800a0