Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:46:12.494249Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 100 of 106 outbound references and 2 inbound Pith citation observations for arXiv:2507.04664.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:46:12.494249Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:19:20.192163Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T15:19:20.262812Z
100 of 106 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d8210e97-721d-4377-98c5-7fbcf411d4a1 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 284f28ca-349f-44b6-a49a-aee047b86458 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Efficient interactive annotation of segmentation datasets with polygon-rnn++
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ed2ba49-0e13-4285-bc0b-562f5b2176d5 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Qwen Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63f16326-a8d3-47a3-9f67-0cafaf669179 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55a2e43a-de79-4c6a-94f2-51ed9ffe9dff · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Multi-task learning for segmentation of building footprints with deep neural networks
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efa7c1c8-6e2f-4fcd-982f-bd06561d09ba · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs J., 2019
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79d0203d-3bb3-4a7b-beee-c4685cd45c41 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs InternLM2 Technical Report
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d361c8c1-8a8f-44a8-a1ab-8d1ea92df04e · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Annotating object instances with a polygon-rnn
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5c7e7e8-8b92-4f60-8cc8-4beaa3a6d456 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs ASF-Net: Adaptive Screening Feature Network for Building Footprint Extraction From Remote-Sensing Images
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9ed3f77-ac4b-48de-844d-899ff0a7eb35 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs RSPrompter: Learning to prompt for remote sensing instance segmentation based on visual foundation model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 676c25f8-9990-409e-9867-bed03307dcbc · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs L., Liu, X., 2020
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f688f2a-513d-4336-8c6b-2217dd18266d · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs CGSANet: A Contour-Guided and Local Structure-Aware Encoder--Decoder Network for Accurate Building Extraction From Very High-Resolution Remote Sensing Imagery
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c0fb206-cf12-4d86-85e9-95a803d45295 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 098777dd-1745-4fc0-8d9d-fa06d2e57f0b · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37d15ac5-d595-4ae6-85ce-e59ac22db8af · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs et al., 2024d
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a946523-bd2f-4276-853f-43490e74626e · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs G., Kirillov, A., Girdhar, R., 2022
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee994916-5335-4e80-af27-bb9bd9bc5ecf · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs E., Stoica, I., Xing, E
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03a7438f-6450-4b4f-b99e-11d5ef313a01 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72388a81-c250-4eec-a565-219e61eab721 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9a3c9a8-4cf6-4041-a84b-2654ce84d911 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs The Llama 3 Herd of Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 402e5d75-ed0f-4680-a137-99357aa52267 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs GeoLLaVA: Efficient Fine-Tuned Vision-Language Models for Temporal Change Detection in Remote Sensing
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2644427a-76f4-458a-ae86-b42183c9ab8c · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Instances as queries
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3417c8d8-f6f3-49f5-80c0-f6220469bb90 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs GPT-3: Its nature, scope, limits, and consequences
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22fb0106-5f5a-4949-b4c3-7d169df1fd6e · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Polygonal building extraction by frame field learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fec73877-e797-40a4-a79b-5ad3e5d28eec · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Remote Sensing ChatGPT: Solving Remote Sensing Tasks with ChatGPT and Visual Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f238ad8e-3b6b-43ad-8503-5e493a6da9c9 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs HigherNet-DST: Higher resolution network with dynamic scale training for rooftop delineation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8a427c6-96d6-4c74-a134-891fa92077fc · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Mask r-cnn
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d390bb96-5275-4feb-ae74-c349d1d523d6 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Rsgpt: A remote sensing vision language model and benchmark
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a2c0461-438f-4b96-87cb-4627b4b0c473 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Sequentially delineation of rooftops with holes from VHR aerial images using a convolutional recurrent neural network
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4336df52-0d5d-43e9-9c6c-c7e0fb34713f · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs OEC-RNN: Object-oriented delineation of rooftops with edges and corners using the recurrent neural network from the aerial images
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09dde30f-882b-419e-83a3-f1e85ca77997 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs TEOChat: A Large Vision-Language Assistant for Temporal Earth Observation Data
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d30da99-0abc-453d-b115-07c3af35c33e · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Fully convolutional networks for multisource building extraction from an open aerial and satellite imagery data set
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcd086c9-29ac-424a-8f5d-3342a2617651 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs A scale robust convolutional neural network for automatic building extraction from aerial and satellite imagery
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 889edfd5-384c-43d1-83c6-c19aae744b8c · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Segment Anything
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54e8859c-9ea6-43ac-868d-0d2187b98f52 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs C., Lo, W.-Y
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62b9df7f-f2c1-4fa1-99c1-dd732c0569ce · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs E., 2017
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a82c35c5-a1ed-4198-92a3-a2b2040f4336 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs S., Naseer, M., Das, A., Khan, S., Khan, F
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf74aaef-e9c3-4173-8e2e-12842ec9cdaa · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Lisa: Reasoning segmentation via large language model
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 83f5ad00-4f16-4c13-ae3c-d5093f990ca2 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs VRSBench: A Versatile Vision-Language Benchmark Dataset for Remote Sensing Image Understanding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7f65599-2566-4ca8-946f-9be25206b819 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Topological Map Extraction from Overhead Images
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3ce7b139-c1b1-4957-8785-0d8f624a62bc · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs D., Lucchi, A., 2019
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f2c10974-566d-4cc7-91ad-596ab7bb7a93 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Polytransform: Deep polygon transformer for instance segmentation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d37fc12e-ba8d-4aed-8492-ab0056e2f0a9 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs L., 2014
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 827c00df-b728-4e03-9d0f-d55d15100aa4 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Draw-and-understand: Leveraging visual prompts to enable mllms to comprehend what you want
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d0bbce28-e468-439e-99ca-15afbfcb788e · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Change-agent: Towards interactive comprehensive remote sensing change interpretation and analysis
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d3ad047d-f059-4b44-aadd-03e46058c842 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs J., 2023a
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c9f2d888-2725-4ade-8a56-de7b935ab042 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs J., 2023b
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 78abfdac-54bd-486c-a69d-ede748ad8ac3 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Path aggregation network for instance segmentation
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e24ef6bb-7b5e-4eb5-b2b1-cf06837ae3e4 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Swin transformer: Hierarchical vision transformer using shifted windows
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ce1e00f3-4929-46f3-b4f9-6d7406b15fc6 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Building Outline Delineation From VHR Remote Sensing Images Using the Convolutional Recurrent Neural Network Embedded With Line Segment Information
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b13b024e-9a19-4236-ac5f-2454e73df9d4 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b14cb8b-87f3-46a5-958f-6c3c4d137129 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Cross-spatiotemporal land-cover classification from VHR remote sensing images with deep learning based domain adaptation
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 686281e2-c5b5-4920-8bb8-32a113dcae3c · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs SAM-RSIS: Progressively adapting SAM with box prompting to remote sensing image instance segmentation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 29af8f0c-1af2-430f-bd2f-6a6bca28fb56 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs P., 2019 (accessed November 10, 2019)
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8df539d9-a371-4125-b177-b49eacccc809 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 128196e5-76c8-4d90-b068-f37701d3c630 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs DINOv2: Learning Robust Visual Features without Supervision
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f4aace0-37a0-4b07-b11b-ba9fbb559ba7 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f887536e-9793-4d10-97a4-26108a357f69 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Deep snake for real-time instance segmentation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4c97ddd7-5460-4792-b3ba-aa1775f656b2 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c54176be-9c7f-44e2-8812-74181bd3f3f5 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs D., Ermon, S., Finn, C., 2024
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0cd113dd-ee83-4552-b263-db1eb4fa46d9 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Am-radio: Agglomerative vision foundation model reduce all domains into one
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation aba2b761-0c8c-4030-8fed-7672c527f1fd · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs M., Xing, E., Yang, M.-H., Khan, F
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9aad2ccb-9034-46e9-a8aa-bd39225ce36f · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Pixellm: Pixel reasoning with large multimodal model
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 540c0d6d-2bb8-433d-807a-fc007055224f · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Geollm-engine: A realistic environment for building geospatial copilots
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation cc1b2af8-ef4a-471e-9c88-355c64381b81 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Internlm: A multilingual language model with progressively enhanced capabilities
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation df3fd4f0-9393-4133-a26b-81021c816039 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Fcos: Fully convolutional one-stage object detection
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c75b4a2d-c816-4c9d-afe0-9087289aa521 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fb2122f-b754-48a1-937a-73536643f55e · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs From image transfer to object transfer: Cross-domain instance segmentation based on center point feature alignment
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 33524a30-b01d-41a9-a075-6ef26917ce84 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4eb0793-f138-4d14-bd7d-67ba16e7bf46 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs et al., 2024b
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 54bb2a9e-a2c6-4242-9dfc-8fcac75dd936 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Solo: Segmenting objects by locations
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 120e7511-4cf5-4df1-9071-b019c8b8f415 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Skyscript: A large and semantically diverse vision-language dataset for remote sensing
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 20acfc7f-f109-4a06-ac92-5b3d82a53cde · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Graph convolutional networks for the automated production of building vector maps from aerial images
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bc4e1410-b3d9-43c9-9dc1-bae1c1a1f4d7 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Toward automatic building footprint delineation from aerial images using CNN and regularization
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b74a44fc-cae6-4e8c-811c-15f90523039e · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs A Concentric Loop Convolutional Neural Network for Manual Delineation-Level Building Boundary Segmentation From Remote-Sensing Images
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 681bfab7-2255-4e1e-8b9c-52960724f664 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs BuildMapper: A fully learnable framework for vectorized building contour extraction
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 33d5c8b8-0ad6-43f7-b51f-561cfea87f99 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs From lines to Polygons: Polygonal building contour extraction from High-Resolution remote sensing imagery
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d6e94cea-aa35-4323-b476-edf1d9b58ca1 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs A., 2022
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9e79cef5-9028-487b-a278-f3bfd73a6cdd · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Vectorizing historical maps with topological consistency: A hybrid approach using transformers and contour-based instance segmentation
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 05abb26e-ca00-4925-a7c1-12782aa2dbdc · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Video instance segmentation is all you need for linking geographic entities from historical maps
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d787037c-3fbf-4ccc-b604-62d610df6fde · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Polarmask: Single shot instance segmentation with polar representation
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 939ef4fa-5a08-4f74-aba5-74d9dfcd08bf · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs HiSup: Accurate polygonal mapping of buildings in satellite imagery with hierarchical supervision
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 61e0fd6c-7086-41b8-a439-e08f2b35bd76 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Qwen3 Technical Report
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45a3c82f-e62d-44f6-be72-1e0b36778ccb · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Qwen2 Technical Report
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02d2e529-04d6-463c-b965-a3a9370fd2c3 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs A review of recurrent neural networks: LSTM cells and network architectures
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b59a960f-b2db-46c5-b692-a14e5a8fb91e · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90967185-4b8f-4448-80df-44aa4b22d059 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Learning building extraction in aerial scenes with convolutional networks
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 184b47f0-cbea-4043-99d8-a666a269c6b0 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Osprey: Pixel understanding with visual instruction tuning
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 22a3060a-2a1b-45da-a5e6-2f7ab87ee0dc · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs SkyEyeGPT: Unifying Remote Sensing Vision-Language Tasks via Instruction Tuning with Large Language Model
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73189925-8303-4776-8873-e0f744cf96dd · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs HiT: Building Mapping with Hierarchical Transformers
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f294c6ba-585a-46ea-8119-03ae8b83043c · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Gpt4roi: Instruction tuning large language model on region-of-interest
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d7f53afa-67ab-4129-8fb1-ac41a3f9eff3 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs OMG-LLaVA: Bridging Image-level, Object-level, Pixel-level Reasoning and Understanding
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d347ad1-9a51-40ec-851f-1a87105d9ce2 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Pixel-SAIL: Single Transformer For Pixel-Grounded Understanding
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1eeaea0-1738-4a20-860f-427110480de5 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs E2ec: An end-to-end contour-based method for high-quality high-speed instance segmentation
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 75124c75-c625-4f0c-9a5d-d9db4f6e4197 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs P2PFormer: A Primitive-to-polygon Method for Regular Building Contour Extraction from Remote Sensing Images
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e38dff98-5106-4db9-b4a2-82426754e431 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Popeye: A Unified Visual-Language Model for Multi-Source Ship Detection from Remote Sensing Imagery
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f2d2a953-4bb8-4b02-a514-c83c161a84e7 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Earthgpt: A universal multi-modal large language model for multi-sensor image comprehension in remote sensing domain
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a09fa0d4-3ad7-4110-88ca-09c776f96563 · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs RS5M and GeoRSCLIP: A large scale vision-language dataset and a large vision-language model for remote sensing
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3a167d90-4e72-487e-a4cc-b679bf24355a · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Building extraction from satellite images using mask r-cnn with building boundary regularization
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c48c7d0f-3e53-4584-8f84-413a13286d3f · outbound
VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs Building instance segmentation and boundary regularization from high-resolution remote sensing images
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 03578315-1c94-4a01-a3a4-eb074b2e5e40 · inbound
HoliTracer: Holistic Vectorization of Geographic Objects from Large-Size Remote Sensing Imagery VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 87847722-2f3c-4d59-87fc-5c8b6bb0e045 · inbound
Actor as Its Own Critic: Unifying Region Understanding and Localization via CycleGRPO VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.