Pith. sign in

Paper Citation Record · LEDGER

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing

As of 18 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2507.08575.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08575 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:19:59.121503Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact4
  • verified fuzzy5
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2f850f75-b021-4635-8d30-f25ed505b968 · outbound

This paper cites GPT-4 Technical Report.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:58.967536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:58.967536Z digest=sha256:cbb7da461c9f7195d2a5eaf41b60b9dc562326a1dddb1a454436a147e12800f2

Observation b5372bbb-cb5c-4103-a6e9-84bb2767805d · outbound

This paper cites PaLM-E: An Embodied Multimodal Language Model.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing PaLM-E: An Embodied Multimodal Language Model

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:58.994170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:58.994170Z digest=sha256:308ed6ba86b9381ed881d88b906ab1b324de12a49bc6e41eaedcb5eeade8a779

Observation 070fce44-a963-48c6-a482-f3dcae919d39 · outbound

This paper cites URL: https://aclanthology.org/P18-1119, doi:10.18653/v1/P18-1119.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing URL: https://aclanthology.org/P18-1119, doi:10.18653/v1/P18-1119

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:58.998910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:58.998910Z digest=sha256:4c688c14f5522981fa058f90443b8df88b17c1259b8c7bff3caa00be8e21f52a

Observation 061f5130-8738-41c6-a169-9c4c2c225d99 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.011087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:59.011087Z digest=sha256:f0ebe09ad9e6b91a00c324b64c9dffe9d8796cde502c7a6f8bd607ae8aaee0cc

Observation d329a34d-baec-401c-858c-244e2a9c0e5b · outbound

This paper cites Grounding spatial named entities for information extraction and question answering.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing Grounding spatial named entities for information extraction and question answering

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:19:59.657870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:19:59.033967Z digest=sha256:bc550b884070d33bb9597670b7eeac3a3d3534f9527f62132db70479cf402ca7

Observation 3ca0bd32-4809-48cb-a01e-0c67a2811234 · outbound

This paper cites Towards Long-Horizon Vision-Language Navigation: Platform, Benchmark and Method.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing Towards Long-Horizon Vision-Language Navigation: Platform, Benchmark and Method

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:19:59.328241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:19:59.057936Z digest=sha256:86be6d5a616f3e2f167a33d5b9b586786b5ae284a0a9a72c2156bcf415c2ba0e

Observation a76411f8-78d3-49d5-91bd-f7578568ac73 · outbound

This paper cites Global Pointer: Novel Efficient Span-based Approach for Named Entity Recognition.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing Global Pointer: Novel Efficient Span-based Approach for Named Entity Recognition

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:19:59.304088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:19:59.062417Z digest=sha256:08e0a91c62a17ffc7cf599d018f3e61ed99f71563c86d2c10808fd488bcf77b8

Observation f2326d49-382f-4aec-a1be-6f96cbcc1ee1 · outbound

This paper cites A Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPT.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing A Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPT

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.088000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:59.088000Z digest=sha256:a373aecd316a1fac5d9c64e29dc9a22030c384db9f033ec0be5476a8a8f4908e

Observation 6cb5e9d4-9e70-4b37-abd8-4c19925aee41 · outbound

This paper cites In-context learning for few-shot nested named entity recognition.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing In-context learning for few-shot nested named entity recognition

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:19:59.634015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:19:59.096943Z digest=sha256:46706cc892c884b8f5f92ce1f76fcd48ed031353d68a0e7d4ab7f7b4c46c37b5

Observation d9f3e58a-0a35-41b3-86b6-11a5327f6e6f · outbound

This paper cites A review on entity relation extraction.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing A review on entity relation extraction

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:19:59.620492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:19:59.103239Z digest=sha256:64b14d6e724f2800b6155cdb67835bdc5bfcf432cdab4e92492bd0a3a09e1004

Observation 493f346d-baa6-456f-8d54-81b62e4951eb · outbound

This paper cites Automatic Chain of Thought Prompting in Large Language Models.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing Automatic Chain of Thought Prompting in Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.108408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:59.108408Z digest=sha256:ef93b43f38adde43f49b29dbcc4505e6720e384897736005d45e03efe4f9f825

Observation c9e9eca3-7b2e-446a-82c4-a49f0f4bb7fd · outbound

This paper cites Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through Logic.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through Logic

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.112934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:59.112934Z digest=sha256:89d43af2d8310065068ffe1f846f07ef129b51765acdd9ae65cdfb4a6ac2e5ea

Observation fce66932-9b5a-478b-b1b1-a15fb387ec3c · outbound

This paper cites A frustratingly easy approach for entity and relation extraction.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing A frustratingly easy approach for entity and relation extraction

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:19:59.598471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:19:59.117455Z digest=sha256:7676ea0ca8f816618d636995c7e3793ce7780b22326a1f25a2fc5d2bae7ec6cd

Observation 577a26f8-d4b6-460a-9845-5754281a89f7 · outbound

This paper cites Geolocation on cartographic maps with multi-modal fusion.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing Geolocation on cartographic maps with multi-modal fusion

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:19:59.579481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:19:59.121503Z digest=sha256:75778043cc53ade5af5a275d78cf72925536f72d6aa204de6cdc6fb6fc5d17fb

Observation 194404aa-8c2c-4111-bd6d-595658e045ed · outbound

This paper cites MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI

Reference 2007

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.092552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:59.092552Z digest=sha256:72b78f59d598368b2c44c7e67c68698da52e5a91a93bcf2015d1f508c46aea6a

Observation 01214886-ca42-4c5d-814d-7f7e8e087fb1 · outbound

This paper cites Deep Embedding for Spatial Role Labeling.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing Deep Embedding for Spatial Role Labeling

Reference 2009

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T18:19:59.362605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:19:59.049593Z digest=sha256:579904c843965c31056cf49cb702501d79af826e509f91e4ce0dea950b0360f3

Observation 086ab75d-abdb-45d0-8780-a0562ec06b05 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 2010

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:58.977609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:58.977609Z digest=sha256:6ea873401983a31e9f8a5cc486513c89a4ef8776aa29b6a73eb3adbe416e56f1

Observation 290b1aad-18df-4aa2-ab3f-cebbf98e35be · outbound

This paper cites 24 Yoonsik Kim, Moonbin Yim, and Ka Yeon Song.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing 24 Yoonsik Kim, Moonbin Yim, and Ka Yeon Song

Reference 2013

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T18:19:59.485488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:19:59.027410Z digest=sha256:6881ee994a66bc7462a3e1c7180d4f25ef94f62519f73ff156cd704e7a521467

Observation be06cf4a-4fb7-4f1a-9c28-5d94b593515f · outbound

This paper cites an unresolved cited work.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing Unresolved cited work

Reference 2015

Resolution
verified exact
doi, observed 2026-08-06T18:19:59.168946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:19:59.070619Z digest=sha256:6970b11279a316f55bd3d12d26edcd1386a61aa8109c1b1bfbbcb7ab1b225162

Observation 48d2d07b-9d23-4fd4-a809-37840541772b · outbound

This paper cites GPT-RE: In-context Learning for Relation Extraction using Large Language Models.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing GPT-RE: In-context Learning for Relation Extraction using Large Language Models

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.074638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:59.074638Z digest=sha256:075ac00baa7a80c9ddb9634aa49eaa99ad6ffd63f1ab1989dffe7c85c4a05e89

Observation ec5fe78f-d0a6-4ab1-90e7-b7527bf3e4e0 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:58.986291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:58.986291Z digest=sha256:a5605acfa7ec8a74a02e9760437d98b0db906307ee39af67ee73c0a5b3afc72a

Observation 2b709d42-0538-4c77-a415-ebb661b4d579 · outbound

This paper cites 59 Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing 59 Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.080445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:59.080445Z digest=sha256:b23688487da4f0bb518eb53d40d5e4324c5f026b9cceedf98cc262b4412dc0cc

Observation fd567148-4df0-4a48-955a-92c6a09183ea · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.066219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:59.066219Z digest=sha256:7f1d07e2251bf9e6077ccd900c02b86bad927e6f39aa5eca051a05ca20611f2a

Observation 445ad10e-b5e3-45ea-bff3-caae10fb8c52 · outbound

This paper cites M5 -- A Diverse Benchmark to Assess the Performance of Large Multimodal Models Across Multilingual and Multicultural Vision-Language Tasks.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing M5 -- A Diverse Benchmark to Assess the Performance of Large Multimodal Models Across Multilingual and Multicultural Vision-Language Tasks

Reference 2021

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:19:59.346091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:19:59.053584Z digest=sha256:b59c5e3bbc6803a0fe9c1a75a8c92b9fe2e5171d65d78f69bd4eba2a6dff4ccc

Observation dbb14d36-a4a6-4345-bfcd-dd00aca3842b · outbound

This paper cites A transformer-based framework for poi- level social post geolocation.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing A transformer-based framework for poi- level social post geolocation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.038022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:59.038022Z digest=sha256:8eb90057d63030088785a16e52336d7c688d76df73284bc75e0662c68fb7f811

Observation 53351287-095b-4972-b180-3e85ad35486f · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:58.990203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:58.990203Z digest=sha256:20aa2a1f25d9362831ed87975f7bf48462d2cb890d6c7b95b9fd89f1e24cb4b7

Observation 1ceb43d4-b75e-4ce5-a815-76d33231b77c · outbound

This paper cites DeepSeek-V3 Technical Report.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing DeepSeek-V3 Technical Report

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.043443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:59.043443Z digest=sha256:609441b2cf00c29315657711647c5c7dbf25f2a2e5e81ce0e6069b8bf6bee006

Observation 99d47f5e-5a25-4663-a847-3451780e21fe · outbound

This paper cites an unresolved cited work.

Large Multi-modal Model Cartographic Map Comprehension for Textual Locality Georeferencing Unresolved cited work

Reference 2025

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:19:59.674004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:19:59.023282Z digest=sha256:9c177b5997c71ff184d612024b397c1fb35f6527a37c2031f30813d8eb1fbbf3

Pith citing papers

No inbound Pith citation observations are available.