Pith. sign in

REVIEW 1 cited by

AI Sees Your Location, But With A Bias Toward The Wealthy World

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2502.11163 v3 pith:ST42HZWG submitted 2025-02-16 cs.CV cs.CL

classification cs.CVcs.CL
keywords imagesvlmsbiasesgeographicperformancedevelopedinformationmodels
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Visual-Language Models (VLMs) have shown remarkable performance across various tasks, particularly in recognizing geographic information from images. However, VLMs still show regional biases in this task. To systematically evaluate these issues, we introduce a benchmark consisting of 1,200 images paired with detailed geographic metadata. Evaluating four VLMs, we find that while these models demonstrate the ability to recognize geographic information from images, achieving up to 53.8% accuracy in city prediction, they exhibit significant biases. Specifically, performance is substantially higher for economically developed and densely populated regions compared to less developed (-12.5%) and sparsely populated (-17.0%) areas. Moreover, regional biases of frequently over-predicting certain locations remain. For instance, they consistently predict Sydney for images taken in Australia, shown by the low entropy scores for these countries. The strong performance of VLMs also raises privacy concerns, particularly for users who share images online without the intent of being identified. Our code and dataset are publicly available at https://github.com/uscnlp-lime/FairLocator.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. GeoChain: Multimodal Chain-of-Thought for Geographic Reasoning

    cs.AI 2025-06 conditional novelty 6.0 of 10

    A new 21-step geographic reasoning benchmark built from 1.46 million street-view images shows current multimodal LLMs handle simple visual questions well but rarely localize precisely.

Pith tools