Pith. sign in

Paper Citation Record · LEDGER

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

As of 17 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 14 inbound Pith citation observations for arXiv:2501.01282.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.01282 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:36:03.857380Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:14:16.972880Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T20:16:12.001782Z

Reference resolution

34 of 34 outbound references displayed

  • verified exact1
  • verified fuzzy3
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b703a96a-d32e-49ea-8089-693a395cc08c · outbound

This paper cites PersianLLaMA: Towards Building First Persian Large Language Model.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries PersianLLaMA: Towards Building First Persian Large Language Model

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.693690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.693690Z digest=sha256:bc1915d51cd19a43d4dd0af70ccaf9bd6f40a4de09315f1856eeb84887e2d45d

Observation 36eab720-581d-4f61-951b-e5d6ca9d659d · outbound

This paper cites Pixtral 12B.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Pixtral 12B

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.709739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.709739Z digest=sha256:47ef5c21fa131f3501af78eaa1277daf3c013af2e7e288099afcdc01159f05af

Observation ce6f9271-f9d7-4c54-a8b5-bdf1c7f639d9 · outbound

This paper cites From Local Concepts to Universals: Evaluating the Multicultural Understanding of Vision-Language Models.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries From Local Concepts to Universals: Evaluating the Multicultural Understanding of Vision-Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.719534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.719534Z digest=sha256:89250f0e5ddd15a0ad913b137ee8706bc95a841c973384ec316d2d5635c006a6

Observation 392d0ad3-7e2c-4496-83b9-877c86c47305 · outbound

This paper cites Enhancing Content Moderation with Culturally-Aware Models.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Enhancing Content Moderation with Culturally-Aware Models

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:36:04.389465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T22:36:03.724985Z digest=sha256:b5ca6ec53eeae38b603f344e3e3a722d6d0356ad14e4b094e0ef576daf4a020b

Observation bde0b0d4-b8f4-447c-8b5b-2ef91d4ace20 · outbound

This paper cites InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.729584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.729584Z digest=sha256:64bf61b341aa32bc763f85ea7f1260f2e014ed0c077a73d6095c65b0630e3bf4

Observation de06cbaf-44cd-4d74-883f-421f3ea58a90 · outbound

This paper cites CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.733945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.733945Z digest=sha256:4062ef1175c37e7968c951805ad8b083be62cc5eb09643ae6560a9d6809d7f8a

Observation d191fd92-d47c-420a-a4a5-9ebb1c42763c · outbound

This paper cites Deconstructing The Ethics of Large Language Models from Long-standing Issues to New-emerging Dilemmas: A Survey.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Deconstructing The Ethics of Large Language Models from Long-standing Issues to New-emerging Dilemmas: A Survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.738526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.738526Z digest=sha256:e1468e16c548c9d3e54c0e588b9f512bb9d3006436af616bddec65825228bea2

Observation d369c2e0-6756-4eb1-8657-c2133cfdc502 · outbound

This paper cites Massively Multi-Cultural Knowledge Acquisition & LM Benchmarking.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Massively Multi-Cultural Knowledge Acquisition & LM Benchmarking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.743295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.743295Z digest=sha256:14b71c7c325e5907d85739e99f8892d460cc64cd8164119223f2d54fcd5d58f6

Observation 500688a7-7ccd-4040-9855-d84822e0db39 · outbound

This paper cites NormSAGE: Multi-Lingual Multi-Cultural Norm Discovery from Conversations On-the-Fly.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries NormSAGE: Multi-Lingual Multi-Cultural Norm Discovery from Conversations On-the-Fly

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.748211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.748211Z digest=sha256:81202059126916820fa8f0d947fbb7599d3e84550f364f0aadcded2b31a8e0c3

Observation c6236d50-2e70-44d7-9fbc-f7e4e07ff15d · outbound

This paper cites GPT-4o System Card.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries GPT-4o System Card

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.753017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.753017Z digest=sha256:04db62852d1e402ab05ff7d0fecb42fb97524dac28ae1cb3a686d8376559b74f

Observation bd6dc3eb-49ad-400d-bb0b-94673e49f68b · outbound

This paper cites The Ghost in the Machine has an American accent: value conflict in GPT-3.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries The Ghost in the Machine has an American accent: value conflict in GPT-3

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.757481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.757481Z digest=sha256:3e0da1bd4333419c531205614f8f7b096ea4ba3390ca99716dfc17558d6d291c

Observation ced4e771-125b-4515-b76e-0d3ba48b78d7 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries LLaVA-OneVision: Easy Visual Task Transfer

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.762120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.762120Z digest=sha256:01b523a6be35050e0d16385f48f01f958be358e9dd0af8a96ca66fc4907cc4e1

Observation 28d21c31-08ca-4e28-96c2-d6c8549d8dd3 · outbound

This paper cites Are Multilingual LLMs Culturally-Diverse Reasoners? An Investigation into Multicultural Proverbs and Sayings.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Are Multilingual LLMs Culturally-Diverse Reasoners? An Investigation into Multicultural Proverbs and Sayings

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.766605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.766605Z digest=sha256:517a6a13892d57dbb56b452977cad4835404dc297f71694de067830f7eef5d2e

Observation d7edca81-f066-49dd-aff0-161460eca7de · outbound

This paper cites The 2012 stein rokkan lecture: Three decades of popu list radical right parties in western europe: so what? In The Populist Radical Right, pages 545–558.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries The 2012 stein rokkan lecture: Three decades of popu list radical right parties in western europe: so what? In The Populist Radical Right, pages 545–558

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:36:04.511847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T22:36:03.780912Z digest=sha256:319cd2752b7cc9c622ab20698b6b6dda2da61d85b42e2989f2d4fff9c9f291c2

Observation 23b6f6d1-9bf4-4278-a6f6-00c57d8ccf17 · outbound

This paper cites Having Beer after Prayer? Measuring Cultural Bias in Large Language Models.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Having Beer after Prayer? Measuring Cultural Bias in Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.790331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.790331Z digest=sha256:c4950ff0f2f938c7edfd8f81f33d6f22734e58921d64960ab6d0e9420b1fd852

Observation 0c3c998a-2872-4efc-903c-f08a9a740455 · outbound

This paper cites SeaLLMs -- Large Language Models for Southeast Asia.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries SeaLLMs -- Large Language Models for Southeast Asia

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.794935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.794935Z digest=sha256:dd83c2bd91ba15f626ea8c8374657ec90fe3a257766331cfcf90a792ee964a47

Observation 004f5fa0-8949-4286-8d76-d5817565537d · outbound

This paper cites Typhoon: Thai Large Language Models.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Typhoon: Thai Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.799848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.799848Z digest=sha256:92f66aae83e8ca250fda42bf6bfce11274bd1c078872cd1e5f5737477a3af6a1

Observation 60af69ee-f49b-41e6-a5a8-e70b951bdf37 · outbound

This paper cites NormAd: A Framework for Measuring the Cultural Adaptability of Large Language Models.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries NormAd: A Framework for Measuring the Cultural Adaptability of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.804437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.804437Z digest=sha256:112ad4f05672b0482bb670415c47f4dc87a21274a6f9f449fe1ee645f7e5f4f6

Observation a91608e1-9925-42b6-a46b-87b2041d3e67 · outbound

This paper cites CVQA: Culturally-diverse Multilingual Visual Question Answering Benchmark.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries CVQA: Culturally-diverse Multilingual Visual Question Answering Benchmark

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.809020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.809020Z digest=sha256:641dc40cad9f07d635a7c3f6c56942843792b1709f547c9e0623f39b9c8d1430

Observation c84d5bb1-a7f7-4bee-8637-4272a4afa976 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Gemini: A Family of Highly Capable Multimodal Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.823476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.823476Z digest=sha256:e4210df73cccf9aeb87ce00b10f6450758472402b04770624148c95fdb2b01b4

Observation 91faec6d-c017-4908-b5a4-826b0d003bcc · outbound

This paper cites A Comprehensive Survey of Hallucination Mitigation Techniques in Large Language Models.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries A Comprehensive Survey of Hallucination Mitigation Techniques in Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.828161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.828161Z digest=sha256:b3ad8f9e4a23177c5f4769aeac827778bd06dd0b9e2921092129b7728c82931d

Observation 189d4403-aaf3-447c-9264-f3e4c5d66c22 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.832979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.832979Z digest=sha256:4a21d5f0aa8af19687f91d4896a6289ce84c6ef5432e797b743974fa0167210c

Observation 4c6f1a82-d5dc-4afd-8d50-1c907be07ddc · outbound

This paper cites Not All Countries Celebrate Thanksgiving: On the Cultural Dominance in Large Language Models.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Not All Countries Celebrate Thanksgiving: On the Cultural Dominance in Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.838421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.838421Z digest=sha256:60c1e5719211534c597c5bd4cdf0f016cdae8aa5abfeb753cceb1bc31196b7ca

Observation 0db2fc3e-8a23-45be-90ea-9913216c4481 · outbound

This paper cites WorldCuisines: A Massive-Scale Benchmark for Multilingual and Multicultural Visual Question Answering on Global Cuisines.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries WorldCuisines: A Massive-Scale Benchmark for Multilingual and Multicultural Visual Question Answering on Global Cuisines

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.843224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.843224Z digest=sha256:f7f147a84728116a3c17b8321055112793f1c56a7bd5b2cfd42064ea312db7ef

Observation e9e8297b-fb45-4ee4-8ef3-3a56a86f5e7e · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.847783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.847783Z digest=sha256:0d55e9c12986bda33de3f03680a85e7b5f2b7e094b7521300df91003a5f4a4ed

Observation 7c5b56fa-ee0f-4cc7-b964-f6bd97b56429 · outbound

This paper cites 14 B.2 Detailed Main Results.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries 14 B.2 Detailed Main Results

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:36:04.496808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T22:36:03.857380Z digest=sha256:8f845b4cdfd627dc85efb8d76c09920d0cef095f2e15ce2c19c94f72e04d7c55

Observation 62c4e83c-4910-4e1d-9a20-d9afa2e9fa68 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Hallucination of Multimodal Large Language Models: A Survey

Reference 1996

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.714782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.714782Z digest=sha256:34bee54dd0402879dfca9fcbd30795ba0b464b4c867361c3ffb8a8dec6cf0399

Observation acb73a92-2334-4750-be28-3d3978e2a2ea · outbound

This paper cites WorldValuesBench: A Large-Scale Benchmark Dataset for Multi-Cultural Value Awareness of Language Models.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries WorldValuesBench: A Large-Scale Benchmark Dataset for Multi-Cultural Value Awareness of Language Models

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.852041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.852041Z digest=sha256:5ece7de56cc88f4d26adb3f9bef825b336b7d10355fccf43fea4d733e668280d

Observation 6fb93bf7-e972-492f-ba2f-57d05f129b30 · outbound

This paper cites BLEnD: A Benchmark for LLMs on Everyday Knowledge in Diverse Cultures and Languages.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries BLEnD: A Benchmark for LLMs on Everyday Knowledge in Diverse Cultures and Languages

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.785654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.785654Z digest=sha256:2b35ce153571eed869e2657c28c3976d3e9ea401711a7b49d449781a0f948cb8

Observation 9ed418d1-e707-44fb-a72d-58467dea94dd · outbound

This paper cites Llama 3.2.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Llama 3.2

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:36:04.526513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T22:36:03.776122Z digest=sha256:4e3adcf153b537d74eb8fd042517aa05c48164d6a8282283422084b469a8bd69

Observation f17c16cb-fc39-45d4-9665-d1ffa17ed9ee · outbound

This paper cites Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.771478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.771478Z digest=sha256:d23d789b313c22f40097dce14ebe35aa0487a2f97b512f679f3ec7ebff351fe2

Observation b0dfbc3c-30b0-46af-93e2-3f087f194d85 · outbound

This paper cites CultureBank: An Online Community-Driven Knowledge Base Towards Culturally Aware Language Technologies.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries CultureBank: An Online Community-Driven Knowledge Base Towards Culturally Aware Language Technologies

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.818665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.818665Z digest=sha256:5d031e7d49fc35c194eaa1618a70d7cca343dc6d38fc6ab31a9c8dfc76ea3a50

Observation b1d9945c-355c-4f40-a0f8-dbe3d1360e60 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.699532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.699532Z digest=sha256:9b922e7bc14c74353cdbc9cd51b9c11b60a140a56f480c62aae7aef28540cfe0

Observation 1c4fc194-ecd0-431a-9fba-5a23b6b5ea51 · outbound

This paper cites Towards Measuring and Modeling "Culture" in LLMs: A Survey.

CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries Towards Measuring and Modeling "Culture" in LLMs: A Survey

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-10T22:36:03.704833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:36:03.704833Z digest=sha256:9e03ae9212ecf7a8e17b30e9956fe9aab28acee0071f887c90d5ba4c0f6cf973

Pith citing papers

Observation 56f58a89-accb-4087-82a8-e4c9fb06670e · inbound

Large Language Model Agent: A Survey on Methodology, Applications and Challenges cites this paper.

Large Language Model Agent: A Survey on Methodology, Applications and Challenges CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 248

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:52:10.327270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-22T21:51:34.309870Z digest=sha256:0ac71b10dbea0aa086d5d12f39b17eef844dbb2f83a25d909f4816a6b367ca36

Observation e6761127-d27f-4325-8902-ba931b0491e7 · inbound

Uncovering Cultural Representation Disparities in Vision-Language Models cites this paper.

Uncovering Cultural Representation Disparities in Vision-Language Models CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T20:14:16.972880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:14:16.972880Z digest=sha256:738247e80691fd5183017507ea3b09d8146cb360a4a7e6c9b4f6e50aa16a7a62

Observation f7677043-799b-4883-b148-6e8915e789c8 · inbound

Evaluation of Cultural Competence of Vision-Language Models cites this paper.

Evaluation of Cultural Competence of Vision-Language Models CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:12.692200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:03:12.692200Z digest=sha256:056814b1edfadc287f7d9bdff44e34759d0ecac9c4ebcd5a273e0bddddf09af1

Observation 2bb404bc-4269-4214-adb8-0f5493d8c35a · inbound

Domain Specific Benchmarks for Evaluating Multimodal Large Language Models cites this paper.

Domain Specific Benchmarks for Evaluating Multimodal Large Language Models CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-07T00:39:42.056277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:39:42.056277Z digest=sha256:b8bc4856c44ecd08e61a981ae547fc9eeaf4bb6c1b628a5fbea953fd5174bbf0

Observation 5dcde609-893b-456c-be9d-dd3b74187c34 · inbound

SARA: Selective and Adaptive Retrieval-augmented Generation with Context Compression cites this paper.

SARA: Selective and Adaptive Retrieval-augmented Generation with Context Compression CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T19:26:58.410428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:26:58.410428Z digest=sha256:98e369e644206ae59fc16361e3515be3f144220090da75770ea466cacf8185aa

Observation 6274f23e-47bc-47da-96b8-c796c8ca706b · inbound

Appear2Meaning: A Cross-Cultural Benchmark for Structured Cultural Metadata Inference from Images cites this paper.

Appear2Meaning: A Cross-Cultural Benchmark for Structured Cultural Metadata Inference from Images CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:05:55.519801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T17:49:25.227568Z digest=sha256:1c5f368adeb4345eda1d209f4ac28806486668dd489e45fc96192db2ada32c38

Observation bcf456a2-4aa1-49a2-a39a-e0dff11e04d2 · inbound

BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources cites this paper.

BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:01:02.085493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T04:22:34.046014Z digest=sha256:5cd929650014ec997c0c285f61d650b029c953e4ec2a69164fb5fe5dcd459e59

Observation bb620059-7136-4826-96eb-ad34ac4750f1 · inbound

When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation cites this paper.

When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T22:02:49.023392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-19T21:59:44.030267Z digest=sha256:5cb2090763f0f1bad3ce78d8bdbf16107ba7c93241309246de293e24147faf97

Observation 2cf76168-110d-4e89-85ec-9599d1807bef · inbound

When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation cites this paper.

When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T19:45:01.074023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-30T19:36:43.654439Z digest=sha256:f13fef4f3c21959dd1a01650f3e51486bacbd3bfdadc668d570edbacc9d2114c

Observation 528c4356-9aa1-47be-8658-c2530d7ab751 · inbound

When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation cites this paper.

When Cultures Move: Measuring and Improving Multicultural Text-to-Video Generation CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T13:52:47.550786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:52:47.550786Z digest=sha256:448b1a2f83adc3fb0155800f5eccd7400358d3d24caf84e76f01256bf781f7cd

Observation 6ab89fc6-1078-4da3-9d4e-e256b772702b · inbound

Computer-Aided Tagging on Wikimedia Commons: Designing for Human-AI Collaboration in Open Knowledge Work cites this paper.

Computer-Aided Tagging on Wikimedia Commons: Designing for Human-AI Collaboration in Open Knowledge Work CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.003355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-28T21:23:51.431142Z digest=sha256:13997d6fe1072635b86d9cfc85218056f04f0ab942413eaebaf8665fe0f3361e

Observation d1674364-8765-43a4-b7d2-96e63035fddd · inbound

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models cites this paper.

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-11T15:14:25.526254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T15:14:25.526254Z digest=sha256:13da06ce26735aa35d98d98bbfd385e0cc37a1d7bcc373903adc0e0e28eaf7f4

Observation 6c6ff0e7-d50f-4986-a70c-519d5ecad198 · inbound

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models cites this paper.

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-13T06:57:34.446788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T06:57:34.446788Z digest=sha256:5b562fa0bd08614dc948a68b6701e61d69d37fd78e1fd157abce2655ad8f25c8

Observation dfd18ba2-00b5-49c3-a827-a3932a04772f · inbound

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models cites this paper.

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T04:32:12.656115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:32:12.656115Z digest=sha256:389e3d99a13b171858ab80d7f06c56a19848cdd877bc23754fefa863eb7044b9