Pith. sign in

Paper Citation Record · LEDGER

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation

As of 8 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2508.16762.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.16762 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T17:15:49.702336Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy26
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b2c894af-23a1-4d9e-874a-f1d50757802f · outbound

This paper cites To- wards measuring and modeling “culture” in LLMs: A sur- vey.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation To- wards measuring and modeling “culture” in LLMs: A sur- vey

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.507144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.487041Z digest=sha256:24633ad2811bd5f01c1a1686debbc36b8ea0419c3ec1aa81a6c52b714f24dbe1

Observation e6385693-90d8-416f-9ec7-ffe9cbb1395b · outbound

This paper cites Investigating Cultural Alignment of Large Language Models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Investigating Cultural Alignment of Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.496287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.496287Z digest=sha256:402aebe64f4f1115fd7cde5472274f048f1b7f47719a2fd8642d85e55e91efd2

Observation e26cacf3-710f-4600-b4af-5286c44cab1c · outbound

This paper cites Probing Pre-Trained Language Models for Cross-Cultural Differences in Values.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Probing Pre-Trained Language Models for Cross-Cultural Differences in Values

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.502065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.502065Z digest=sha256:bd9f8c6e8effeed52e7efe9e68f91eb0398bde5d496ecb49593d38a28b7fb6db

Observation eeff2b56-5420-42a3-a3ee-a0d9ff2c354e · outbound

This paper cites Probing pre-trained language models for cross-cultural dif- ferences in values.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Probing pre-trained language models for cross-cultural dif- ferences in values

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.488957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.511930Z digest=sha256:4190bc7b4d2c8097aa2252935e8650466913a467ff06910b1fec8600ffa25d22

Observation 66a2cf49-afc4-4e96-b9ad-aeec69601ceb · outbound

This paper cites Qwen2.5-vl technical report, 2025.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Qwen2.5-vl technical report, 2025

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.470713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.518870Z digest=sha256:793ed5d1f0f02339871100a67e6876cd700dbb042ff37d05a198fc3cb8473946

Observation 4f13b1f1-b300-4206-8e75-309ffe398716 · outbound

This paper cites Venkatesh Babu, and Danish Pruthi.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Venkatesh Babu, and Danish Pruthi

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.451398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.526751Z digest=sha256:c12db64fefd7b458d30e20ebc582e901392d7953ca28ce3563e072aaa7cff2c5

Observation 91957dbb-4d85-4206-aead-d19ce5c8ec25 · outbound

This paper cites From local concepts to univer- sals: Evaluating the multicultural understanding of vision- language models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation From local concepts to univer- sals: Evaluating the multicultural understanding of vision- language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.434223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.534440Z digest=sha256:70a735cfba14e82199903ce9ee27cd656a05eed2452a3812546cb9d8c142d354

Observation 97432fa6-e243-42be-9367-9868606a8645 · outbound

This paper cites Extrinsic evaluation of cultural competence in large language models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Extrinsic evaluation of cultural competence in large language models

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.419391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.540424Z digest=sha256:48b7433dc5372373e5359c53c3ca7568d2ba4096dee0f5e6334ea0ba772276ac

Observation ce25a871-4a72-407e-9cf4-76306ed76a63 · outbound

This paper cites Language (technology) is power: A critical sur- vey of “bias” in NLP.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Language (technology) is power: A critical sur- vey of “bias” in NLP

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.400342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.545812Z digest=sha256:f40f37f377754ddfab70599b00afaeec925f5c2c4c431424e8a21419cdaeda3a

Observation b419d52b-e392-4a3a-9732-fd291315fdf8 · outbound

This paper cites Assessing Cross-Cultural Alignment between ChatGPT and Human Societies: An Empirical Study.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Assessing Cross-Cultural Alignment between ChatGPT and Human Societies: An Empirical Study

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.550399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.550399Z digest=sha256:4c01ee0a3d4d18e250e1979b4902eafd074ef74a816e9639c267b615c77d8a63

Observation 9f93418e-cef3-412c-af0f-cd8c33a085aa · outbound

This paper cites MaXM: Towards multilingual visual question an- swering.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation MaXM: Towards multilingual visual question an- swering

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.244100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.556238Z digest=sha256:3cf1af4f96853a964df55b59897b749cfcea50b4936f45de8cf79e07f2c745a6

Observation bd15319a-ccc5-41a6-b263-edc9fac94030 · outbound

This paper cites The SAGE handbook of intercultural competence.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation The SAGE handbook of intercultural competence

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.217411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.566255Z digest=sha256:8dd9ac749e1c6f6801a9b8cec7d466d0fd5f010fb8c208f21d8daf3e37a099fa

Observation 44af451b-f7bc-47fb-ac35-773a8126fcff · outbound

This paper cites Towards Measuring the Representation of Subjective Global Opinions in Language Models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Towards Measuring the Representation of Subjective Global Opinions in Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.570715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.570715Z digest=sha256:753846d108251ed4af97ec7a825f9778f98070b9a9b803f9e738fd362e3347c9

Observation f5c7ff0f-a9b3-41d4-9292-2a671d4991b2 · outbound

This paper cites ”i wouldn’t say offensive but...”: Disability-centered perspectives on large language models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation ”i wouldn’t say offensive but...”: Disability-centered perspectives on large language models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.201890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.575968Z digest=sha256:72791b5253be8a7baa078d7923733323e540ddc786b1f9a69df70d4247f23ca2

Observation 36a82315-2055-48ba-be2e-18e18cb78857 · outbound

This paper cites World values survey: Round seven - country-pooled datafile version 5.0, 2022.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation World values survey: Round seven - country-pooled datafile version 5.0, 2022

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.183490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.581681Z digest=sha256:ae653abe0e095a131d110fe92419e767007dbd9aa81d598da364fadcb995306d

Observation 6a98888b-6b7f-41f2-9525-77fb27da3e41 · outbound

This paper cites Challenges and strategies in cross- cultural NLP.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Challenges and strategies in cross- cultural NLP

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.161813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.586036Z digest=sha256:c9a9048fe0cfbe9f44f57a1ede6b1fb7fd50c938078c264e6c740aa490882fb4

Observation ab7e2c3b-6d81-492f-8621-97fa4e3f144a · outbound

This paper cites Clipscore: A reference-free evaluation met- ric for image captioning, 2022.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Clipscore: A reference-free evaluation met- ric for image captioning, 2022

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.147569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.591603Z digest=sha256:996e7e66e7512b75572755f54bcd1f3713be6ffa008e0cf2afd99c90a3574a81

Observation 8a2bcde3-1a01-4517-bc7d-f93ac753c396 · outbound

This paper cites Dimensionalizing cultures: The hofstede model in context.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Dimensionalizing cultures: The hofstede model in context

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.133979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.601101Z digest=sha256:46e6796688cdc21a06cb04d2ec1b22c77aa7f840661f0e3f747014e6ca1eb6fa

Observation 07b5b54a-1e3c-40a4-961c-e0fe3acb6855 · outbound

This paper cites The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.605724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.605724Z digest=sha256:c30da6074d705044ef1d0639e78620ed204da96a010d16f3bfe707db939deeb5

Observation 51157317-dd22-42c3-8893-0d9745f895a1 · outbound

This paper cites Visually 9 grounded reasoning across languages and cultures.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Visually 9 grounded reasoning across languages and cultures

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.120240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.613174Z digest=sha256:226bf321c74c53d008837f5422834f110d6869e00102cc1c62e7ef59023f2f4c

Observation 7b1f3961-56ea-4081-9caa-e6ffaeb19ea3 · outbound

This paper cites Wong, Qing- song Wen, Lichao Sun, Haipeng Chen, Xing Xie, and Jin- dong Wang.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Wong, Qing- song Wen, Lichao Sun, Haipeng Chen, Xing Xie, and Jin- dong Wang

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.106486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.619857Z digest=sha256:d659811c7a41ad90caa8f66a953c834f4e580a223cdbdddbd93c87ffae30636d

Observation feb9ebcd-dc92-4475-8b1b-1c7fc41cec29 · outbound

This paper cites Smolvlm: Re- defining small and efficient multimodal models, 2025.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Smolvlm: Re- defining small and efficient multimodal models, 2025

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.089084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.624563Z digest=sha256:6c3ddb7b865984e30bc8bc93049266bb3841ffff53633ccd5b25a27ed43b6d05

Observation 382ea78f-496d-486c-b6c1-0567dd6a0531 · outbound

This paper cites Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.632259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.632259Z digest=sha256:8c0e65bb2f9d27e97e9e7a5dcc6545a84801df05041ccd2b94a020e3af245565

Observation b035525d-fed5-4b4d-b39f-9d007f0fa8ef · outbound

This paper cites Assessing demographic bias in named entity recognition, 2020.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Assessing demographic bias in named entity recognition, 2020

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.071548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.638211Z digest=sha256:ac5835f2052e9e656af54ae96963f095d28aa11908be6320f7da0b980e2d6d87

Observation e4b1d5bd-0ba2-4192-bea4-f45810b8508b · outbound

This paper cites Benchmarking vision language models for cultural understanding.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Benchmarking vision language models for cultural understanding

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.053582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.642962Z digest=sha256:f557ac3f2a0670ebf74c789d8e6a1cdc2e7c54df6c65b3984fd5c48da8a08a75

Observation 56c09c94-b480-49b8-b306-f218cdf8f666 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Bleu: a method for automatic evaluation of machine translation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.034075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.647539Z digest=sha256:1fc0ccfd056560ff14d99a3fc2a65f235904fc35d6d5d7d552718af80bf383cf

Observation 6f975014-75c2-46a3-9868-a7711e77cba0 · outbound

This paper cites Knowledge of cultural moral norms in large language models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Knowledge of cultural moral norms in large language models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.655207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.655207Z digest=sha256:49ddce7a9875d7229b8637a26955f4503899d7f127e0814d7be30f424104f879

Observation 1694d9d3-b235-4ed8-beb3-5dc8c4667c2c · outbound

This paper cites Normad: A benchmark for measuring the cultural adaptability of large language mod- els.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Normad: A benchmark for measuring the cultural adaptability of large language mod- els

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.007915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.664060Z digest=sha256:da239d61538220200c352e4d8c32e740060c5d9a8e3c499691df066182159dad

Observation bcbf7f3f-335a-4d07-8dfa-752d6484d6a3 · outbound

This paper cites Cvqa: culturally-diverse multilingual visual question answering benchmark.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Cvqa: culturally-diverse multilingual visual question answering benchmark

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.988507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.669594Z digest=sha256:95b63a59eeca20b37a7eb8adffe61162cc66fe657f3137623b3f6834167eb6e6

Observation 4caac69a-eeaa-4e52-9984-a26eca4641b5 · outbound

This paper cites Geographical erasure in language generation.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Geographical erasure in language generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.969879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.676360Z digest=sha256:a2eca1cd6e259252fd94ec05086411ffd6807b88f662a638da1910a2c43a2b00

Observation 8ea07621-8980-4273-b82c-8f15f33284f4 · outbound

This paper cites Cultural bias and cultural alignment of large language mod- els.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Cultural bias and cultural alignment of large language mod- els

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.953680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.683290Z digest=sha256:202f5c4d1a1d866ee6f3159ed2a28df7b8a30aa49e5cf73092740b5e39832b26

Observation 79701dd4-3670-4960-82c9-139fbec99289 · outbound

This paper cites an unresolved cited work.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-05T17:15:49.933681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.689074Z digest=sha256:9d5ae5328cb36fa7825bcb926a4b891a0cdadfb4cadb48d7ac2c182ee2551967

Observation cc5f8c52-4426-499d-8f33-852afcc446fe · outbound

This paper cites Broaden the vision: Geo-diverse visual com- monsense reasoning.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Broaden the vision: Geo-diverse visual com- monsense reasoning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.914244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.694750Z digest=sha256:3f0353e59cb8109c5b5374bbce609c3a6c0ae0537fdf240d365063651784c2d3

Observation 76a60ec6-aed8-4285-85f8-955bc4c04e0c · outbound

This paper cites Internvl3: Exploring advanced training and test-time recipes for open-source multimodal models, 2025.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Internvl3: Exploring advanced training and test-time recipes for open-source multimodal models, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.897339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T17:15:49.702336Z digest=sha256:32ed13c2e70305552d97715401c44e1d1075f4637cb592b2adca2b7007d4b774

Observation 0e39aa5f-683e-468c-932e-56a15a210dc5 · outbound

This paper cites an unresolved cited work.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Unresolved cited work

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.561069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.561069Z digest=sha256:76f8ffdf5b6098396e0b960f56b66bfed9aefe1a012bfebbf7e0be3409d787bd

Pith citing papers

No inbound Pith citation observations are available.