Pith. sign in

Paper Citation Record · LEDGER

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation

As of 19 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2508.16762.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.16762 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T17:15:49.702336Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy26
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b2c894af-23a1-4d9e-874a-f1d50757802f · outbound

This paper cites To- wards measuring and modeling “culture” in LLMs: A sur- vey.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation To- wards measuring and modeling “culture” in LLMs: A sur- vey

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.507144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.487041Z digest=sha256:b64c137c2c5ee021ade55cfa91b82d69fc89db3649829b2600b0cc0797c13142

Observation e6385693-90d8-416f-9ec7-ffe9cbb1395b · outbound

This paper cites Investigating Cultural Alignment of Large Language Models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Investigating Cultural Alignment of Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.496287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.496287Z digest=sha256:684d29ed3b7e1e5a88853853c6297813b9039cd1f24cb367159ef90119af68bb

Observation e26cacf3-710f-4600-b4af-5286c44cab1c · outbound

This paper cites Probing Pre-Trained Language Models for Cross-Cultural Differences in Values.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Probing Pre-Trained Language Models for Cross-Cultural Differences in Values

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.502065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.502065Z digest=sha256:fa62e00d248f0ebead506057a88b9f3c3e12298a598900e307fdf3fd2027e996

Observation eeff2b56-5420-42a3-a3ee-a0d9ff2c354e · outbound

This paper cites Probing pre-trained language models for cross-cultural dif- ferences in values.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Probing pre-trained language models for cross-cultural dif- ferences in values

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.488957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.511930Z digest=sha256:c949a47a3f05bc348dac300951a87edd55e7c75cc51faa2a974c45d12d7ef696

Observation 66a2cf49-afc4-4e96-b9ad-aeec69601ceb · outbound

This paper cites Qwen2.5-vl technical report, 2025.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Qwen2.5-vl technical report, 2025

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.470713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.518870Z digest=sha256:c735795bbefc008b0b814407978727de6abab63421edf46072f8557ce58c93bf

Observation 4f13b1f1-b300-4206-8e75-309ffe398716 · outbound

This paper cites Venkatesh Babu, and Danish Pruthi.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Venkatesh Babu, and Danish Pruthi

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.451398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.526751Z digest=sha256:5a8aa9cc7bdfc33479d26ca56245a978b95188a03ce96e3d0ded44cbcd20aba5

Observation 91957dbb-4d85-4206-aead-d19ce5c8ec25 · outbound

This paper cites From local concepts to univer- sals: Evaluating the multicultural understanding of vision- language models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation From local concepts to univer- sals: Evaluating the multicultural understanding of vision- language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.434223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.534440Z digest=sha256:8c9fc0f1eadf3beac2a7d0d46a7a4a07274529cd033061770a6a04538304ed8b

Observation 97432fa6-e243-42be-9367-9868606a8645 · outbound

This paper cites Extrinsic evaluation of cultural competence in large language models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Extrinsic evaluation of cultural competence in large language models

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.419391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.540424Z digest=sha256:73c3b9bb2fe704d49a2d80d70ea5a000ac2ac7a7299ccef844aca78f17339196

Observation ce25a871-4a72-407e-9cf4-76306ed76a63 · outbound

This paper cites Language (technology) is power: A critical sur- vey of “bias” in NLP.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Language (technology) is power: A critical sur- vey of “bias” in NLP

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.400342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.545812Z digest=sha256:31890934958343debf40af2270beadc580c6617c6733d7f3d39e526becab00e4

Observation b419d52b-e392-4a3a-9732-fd291315fdf8 · outbound

This paper cites Assessing Cross-Cultural Alignment between ChatGPT and Human Societies: An Empirical Study.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Assessing Cross-Cultural Alignment between ChatGPT and Human Societies: An Empirical Study

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.550399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.550399Z digest=sha256:5d9ab506feb96aba12c93c2577a88db10422f038b7ee4e2b694fc9401dbcc193

Observation 9f93418e-cef3-412c-af0f-cd8c33a085aa · outbound

This paper cites MaXM: Towards multilingual visual question an- swering.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation MaXM: Towards multilingual visual question an- swering

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.244100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.556238Z digest=sha256:d1240938a8c47baff4af61e01c482fbab2a15f70702d738ef0e30c8e014be53b

Observation bd15319a-ccc5-41a6-b263-edc9fac94030 · outbound

This paper cites The SAGE handbook of intercultural competence.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation The SAGE handbook of intercultural competence

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.217411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.566255Z digest=sha256:75f4b2ad21c1b3398423ff85c5be7f8934a3e1d479955760f2f688bd18c282b4

Observation 44af451b-f7bc-47fb-ac35-773a8126fcff · outbound

This paper cites Towards Measuring the Representation of Subjective Global Opinions in Language Models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Towards Measuring the Representation of Subjective Global Opinions in Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.570715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.570715Z digest=sha256:7e1c89a878b10b3c370975195fe11c12efa15493cfad74f57e67cbd3d599411c

Observation f5c7ff0f-a9b3-41d4-9292-2a671d4991b2 · outbound

This paper cites ”i wouldn’t say offensive but...”: Disability-centered perspectives on large language models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation ”i wouldn’t say offensive but...”: Disability-centered perspectives on large language models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.201890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.575968Z digest=sha256:7328109b605ea6f8ca8194e9ecfbbfbcceb2f290283a1705e5d6bf90600f538f

Observation 36a82315-2055-48ba-be2e-18e18cb78857 · outbound

This paper cites World values survey: Round seven - country-pooled datafile version 5.0, 2022.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation World values survey: Round seven - country-pooled datafile version 5.0, 2022

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.183490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.581681Z digest=sha256:b8e41985e0a6c7dfa21709988654104548e21137d66d45bab603029ff8a4b6ed

Observation 6a98888b-6b7f-41f2-9525-77fb27da3e41 · outbound

This paper cites Challenges and strategies in cross- cultural NLP.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Challenges and strategies in cross- cultural NLP

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.161813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.586036Z digest=sha256:43651459c0238526a77bc1ade4a8967a10bdc404664eac13698562f3af30b55f

Observation ab7e2c3b-6d81-492f-8621-97fa4e3f144a · outbound

This paper cites Clipscore: A reference-free evaluation met- ric for image captioning, 2022.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Clipscore: A reference-free evaluation met- ric for image captioning, 2022

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.147569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.591603Z digest=sha256:9c364b4c4f57da3479228689c27720739d76960ef3a517612942a0b9f5afcd82

Observation 8a2bcde3-1a01-4517-bc7d-f93ac753c396 · outbound

This paper cites Dimensionalizing cultures: The hofstede model in context.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Dimensionalizing cultures: The hofstede model in context

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.133979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.601101Z digest=sha256:383b9455ce8e2520c62b5bd15ad37022c04e3c1285bc0423007ac5af83c80b72

Observation 07b5b54a-1e3c-40a4-961c-e0fe3acb6855 · outbound

This paper cites The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.605724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.605724Z digest=sha256:103e1e824c82ce5c174722a6611ee0077d5d6d2ee3da1ed3238a6b35c825e167

Observation 51157317-dd22-42c3-8893-0d9745f895a1 · outbound

This paper cites Visually 9 grounded reasoning across languages and cultures.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Visually 9 grounded reasoning across languages and cultures

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.120240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.613174Z digest=sha256:4277a42bc120ca4f04ab2fce0e5b6d4289d3553f50632ea2f41b21f3e48834fb

Observation 7b1f3961-56ea-4081-9caa-e6ffaeb19ea3 · outbound

This paper cites Wong, Qing- song Wen, Lichao Sun, Haipeng Chen, Xing Xie, and Jin- dong Wang.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Wong, Qing- song Wen, Lichao Sun, Haipeng Chen, Xing Xie, and Jin- dong Wang

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.106486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.619857Z digest=sha256:82ac4bf62c921b63531440f51a0678cfa3f0af71870b8e84f1dcb560008a8142

Observation feb9ebcd-dc92-4475-8b1b-1c7fc41cec29 · outbound

This paper cites Smolvlm: Re- defining small and efficient multimodal models, 2025.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Smolvlm: Re- defining small and efficient multimodal models, 2025

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.089084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.624563Z digest=sha256:1e3b21011b58c871e2973f05d535ca4867030932b4e736489be021e31513965a

Observation 382ea78f-496d-486c-b6c1-0567dd6a0531 · outbound

This paper cites Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.632259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.632259Z digest=sha256:60215a061c4dfb78f686a50a0a8a6fe1cd26b6c159c4f6201bc6893b4bde7271

Observation b035525d-fed5-4b4d-b39f-9d007f0fa8ef · outbound

This paper cites Assessing demographic bias in named entity recognition, 2020.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Assessing demographic bias in named entity recognition, 2020

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.071548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.638211Z digest=sha256:c6004bb8e515a3163d5f0e7626ceec8a126fe9266a69512ec3d510bb69b527d1

Observation e4b1d5bd-0ba2-4192-bea4-f45810b8508b · outbound

This paper cites Benchmarking vision language models for cultural understanding.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Benchmarking vision language models for cultural understanding

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.053582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.642962Z digest=sha256:e87171a60e8079142232242889e53e0e50feb0e111a574fd2fa53eb775a0710f

Observation 56c09c94-b480-49b8-b306-f218cdf8f666 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Bleu: a method for automatic evaluation of machine translation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.034075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.647539Z digest=sha256:8126a990ba31f69f3c687ffe8d328720322ec35725ff0a200d1789671dd6fa5b

Observation 6f975014-75c2-46a3-9868-a7711e77cba0 · outbound

This paper cites Knowledge of cultural moral norms in large language models.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Knowledge of cultural moral norms in large language models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.655207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.655207Z digest=sha256:9f327f64e73b1293250f6ce19e90a1341afb207cf225b81f8abe4315bfcd9255

Observation 1694d9d3-b235-4ed8-beb3-5dc8c4667c2c · outbound

This paper cites Normad: A benchmark for measuring the cultural adaptability of large language mod- els.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Normad: A benchmark for measuring the cultural adaptability of large language mod- els

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:50.007915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.664060Z digest=sha256:f12df0821f54f871ea305266f8b65cbf448d70104f0db8bf1fa2b61fecdec0c3

Observation bcbf7f3f-335a-4d07-8dfa-752d6484d6a3 · outbound

This paper cites Cvqa: culturally-diverse multilingual visual question answering benchmark.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Cvqa: culturally-diverse multilingual visual question answering benchmark

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.988507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.669594Z digest=sha256:132405bedb4dafadf965675a8f0d47d996debb5e0de4d47010c348bb30c2f0b3

Observation 4caac69a-eeaa-4e52-9984-a26eca4641b5 · outbound

This paper cites Geographical erasure in language generation.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Geographical erasure in language generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.969879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.676360Z digest=sha256:8850a193699d45e1cc8e18ff05f57116d28d5d5c177ff1647ec736ba12de7f22

Observation 8ea07621-8980-4273-b82c-8f15f33284f4 · outbound

This paper cites Cultural bias and cultural alignment of large language mod- els.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Cultural bias and cultural alignment of large language mod- els

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.953680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.683290Z digest=sha256:97c9f06ba0b54b524646690431111d589e03baacadcfccbf2ffa552559cd4d8e

Observation 79701dd4-3670-4960-82c9-139fbec99289 · outbound

This paper cites an unresolved cited work.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-05T17:15:49.933681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.689074Z digest=sha256:fe1041ba6d7c29ef89c5b5c8f1f0decd1b03107f1a1dca079e19c8264b1d78b9

Observation cc5f8c52-4426-499d-8f33-852afcc446fe · outbound

This paper cites Broaden the vision: Geo-diverse visual com- monsense reasoning.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Broaden the vision: Geo-diverse visual com- monsense reasoning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.914244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.694750Z digest=sha256:c8e6a8ea397da29c6be8ed485332e297fe1f5e54350f1b91feb62cca56cc1641

Observation 76a60ec6-aed8-4285-85f8-955bc4c04e0c · outbound

This paper cites Internvl3: Exploring advanced training and test-time recipes for open-source multimodal models, 2025.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Internvl3: Exploring advanced training and test-time recipes for open-source multimodal models, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:15:49.897339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T17:15:49.702336Z digest=sha256:27b5843835e6b834ae63bc66b74a0d79197a0668583de6d333824bf8cfddec2b

Observation 0e39aa5f-683e-468c-932e-56a15a210dc5 · outbound

This paper cites an unresolved cited work.

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation Unresolved cited work

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T17:15:49.561069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:15:49.561069Z digest=sha256:5bdc79393b6b96d6aa8a5a38369daffac2a0611954d185a4d771bc047e1b78d2

Pith citing papers

No inbound Pith citation observations are available.