Pith. sign in

Paper Citation Record · LEDGER

CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2410.02677.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.02677 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:52:40.917468Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T07:08:07.004760Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 035a9400-9f93-4642-8009-c83a44ba2d8e · inbound

Language Matters: How Do Multilingual Input and Reasoning Paths Affect Large Reasoning Models? cites this paper.

Language Matters: How Do Multilingual Input and Reasoning Paths Affect Large Reasoning Models? CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:40.917468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:40.917468Z digest=sha256:a8776bc2570932f60559a74cd8c7f8b28e3a783b87eb0fcf1ca945ed6d239d6e

Observation 3f844e7c-31c0-4c5b-9464-60be003f435f · inbound

CulFiT: A Fine-grained Cultural-aware LLM Training Paradigm via Multilingual Critique Data Synthesis cites this paper.

CulFiT: A Fine-grained Cultural-aware LLM Training Paradigm via Multilingual Critique Data Synthesis CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:26.559303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:17:26.559303Z digest=sha256:a108edb9765d5381533d32b9db4cf9db42e438f627cbc8fa1b56c7755096ae43

Observation d6e1a612-e347-49a3-b5e8-66b55159d98f · inbound

Disentangling Language and Culture for Evaluating Multilingual Large Language Models cites this paper.

Disentangling Language and Culture for Evaluating Multilingual Large Language Models CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.577676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.577676Z digest=sha256:3fc1dae9b49533567cbbef2abb933674cb90667f4d1ae03000e2e5bcca8e063b

Observation 6deff245-bd09-4939-867c-5908395d103e · inbound

BenchHub: A Unified Benchmark Suite for Holistic and Customizable LLM Evaluation cites this paper.

BenchHub: A Unified Benchmark Suite for Holistic and Customizable LLM Evaluation CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:10.604298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:08:10.604298Z digest=sha256:8eb8d486223c9e4a7467c9b70bd52d559c9a0a7080581acdeb2402ba31964b6c

Observation f3af707f-9044-46bc-83af-a6876ca35b78 · inbound

Marco-Bench-MIF: On Multilingual Instruction-Following Capability of Large Language Models cites this paper.

Marco-Bench-MIF: On Multilingual Instruction-Following Capability of Large Language Models CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T17:04:22.448754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:04:22.448754Z digest=sha256:3cd80cdddb329314266eb7e8dd96bc237dbf023d3048b15e92fb49e56a9c1d77

Observation 5d3cb2c3-efbb-4b55-ba5e-c7e9390f0745 · inbound

MyCulture: Exploring Malaysia's Diverse Culture under Low-Resource Language Constraints cites this paper.

MyCulture: Exploring Malaysia's Diverse Culture under Low-Resource Language Constraints CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T23:25:38.188067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:25:38.188067Z digest=sha256:7497eeadb8431e3bb1031886f6fadbbab65206701d96b2383da5e174c3d18d64

Observation 0fcd6827-6d75-4f62-b2c3-88fc8c886dfc · inbound

SEADialogues: A Multilingual Culturally Grounded Multi-turn Dialogue Dataset on Southeast Asian Languages cites this paper.

SEADialogues: A Multilingual Culturally Grounded Multi-turn Dialogue Dataset on Southeast Asian Languages CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T22:22:20.230941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:22:20.230941Z digest=sha256:59badc629b606edab8b619dbc19f1a1a32485da47d004834107beb556cdef2f8

Observation 628cc7dc-53e5-4456-b354-d9daafdaf60a · inbound

CultureSynth: A Hierarchical Taxonomy-Guided and Retrieval-Augmented Framework for Cultural Question-Answer Synthesis cites this paper.

CultureSynth: A Hierarchical Taxonomy-Guided and Retrieval-Augmented Framework for Cultural Question-Answer Synthesis CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:37.982081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:31:37.982081Z digest=sha256:ddd14789bac94f8bb7236d3f0d9293aa951af781d04a3f141a64f42b4851afa1

Observation 847a70e5-09c1-4df1-9ba9-7b39aef4b5ed · inbound

MoCo: A One-Stop Shop for Model Collaboration Research cites this paper.

MoCo: A One-Stop Shop for Model Collaboration Research CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:17:43.695656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T10:17:37.129753Z digest=sha256:b7febec83af506fc30a3bcafd47516cb4c608d67b1331751b84d0ad0b5181b0d

Observation f29d1516-39c1-4250-a220-9a0c3a1257e8 · inbound

Geographic Blind Spots in AI Control Monitors: A Cross-National Audit of Claude Opus 4.6 cites this paper.

Geographic Blind Spots in AI Control Monitors: A Cross-National Audit of Claude Opus 4.6 CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T08:35:19.309206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T08:30:47.680521Z digest=sha256:693c0535efe66ee3fbe0759d0e81093bf6b11396f0eb881e6a62e5fca097c99a

Observation baa2b86f-f80a-41f3-8df2-48b898dd0e85 · inbound

Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South cites this paper.

Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:08:07.006454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T07:06:57.555070Z digest=sha256:8e88ec5bcdd9dec1bb2932d23fa84e5f41832c052503c2e52f44e42c8bbbc9d8

Observation 1f2adfbd-078e-45a4-bd4e-a002195e50b2 · inbound

Prompt Robustness Is Task-Dependent: Comparing Objective and Belief-Style Questions in LLM Evaluation cites this paper.

Prompt Robustness Is Task-Dependent: Comparing Objective and Belief-Style Questions in LLM Evaluation CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-11T05:52:45.392115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T05:52:45.392115Z digest=sha256:50eaa6759a6150e363a57b465ed3da502b94c31e36e0ca489f8ab3323be0b71f