Pith. sign in

Paper Citation Record · LEDGER

KoBALT: Korean Benchmark For Advanced Linguistic Tasks

As of 19 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 2 inbound Pith citation observations for arXiv:2505.16125.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16125 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:10:18.937297Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:28:03.594544Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T09:39:54.392252Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d09e1769-c78d-4397-ab6a-66223b7e3020 · outbound

This paper cites Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:12.530788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:12.530788Z digest=sha256:06ff29872bc5ed71335cdcadd567c0ccc93ed3ddacf99f3fab7fbdeddd008871

Observation 04a168de-7543-45ab-8a25-5d5d2e6d2240 · outbound

This paper cites GPT-4 Technical Report.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:12.590589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:12.590589Z digest=sha256:8df06513c3530401157a004a271385df1f67312270e31cfaacaeda6807cba6cb

Observation ea0cd3aa-d7e4-4a3d-82d5-ae817dd04642 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:12.602073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:12.602073Z digest=sha256:4382e44e6f0f012d08e62d3d6e08a7b3727d9812e0439bfb4867f494bf223d90

Observation fc996665-495f-4aa0-9fac-7725fb779284 · outbound

This paper cites Language Models are Few-Shot Learners.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Language Models are Few-Shot Learners

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:12.697859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:12.697859Z digest=sha256:3f77885fd0f213eb1e35d0ffd8d0effcdda5dfed27d92818a7f07de3dd1fc7f5

Observation b232d927-0eaf-4de6-a479-afa29610c04f · outbound

This paper cites an unresolved cited work.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:10:20.532713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:10:12.840303Z digest=sha256:f2047164861ffbae78b149a9cd4756f9327bea11fd8d3f1349a8b6c873318f07

Observation d9523693-0231-4722-b725-bc6a7edab369 · outbound

This paper cites Holmes Recorder a benchmark to assess the linguistic competence of language models.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Holmes Recorder a benchmark to assess the linguistic competence of language models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:12.964354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:12.964354Z digest=sha256:7bc2d4a4b77b006a76c9de83a188971ee0b02bd1c9004c3451689fcfb4f0f0c6

Observation 89ffbc7e-642f-43f6-a77d-60cca45f4cf9 · outbound

This paper cites Syntaxgym: An online platform for targeted evaluation of language models.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Syntaxgym: An online platform for targeted evaluation of language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:10:20.499389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:10:13.135458Z digest=sha256:3ef1651e709f70926d6fe80ce1af5f98bd40e0adebc3200fb205c62794e962cc

Observation be8b3828-e619-47a5-a330-b72e87497b38 · outbound

This paper cites Linguini: A benchmark for language-agnostic linguistic reasoning.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Linguini: A benchmark for language-agnostic linguistic reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:13.254934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:13.254934Z digest=sha256:57d62fa649ab8869a892c7c0b26c37930423d4a9f8b980ff91eab48b9e8d4f69

Observation d9d1087e-4961-4f2f-85a1-30f6987d7bb2 · outbound

This paper cites Iolbench: Benchmarking llms on linguistic reasoning, 2025.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Iolbench: Benchmarking llms on linguistic reasoning, 2025

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:13.361162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:13.361162Z digest=sha256:645379d3e7afd92260a42b9071b08095ad954f305d1fe01c69f7f2d307fbf14f

Observation a102fb66-db7c-48e0-be66-57460d022240 · outbound

This paper cites LINGOLY: A Benchmark of Olympiad-Level Linguistic Reasoning Puzzles in Low-Resource and Extinct Languages.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks LINGOLY: A Benchmark of Olympiad-Level Linguistic Reasoning Puzzles in Low-Resource and Extinct Languages

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:10:19.679226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:10:13.474052Z digest=sha256:c591857956c60de48b5533e7e12ea18c3042812e46cb58fcfe4213dc58a4ddba

Observation d6f0e28a-a2e5-4511-9fba-4c3f5cc37021 · outbound

This paper cites an unresolved cited work.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:13.611956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:13.611956Z digest=sha256:ce58f115348f7f262c0a1d5b5ce94cba09c2ab4d2bbb6a4c6cd3454de0b0cba0

Observation dbc8975f-4a6e-48a3-a6a4-4eab4105c6d8 · outbound

This paper cites SuperGLUE: A Stickier Benchmark for General-Purpose Language Understanding Systems.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks SuperGLUE: A Stickier Benchmark for General-Purpose Language Understanding Systems

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:13.730046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:13.730046Z digest=sha256:c1925c68618f5dea6310ad8e552529564dbee5f36f77659ddb584b08deed66dc

Observation ae80c1e2-4b98-4f5a-aa22-da12bf5c2dc9 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Measuring Massive Multitask Language Understanding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:13.903057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:13.903057Z digest=sha256:5612af46ab6806671bb4149d4f0a5c687a60eeca1bc4f6edd2fa6a6d63443e09

Observation 74c689d1-1ff9-4eb4-91a3-d57cd9968a9e · outbound

This paper cites Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:14.024926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:14.024926Z digest=sha256:adbdfcec82beb6a251aff8983dafd060ec9b8f0ae8a762f12a24aa5826047554

Observation fb230159-194c-48a0-9f08-bc10033b3ba5 · outbound

This paper cites Pub: A pragmatics understanding benchmark for assessing llms’ pragmatics capabilities.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Pub: A pragmatics understanding benchmark for assessing llms’ pragmatics capabilities

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:10:20.443047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:10:14.103530Z digest=sha256:04240d5708641424f18f04b554488d426553facaf99d0bef382f0d7c25f14d7f

Observation 17042f30-904b-4c36-8644-9601ce4f5356 · outbound

This paper cites MultiPragEval: Multilingual Pragmatic Evaluation of Large Language Models.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks MultiPragEval: Multilingual Pragmatic Evaluation of Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:14.195846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:14.195846Z digest=sha256:6226c2de735f96939597fcc1343c7e4fe01e6efd30928f8bc68b3e4d75ed46a7

Observation ae376289-562c-45e4-970e-5b4f00c83c8c · outbound

This paper cites PhonologyBench: Evaluating Phonological Skills of Large Language Models.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks PhonologyBench: Evaluating Phonological Skills of Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:14.314620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:14.314620Z digest=sha256:9e3b13589fd698c74c04d4742656dba079ebd20ee0dc949c146a60b3e6467c63

Observation e4bdb0c9-e6ce-4a8a-a1f9-a4360de2a64e · outbound

This paper cites KLUE: Korean Language Understanding Evaluation.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks KLUE: Korean Language Understanding Evaluation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:14.396034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:14.396034Z digest=sha256:c15d394ef2012420e948e16e490ec2729a78a56eecc3e1b5fb68886e7153d89e

Observation 4298c0aa-d7b4-494f-bc30-3e688b699902 · outbound

This paper cites Open Ko-LLM Leaderboard: Evaluating Large Language Models in Korean with Ko-H5 Benchmark.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Open Ko-LLM Leaderboard: Evaluating Large Language Models in Korean with Ko-H5 Benchmark

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:14.473160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:14.473160Z digest=sha256:334cb1a73c8269a41555b63697884491984efe4f7c9ff945002461cd05c2277d

Observation 09125917-6f0b-42dd-aa1a-025f43e66566 · outbound

This paper cites K o BEST : K orean balanced evaluation of significant tasks.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks K o BEST : K orean balanced evaluation of significant tasks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:10:20.413853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:10:14.590704Z digest=sha256:e879fc3731162e40badf9e278d09c1a4d8606c9361c38d884d6de52b0f7dc00e

Observation 78f336cc-f88a-4412-bfcf-c35222234bfb · outbound

This paper cites HAE - RAE bench: Evaluation of K orean knowledge in language models.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks HAE - RAE bench: Evaluation of K orean knowledge in language models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:10:20.372281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:10:14.718539Z digest=sha256:d119bbdae1e3813ec5b7567e4ed304e12197ec54573dd5e2e01aa4eddb18a40f

Observation 39232473-5816-4950-be63-5103f5a0a70d · outbound

This paper cites CLI c K : A benchmark dataset of cultural and linguistic intelligence in K orean.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks CLI c K : A benchmark dataset of cultural and linguistic intelligence in K orean

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:10:20.345710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:10:14.803387Z digest=sha256:19ac6c51f8fffaa4e70931aad47077df72440643fa7a76ff0e0c91233dae5656

Observation 35e96e00-2ba7-45ae-822e-e26d5dd8ac7f · outbound

This paper cites The Llama 3 Herd of Models.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks The Llama 3 Herd of Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:14.853180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:14.853180Z digest=sha256:2540331856e85eff8b9cd91c0665e08c28b717d55a9484fbb9f91ac1aa449d2d

Observation 4ebaf431-d609-497d-81ba-1adec09bfda9 · outbound

This paper cites Mistral 7B.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Mistral 7B

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:14.966529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:14.966529Z digest=sha256:38727225960f62e3cf56c002b16e4752a652bd1e378c6cfbe950e69cf740b54c

Observation 5e5b084c-f3d3-4eca-82b2-e8052c8a1a06 · outbound

This paper cites Mistral small 3.1, 2025.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Mistral small 3.1, 2025

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:10:20.323052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:10:15.140331Z digest=sha256:afd0d57dbaea204df4f1e627583bce48ce99ca4ab787ee2454c982ca791129b6

Observation dafbaa34-a0d9-4c14-b483-e3b301200010 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Gemma 2: Improving Open Language Models at a Practical Size

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:15.220977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:15.220977Z digest=sha256:87dda4178da0f51da63f051cb91feea79259e887caffbbaa9113b3c21e762ac6

Observation 049f3b0d-f3ff-4a0d-b86c-327511addd75 · outbound

This paper cites Gemma 3 Technical Report.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Gemma 3 Technical Report

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:15.318958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:15.318958Z digest=sha256:55ea7ae1518140844133eaa416efb5ae17b5052dd9745c690830182a8619b940

Observation 48ee745a-4605-4fe7-b3c2-d0be1828dc62 · outbound

This paper cites Qwen2.5 Technical Report.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Qwen2.5 Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:15.388040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:15.388040Z digest=sha256:55d52ddf93acbec687157c4c2029d8f9dba0fd66430fddce6ef51a725381e40c

Observation ca64d69b-562e-4669-af73-72796e5b2478 · outbound

This paper cites Aya Expanse: Combining Research Breakthroughs for a New Multilingual Frontier.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Aya Expanse: Combining Research Breakthroughs for a New Multilingual Frontier

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:15.554873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:15.554873Z digest=sha256:545218347780bfaad7362340e9246714db0ca9b8753654cd5d939bf58f3707a0

Observation f7796360-c9fd-4f7a-8d14-b757aeced216 · outbound

This paper cites Claude 3.5 sonnet, 2024.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Claude 3.5 sonnet, 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:10:20.223974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:10:17.270321Z digest=sha256:20a013847592af8584783b6a4c5a9e501a8f00946fcb06f683dcadd72b77c793

Observation d8ead1df-48dc-47f6-810d-302f1513e76c · outbound

This paper cites Claude 3.7 sonnet, 2025.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Claude 3.7 sonnet, 2025

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:10:20.051466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:10:18.315796Z digest=sha256:55ad08fedbca29a6e324ec81b03f304cd1681200cc61418806bb706b3e4ff81a

Observation 27b4b4dd-739d-448c-a43a-8cb83dc50de8 · outbound

This paper cites Command A: An Enterprise-Ready Large Language Model.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Command A: An Enterprise-Ready Large Language Model

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:18.443678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:18.443678Z digest=sha256:1e43e9440481510b226fe882731e7b7fb807c43e3d8b43e3df5deddf97ff7e3d

Observation b2ed16ec-f54d-48c7-9cf4-e2674101cbce · outbound

This paper cites DeepSeek-V3 Technical Report.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks DeepSeek-V3 Technical Report

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:18.557472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:18.557472Z digest=sha256:3d1ed4f42a6435e049252f1f359ee3b9aa5c733975fb265295684f63e8c0e5e6

Observation c234c631-4ec7-4a6e-ab0b-a7ba9a6b34f5 · outbound

This paper cites Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:18.644197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:18.644197Z digest=sha256:0b3abd1debe87c3f1fd0608732979a15698fba4668d18e3cebec348c960da527

Observation b6fea1fb-d494-4cb3-90fa-606d102a79b7 · outbound

This paper cites Dissecting human and LLM preferences.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Dissecting human and LLM preferences

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:18.747732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:18.747732Z digest=sha256:6389ee0c171612fa839412513b3a9541e4908bc3026422278b7b355b28f49dc0

Observation 6e9a150b-4351-4986-8eb0-f49e26e01edd · outbound

This paper cites Uncovering factor level preferences to improve human-model alignment, 2024.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Uncovering factor level preferences to improve human-model alignment, 2024

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:18.890327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:18.890327Z digest=sha256:836ea507fe499ae9543a99be3d27d39875ebbd217514d7dc0737f75bd302fdc0

Observation af71041c-2a60-4759-af02-fc7b4135293b · outbound

This paper cites Human Feedback is not Gold Standard.

KoBALT: Korean Benchmark For Advanced Linguistic Tasks Human Feedback is not Gold Standard

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:18.937297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:10:18.937297Z digest=sha256:bcb07188be57cd08d857d3817b69c30704641425a2abfea7ba84a548ee7dc3fa

Pith citing papers

Observation 718795ec-14aa-4c83-93e0-9f634b776ed5 · inbound

HanjaBridge: Resolving Semantic Ambiguity in Korean LLMs via Hanja-Augmented Pre-Training cites this paper.

HanjaBridge: Resolving Semantic Ambiguity in Korean LLMs via Hanja-Augmented Pre-Training KoBALT: Korean Benchmark For Advanced Linguistic Tasks

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T17:28:03.594544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:28:03.594544Z digest=sha256:568a171dd282ee60786432cfe48508e24e7a217972e7aa91f5ccc192b2fe0e93

Observation 656a7e4d-8394-4154-b8f0-231d48098add · inbound

BanglaSocialBench: A Benchmark for Evaluating Sociopragmatic and Cultural Alignment of LLMs in Bangladeshi Social Interaction cites this paper.

BanglaSocialBench: A Benchmark for Evaluating Sociopragmatic and Cultural Alignment of LLMs in Bangladeshi Social Interaction KoBALT: Korean Benchmark For Advanced Linguistic Tasks

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:39:54.394861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T09:38:29.597205Z digest=sha256:4be8c5850cee51a70d7780256288f30cd8319155cd3a1ed5e28bf7429a0aec7a