Pith. sign in

Paper Citation Record · LEDGER

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations

As of 21 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 1 inbound Pith citation observation for arXiv:2506.11114.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11114 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:39:42.506184Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:21:13.443456Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-15T23:21:14.448985Z

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 060c174e-c991-4b8e-87d0-2019120b2618 · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:45.354572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:40.369014Z digest=sha256:c5f78a995fa3d5f8c305afb735fd77bd327efe44ceae1fecb7d4bb394c94a10b

Observation 27856649-cae8-4e56-911f-f832df18cab1 · outbound

This paper cites Computer software (2024),https://acrobat.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Computer software (2024),https://acrobat

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:45.340053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:40.467025Z digest=sha256:4fbd2112091eda072e30b73688236d173105046f30368ee4a5ea51030cce5a66

Observation 96fda058-952c-40bf-8a73-30b661501932 · outbound

This paper cites Clinical Anatomy38(2), 186–199 (2025).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Clinical Anatomy38(2), 186–199 (2025)

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:45.323735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:40.638781Z digest=sha256:eb972140c3ccc79ebd7b7a7ee8bf79456e737f337ca32429bc85f3a1486d2207

Observation 3795bcb1-f27c-41a8-91e8-033cfa954105 · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:45.307733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:40.809202Z digest=sha256:32b544fb28a9e1c9b76e657f37103eb612e1e12d335f0d05d88d4161172e5c22

Observation b5464ab6-122e-485f-b741-99726d5abbfe · outbound

This paper cites ishiyaku-dental.jp/, accessed: 2025-06-05.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations ishiyaku-dental.jp/, accessed: 2025-06-05

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:45.292741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:41.025291Z digest=sha256:a8a31b53e7b5b4b310ccabf0204f38e35ab8d2af202ff4d168c25e7f673aafa4

Observation 464a702b-addd-466d-9fb8-246070c98a74 · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:45.088649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:41.204662Z digest=sha256:bc2da46f34025d211e32f07dc91172fa9d4a96a7a7a4cb806f6146562cbed9ae

Observation 94858ee3-b033-4a99-ad26-05bfe244faa4 · outbound

This paper cites mynavi.jp/conts/kokushi_kouryaku/, accessed: 2025-06-05 KokushiMD-10 9.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations mynavi.jp/conts/kokushi_kouryaku/, accessed: 2025-06-05 KokushiMD-10 9

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:44.891693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:41.344767Z digest=sha256:3cab3ff68fd97da7ca285c9f2d75989f66e3f2c6bf6de9f2c2d9624df7b25d21

Observation 9b2af865-a51f-4637-9c95-60be196d77d3 · outbound

This paper cites Applied Sciences11(14), 6421 (2021).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Applied Sciences11(14), 6421 (2021)

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:39:41.416728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:39:41.416728Z digest=sha256:c6d6b8eb5cf65f071a221e994abc01d7ca1e50ffe03ce608720eae64b68365a2

Observation 15ea280b-ca5b-4968-bfa2-489e37eabb39 · outbound

This paper cites Evaluating GPT-4 and ChatGPT on Japanese Medical Licensing Examinations.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Evaluating GPT-4 and ChatGPT on Japanese Medical Licensing Examinations

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:39:41.561909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:39:41.561909Z digest=sha256:a5d00914677e367718cc62d70dea2ab1d3d5122423eaa297bf3f6bdd7d8198fb

Observation 6eb9a93d-b5df-4203-8c67-3eb414012bdd · outbound

This paper cites Advances in Neural Information Processing Systems 36, 52430–52452 (2023).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Advances in Neural Information Processing Systems 36, 52430–52452 (2023)

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:44.516979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:41.728068Z digest=sha256:fc2bd6a06b9ce2d1b19cbde8886a151f7f66bb22e3c8de34d98ce12c76b0153b

Observation 25469ee5-777c-4eff-81d7-19685c55b6de · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:44.269056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:41.844522Z digest=sha256:2b958f651a2dd18382ba5bbc4978c6cc3c062431961e20b8a6754061d64cfc70

Observation 71c0fcb2-ec5c-4f2f-a345-4377efcd3715 · outbound

This paper cites html, accessed 2025-06-07.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations html, accessed 2025-06-07

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:44.016572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:41.879923Z digest=sha256:8846e03eee2bafc84aecda7fa0c14db43db10b88afde0102be2390b076f49938

Observation 88b648ca-0929-4668-a34f-202c24122b8a · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:43.756452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:41.927605Z digest=sha256:3dc67a3efa8979edee7e5aeabf5f966c4c498e8c99d24f25401f5106e29f1217

Observation fed79d2e-52f3-4942-8150-f4e0437c148a · outbound

This paper cites JJDEA40, 3–10 (2024).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations JJDEA40, 3–10 (2024)

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:43.589654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:41.989459Z digest=sha256:76087a190f7f6038c55e5f9f6bd813daa69adfe0d805633a98615b6af8e65455

Observation 7494be47-5966-49ea-8042-2d749708eb89 · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:43.479864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:42.034777Z digest=sha256:fd216462b6d0456a744a1671eb16633eb29ceee5c5d2bc5da1527248d9cfe93f

Observation 13117dba-ec30-4cf4-aeca-4b095e4f3f0c · outbound

This paper cites In: Conference on health, inference, and learning.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations In: Conference on health, inference, and learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:43.328755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:42.087521Z digest=sha256:f3d3cf203798702fb2c2f060e77d48ff7de03b6d05d459e1e46e18575f639c69

Observation 46ab7c81-32c5-40c9-a13c-d56e4dcf76c1 · outbound

This paper cites guppy.jp/, accessed: 2025-06-05.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations guppy.jp/, accessed: 2025-06-05

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:43.154140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:42.148285Z digest=sha256:35a888abef64cf421648173b8d2a204bc4af4385334fcb86a53d168a8e33cf97

Observation 79da25e9-5225-4eaa-afb2-ec965edc9699 · outbound

This paper cites Bell System Technical Journal27(3), 379–423 (1948).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Bell System Technical Journal27(3), 379–423 (1948)

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:39:42.193816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:39:42.193816Z digest=sha256:4a07525d73203502bfd65d6a12ba3c69f4187e4951d132ebe67cc8648118006a

Observation fc12455e-3dac-4895-a750-c48eaca57604 · outbound

This paper cites Scientific Reports14(1), 9330 (2024).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Scientific Reports14(1), 9330 (2024)

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:43.040530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:42.253208Z digest=sha256:8022f10683531bd35ac0ab50a1f0435bea71991c1a98a0167cf5be7e86286817

Observation 7798c562-97c6-46ae-a591-2a1f0aa362b5 · outbound

This paper cites A Japanese Language Model and Three New Evaluation Benchmarks for Pharmaceutical NLP.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations A Japanese Language Model and Three New Evaluation Benchmarks for Pharmaceutical NLP

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:39:42.647730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:42.317690Z digest=sha256:314f94ee9f610ede59b8bb0b9cf5649bdc668bbbe0c692230dc7e51305f54db5

Observation f3a30e7f-1d35-4b32-93e0-0085db6287a8 · outbound

This paper cites JMIR medical education9(1), e48002 (2023).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations JMIR medical education9(1), e48002 (2023)

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:42.891688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:42.396218Z digest=sha256:2e8faf959275381a8e3e4ce0a46dfa691bdc84fdbdd8e529d82f086b57d8213b

Observation dceb8435-7d19-409c-bbb2-7283f6464edb · outbound

This paper cites Advances in Neural Information Processing Systems35, 24824–24837 (2022).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Advances in Neural Information Processing Systems35, 24824–24837 (2022)

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:42.767195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:39:42.448581Z digest=sha256:e76e1c2fe798baf39659e79579faad025726c37917d1ef61c713b6c97df2a144

Observation 4c4cdb70-7b00-4997-831d-c99d919ca950 · outbound

This paper cites MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:39:42.506184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:39:42.506184Z digest=sha256:dd717502427d024f44d7cf07fe84d0c638aa00a624c6a4db742f3dd021ae69c1

Pith citing papers

Observation be2dbc15-0c43-4ea1-85ab-a6312380732f · inbound

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models cites this paper.

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations

Reference 292

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:21:14.454515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-15T23:21:13.443456Z digest=sha256:73485f5a69550b0aff02645294751d5b0827adb1b169588253d8cb529ab1696b