Pith. sign in

Paper Citation Record · LEDGER

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations

As of 9 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2506.11114.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11114 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:39:42.506184Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 060c174e-c991-4b8e-87d0-2019120b2618 · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:45.354572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:40.369014Z digest=sha256:9ed4862f69a7cdc9b5b0024b81fa966bc48032264bb2cf66ac3b34f707ede258

Observation 27856649-cae8-4e56-911f-f832df18cab1 · outbound

This paper cites Computer software (2024),https://acrobat.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Computer software (2024),https://acrobat

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:45.340053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:40.467025Z digest=sha256:be1dc305a5041e99b0aa94e9b190f931746e51fa6e7b2f51b52f611ff3362bb4

Observation 96fda058-952c-40bf-8a73-30b661501932 · outbound

This paper cites Clinical Anatomy38(2), 186–199 (2025).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Clinical Anatomy38(2), 186–199 (2025)

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:45.323735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:40.638781Z digest=sha256:5442fe4547a7af17434e0021ee490fab9ff98a2844f30121d678b77f7c0bb77a

Observation 3795bcb1-f27c-41a8-91e8-033cfa954105 · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:45.307733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:40.809202Z digest=sha256:6167dd6a8fe38dfab8ef329d16f5335c48d93b4325386743b542bdd2924dc2e5

Observation b5464ab6-122e-485f-b741-99726d5abbfe · outbound

This paper cites ishiyaku-dental.jp/, accessed: 2025-06-05.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations ishiyaku-dental.jp/, accessed: 2025-06-05

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:45.292741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:41.025291Z digest=sha256:0eb4e84dccd20105c4930a36c028b4d316b820e820c4cddc23165ff68325838e

Observation 464a702b-addd-466d-9fb8-246070c98a74 · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:45.088649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:41.204662Z digest=sha256:5781f3adf96c21faf558e3cb72a210c532c87e2931a5f05bfa08b75f6325e9f4

Observation 94858ee3-b033-4a99-ad26-05bfe244faa4 · outbound

This paper cites mynavi.jp/conts/kokushi_kouryaku/, accessed: 2025-06-05 KokushiMD-10 9.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations mynavi.jp/conts/kokushi_kouryaku/, accessed: 2025-06-05 KokushiMD-10 9

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:44.891693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:41.344767Z digest=sha256:1e8469f63e0e24979b3034b26b088ef930b898f86658c3672e7abb567e8ad9d0

Observation 9b2af865-a51f-4637-9c95-60be196d77d3 · outbound

This paper cites Applied Sciences11(14), 6421 (2021).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Applied Sciences11(14), 6421 (2021)

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:39:41.416728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:39:41.416728Z digest=sha256:e3e323dc168c489badef2a2a543d81131f28ab7550a8e26d831d1e0611a97319

Observation 15ea280b-ca5b-4968-bfa2-489e37eabb39 · outbound

This paper cites Evaluating GPT-4 and ChatGPT on Japanese Medical Licensing Examinations.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Evaluating GPT-4 and ChatGPT on Japanese Medical Licensing Examinations

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:39:41.561909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:39:41.561909Z digest=sha256:dc5e3333bdc894fd803327f171bc4c21260d43908f757ba55600d596919ef96c

Observation 6eb9a93d-b5df-4203-8c67-3eb414012bdd · outbound

This paper cites Advances in Neural Information Processing Systems 36, 52430–52452 (2023).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Advances in Neural Information Processing Systems 36, 52430–52452 (2023)

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:44.516979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:41.728068Z digest=sha256:8f40fb6637ddffccda928f07ca6f1c1f4feced02a624274138f28f756f53770d

Observation 25469ee5-777c-4eff-81d7-19685c55b6de · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:44.269056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:41.844522Z digest=sha256:59b793912d43a653022e52900afa55a955207a7cb832b8b70a15a167bb733c9a

Observation 71c0fcb2-ec5c-4f2f-a345-4377efcd3715 · outbound

This paper cites html, accessed 2025-06-07.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations html, accessed 2025-06-07

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:44.016572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:41.879923Z digest=sha256:901a5119f38e21cc90eea2a3d733b14789412717686c2ff68ecc6404ab91caa2

Observation 88b648ca-0929-4668-a34f-202c24122b8a · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:43.756452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:41.927605Z digest=sha256:53ecc9c2a3da01259c182429fa09bdacae8a5d8e597e2aa2d35add3a9dd985f6

Observation fed79d2e-52f3-4942-8150-f4e0437c148a · outbound

This paper cites JJDEA40, 3–10 (2024).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations JJDEA40, 3–10 (2024)

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:43.589654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:41.989459Z digest=sha256:7b197abaa10ded5e05236b6d73a99aa3e13f46873acd18d63df629c079c80cb1

Observation 7494be47-5966-49ea-8042-2d749708eb89 · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:43.479864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:42.034777Z digest=sha256:669e42d80fe8e69fa8ef483093c9c370c4b02c91d896f4d8fefdbf007a4fdead

Observation 13117dba-ec30-4cf4-aeca-4b095e4f3f0c · outbound

This paper cites In: Conference on health, inference, and learning.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations In: Conference on health, inference, and learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:43.328755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:42.087521Z digest=sha256:8aa2a9c551c4cb7df0ae2458aba6e50291a379dc77c507d24ddd4aede748fe61

Observation 46ab7c81-32c5-40c9-a13c-d56e4dcf76c1 · outbound

This paper cites guppy.jp/, accessed: 2025-06-05.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations guppy.jp/, accessed: 2025-06-05

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:43.154140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:42.148285Z digest=sha256:a972ce640da6bdea6c46413248a6a8f54c1c7a5bae47c0212c0e0811f287d76a

Observation 79da25e9-5225-4eaa-afb2-ec965edc9699 · outbound

This paper cites Bell System Technical Journal27(3), 379–423 (1948).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Bell System Technical Journal27(3), 379–423 (1948)

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:39:42.193816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:39:42.193816Z digest=sha256:bd5cae2bd22ee58823a353f0663c5d6c3e29c72b151d792661d7f8f6dda050ea

Observation fc12455e-3dac-4895-a750-c48eaca57604 · outbound

This paper cites Scientific Reports14(1), 9330 (2024).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Scientific Reports14(1), 9330 (2024)

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:43.040530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:42.253208Z digest=sha256:87e74e1a9d36231babfd3b6a2c0f21e6e7e1947b06130995939ca71c20f39010

Observation 7798c562-97c6-46ae-a591-2a1f0aa362b5 · outbound

This paper cites A Japanese Language Model and Three New Evaluation Benchmarks for Pharmaceutical NLP.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations A Japanese Language Model and Three New Evaluation Benchmarks for Pharmaceutical NLP

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:39:42.647730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:42.317690Z digest=sha256:827b73756b17b69c4b0db46262e035f8889e24f6cfbfb55ea810c3ef38749b41

Observation f3a30e7f-1d35-4b32-93e0-0085db6287a8 · outbound

This paper cites JMIR medical education9(1), e48002 (2023).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations JMIR medical education9(1), e48002 (2023)

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:42.891688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:42.396218Z digest=sha256:ed1aed7b6387cb46883459eb22f021471e3e4339312f4418db6e912f00d9db3b

Observation dceb8435-7d19-409c-bbb2-7283f6464edb · outbound

This paper cites Advances in Neural Information Processing Systems35, 24824–24837 (2022).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Advances in Neural Information Processing Systems35, 24824–24837 (2022)

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:42.767195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:39:42.448581Z digest=sha256:657dfc6a32d6c1cf62a14bf6625149693d752f4d8e6c30728277742dd65c9762

Observation 4c4cdb70-7b00-4997-831d-c99d919ca950 · outbound

This paper cites MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:39:42.506184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:39:42.506184Z digest=sha256:7792d967382e2f5928c955eb85502266a0c1f199faec4f1de1692f535d5ef1b1

Pith citing papers

No inbound Pith citation observations are available.