Pith. sign in

Paper Citation Record · LEDGER

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving

As of 17 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2506.08349.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08349 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:19:33.245710Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact1
  • verified fuzzy16
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a78a2827-c586-419e-b5d6-48029857868f · outbound

This paper cites Phi-4 Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Phi-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.126829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.126829Z digest=sha256:ecd547917b35002e65e06379fd12aa70768baf7070e13201243dc3fd011769c1

Observation baaad06f-1a16-4614-83fb-429004cf21ea · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.130742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.130742Z digest=sha256:053e6e2117002f5e9e6af3bf59d0831672df01114292db6f479ac8c4ceb0b54b

Observation e6fd8e85-c75c-4a90-b8dd-048bf80acbe6 · outbound

This paper cites GPT-4 Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.134332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.134332Z digest=sha256:9f14981a4dabc481651028d3b9aa407aedb282df5a652131d925b7dfdf818ff5

Observation e1edc71c-ca9c-4b06-9973-ca92efbb74a8 · outbound

This paper cites an unresolved cited work.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:19:33.594125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.138904Z digest=sha256:51c85883ec71917e8f7ca813867fdbb5cc5da066f1d6935e973107053348d8c1

Observation 73a37a47-5b85-4ec6-8b57-89151924ac1a · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Gemini: A Family of Highly Capable Multimodal Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.142732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.142732Z digest=sha256:14e14e4664ee776947fff334510605bfd302c6408de334fb808eb756f475c017

Observation b8962460-be4d-4715-a13a-e32a85b439d3 · outbound

This paper cites Qwen Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Qwen Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.146797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.146797Z digest=sha256:85d989200e0320906b988ff6a1582a1cfc7a4429db27ddbf0f1bbbbfabaf9118

Observation a1e12a8d-ce13-4f17-b6a5-19d07d4e3f56 · outbound

This paper cites Overview of the medical question answering task at trec 2017 liveqa.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Overview of the medical question answering task at trec 2017 liveqa

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.586684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.150105Z digest=sha256:506d73b816abb0bd055f35f7dd8c59673ad0bb4ce62ad4d953d1c8f49b2dea0e

Observation e4be78c0-33bf-47be-a4e8-95f3a2c3a249 · outbound

This paper cites S., Englehart, M.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving S., Englehart, M

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.578962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.153164Z digest=sha256:68509217a651563f3cf62ce63f8da822d665676ac1535ac11127c3841318775d

Observation 5d9070cc-4f18-475a-a428-33118c443882 · outbound

This paper cites The unified medical language system (umls): integrating biomedical terminology.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving The unified medical language system (umls): integrating biomedical terminology

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.571091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.155904Z digest=sha256:f729439fbba44d0f866acb0c80b781bdc42f9563dc0c4b4de88d225b909fe5dc

Observation cbc39d3a-86ce-4bad-803a-46ab91b44a48 · outbound

This paper cites an unresolved cited work.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:19:33.562957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.158633Z digest=sha256:c188ed505b99a5dc3b779aa7f2eae7bfcf46bc082494b44d14f1499b262f7e32

Observation 72be59a9-6e3a-4dc2-bef3-47e54e26ec34 · outbound

This paper cites Medbench: A large-scale chinese benchmark for evaluating medical large language models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Medbench: A large-scale chinese benchmark for evaluating medical large language models

Reference 11

Resolution
verified exact
doi, observed 2026-08-07T05:19:33.554631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.161403Z digest=sha256:bbf429c7cdbcacf02351b6e7bf0308b4bdc1099e454aa0d0fe9b254aefb550fb

Observation 7b449a13-136e-4b1c-8e37-a8137733ce75 · outbound

This paper cites MEDITRON-70B: Scaling Medical Pretraining for Large Language Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving MEDITRON-70B: Scaling Medical Pretraining for Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.164270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.164270Z digest=sha256:e983f16609f26dea46fd775a3a9e2bad3ced81fdceddb3bf512f67c83ef1a677

Observation 8d05d94c-dc3b-458f-b289-ff7c5e008b1b · outbound

This paper cites U., Pimentel, M.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving U., Pimentel, M

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.546163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.167785Z digest=sha256:0526d9116785db1f04c8f567990cab915f229816d483c58b75b6be10c9e46f9b

Observation 9a8cf196-9e57-404c-9580-351919d5c879 · outbound

This paper cites Med42-v2: A Suite of Clinical LLMs.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Med42-v2: A Suite of Clinical LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.170913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.170913Z digest=sha256:702b1fd8c92f1fe6673810e35d2867a319898df2059de1e0f8c34f45121d59cc

Observation 2f11382f-c86c-4fa6-a23b-6ceab1eb99e8 · outbound

This paper cites The Llama 3 Herd of Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving The Llama 3 Herd of Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.173687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.173687Z digest=sha256:e3402555fb4d746ad42f639920af155ae29c8f5f753fd37a6596dd3c33a05240

Observation 105f22a7-a464-4273-91fb-2c2102d96403 · outbound

This paper cites Evaluation and mitigation of the limitations of large language models in clinical decision-making.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Evaluation and mitigation of the limitations of large language models in clinical decision-making

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.538983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.176259Z digest=sha256:92fd893582080b5fd13409163d2c10d89d1b483e2182ce50a76e94c87dbeca70

Observation 0fcbfc62-bd5c-45bf-aff3-08c2a3a698a8 · outbound

This paper cites Qwen2.5-Coder Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Qwen2.5-Coder Technical Report

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.178782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.178782Z digest=sha256:4c8d4bc725df16e1681ff8faab1f9e2a65f9fd4a063a890c5e1673ca81c5e8bc

Observation f86ec334-4259-4823-86f7-b98d40e9386d · outbound

This paper cites What disease does this patient have? a large-scale open domain question answering dataset from medical exams.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving What disease does this patient have? a large-scale open domain question answering dataset from medical exams

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.531079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.181345Z digest=sha256:9f8d7d302f8b883336f652826ef4d8f60c23e829bb70527dc004379a61c588e9

Observation 4c0cd703-4887-4223-9258-4fd8c83a4fd0 · outbound

This paper cites P ub M ed QA : A dataset for biomedical research question answering.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving P ub M ed QA : A dataset for biomedical research question answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.184334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.184334Z digest=sha256:69fd0b0511a63e7607f7d0db70fd006be680bb9f029c75772422e9940e70b24b

Observation eaca3cc6-1e9f-43ff-a53f-6ab39b6068f4 · outbound

This paper cites E., Bulgarelli, L., Shen, L., Gayles, A., Shammout, A., Horng, S., Pollard, T.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving E., Bulgarelli, L., Shen, L., Gayles, A., Shammout, A., Horng, S., Pollard, T

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.522628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.187420Z digest=sha256:3febed6be95cca43e1a9aec48c06a8634d158062176ebb577d579f357a3af174

Observation b9c83da7-d8bd-4a6c-b933-c0c64d238864 · outbound

This paper cites A., Roberts, A., et al.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving A., Roberts, A., et al

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.514149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.190317Z digest=sha256:068928cbcae3d398bff7cb25c36c271f00e4650636d3008c2c405610021c2501

Observation c46d6e30-cdb6-46b1-82cf-70978be20cd9 · outbound

This paper cites DeepSeek-V3 Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving DeepSeek-V3 Technical Report

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.193159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.193159Z digest=sha256:8d937d845694050b53a16da1f26db270800b6b058a049a96b0516cf38c4c81e6

Observation 97da048d-0f1c-41db-bd60-efe06806f59d · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Gemma: Open Models Based on Gemini Research and Technology

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.196012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.196012Z digest=sha256:097c7fead7b075b4fcf66191f9b91775da053db316d7dd3a704cd05f41bd205c

Observation f6b9cd02-c8d5-4bcd-a317-cb386966799b · outbound

This paper cites Capabilities of GPT-4 on Medical Challenge Problems.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Capabilities of GPT-4 on Medical Challenge Problems

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.198772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.198772Z digest=sha256:30a026ea384af450bdf1d4e546a1677853685a34c6dc789841ee8a84697bbe82

Observation 3cd2eaad-b15f-42f2-bf48-c9c9468ef675 · outbound

This paper cites Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.202015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.202015Z digest=sha256:604dfd69480593e1b484faa00a7403289dceee571f896fa2010538e0d7eb768e

Observation a5c9893b-15d8-47a0-bc03-e45d70f23d47 · outbound

This paper cites Gpt-4o mini: advancing cost-efficient intelligence, 2024.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Gpt-4o mini: advancing cost-efficient intelligence, 2024

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.506444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.205196Z digest=sha256:8b50ffbdd4dc640b90345e5534d8bb983b75a7c37290bf40f4c0c3d521906d22

Observation c3795670-7509-4275-a135-76acf6b41bc6 · outbound

This paper cites L., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., Schulman, J., Hilton, J., Kelton, F., Miller, L., Simens, M., Askell, A., Welinder, P., Christiano, P.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving L., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., Schulman, J., Hilton, J., Kelton, F., Miller, L., Simens, M., Askell, A., Welinder, P., Christiano, P

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.498276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.207683Z digest=sha256:6ea7f43e12cc92b8115176c616321f4ea84894f05cacf5f32eef657220f6f5bd

Observation bfc94c8e-6edd-48f9-8d71-382c70598620 · outbound

This paper cites Climedbench: A large-scale chinese benchmark for evaluating medical large language models in clinical scenarios.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Climedbench: A large-scale chinese benchmark for evaluating medical large language models in clinical scenarios

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.489846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.210803Z digest=sha256:9c31429251a5b43aeb2bb0cc90e01d96cc30cf5baa27d1b6de5c682b8c2d0348

Observation c8987b34-d0f1-4c58-9a8f-96e2a294843f · outbound

This paper cites K., and Sankarasubbu, M.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving K., and Sankarasubbu, M

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.480318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.213449Z digest=sha256:59a29642aa5ab10361f445992f8ca12b37d15816e34165621a2feaabb130ed78

Observation 248fa337-d7e6-4676-bf79-384226bffc17 · outbound

This paper cites Towards building multilingual language model for medicine.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Towards building multilingual language model for medicine

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.472154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.216322Z digest=sha256:3066caf4c1680896e49bebf4008f97a7040f15c0ad2f2411d9a02f25ac32dd28

Observation 5e131aa7-9c55-476c-9375-e08872162d29 · outbound

This paper cites S., Wei, J., Chung, H.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving S., Wei, J., Chung, H

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.463569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.218853Z digest=sha256:8a3e5f608dddae01b1f93421217373527e8b927d22276ed24cd9fd14b41ec6b0

Observation af6a1063-7661-448a-adea-d5ba1c302f63 · outbound

This paper cites Towards Expert-Level Medical Question Answering with Large Language Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Towards Expert-Level Medical Question Answering with Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.221544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.221544Z digest=sha256:05c8ccbf7de775cb86078a92fa9f474c61d0fc6d30195d42e997a8b6090671cf

Observation 3ffb2cb6-3710-49b2-8cef-d42da69b3dad · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Gemma 2: Improving Open Language Models at a Practical Size

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.224344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.224344Z digest=sha256:dfcb0498a69bb04f0f33cc934ddcfb04a5b40d8db9af204ce407e4c45f62d2b7

Observation 8336a65d-d5a4-4049-9261-69fb63164818 · outbound

This paper cites Clinical Camel: An Open Expert-Level Medical Language Model with Dialogue-Based Knowledge Encoding.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Clinical Camel: An Open Expert-Level Medical Language Model with Dialogue-Based Knowledge Encoding

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.227231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.227231Z digest=sha256:e527ce0823e5a83b1df35d76d537c377f9c3975dadafee31443fcf2319d608cc

Observation 0acfdcce-ad1c-497e-8b42-1212f8ceea78 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving LLaMA: Open and Efficient Foundation Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.231310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.231310Z digest=sha256:ae2468a831d57ec352a29421d25da4799e87c2386878f845746a823124cf5027

Observation d8504776-e4c1-4978-96c2-9dc051e9a5e8 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.234201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.234201Z digest=sha256:784f9c076395d6f4d0ebf3b852a0cbb3003544994532aac52e7f9818506cfcaa

Observation 17dcdb23-6cf0-47ba-9eb7-35b378b4663f · outbound

This paper cites CMB : A comprehensive medical benchmark in C hinese.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving CMB : A comprehensive medical benchmark in C hinese

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.454966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.236963Z digest=sha256:1179785e40c0fad844dce7f2a53d020e0d1548bb437639da60660821db625cb9

Observation ef377819-65cf-4f80-9436-19ae876ac2df · outbound

This paper cites C., Wu, J., and Liu, X.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving C., Wu, J., and Liu, X

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.445187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.239860Z digest=sha256:3275ebba98fcb5a17c2f6a5f7d5fb36a64c2b7448437b94342aa75d748545020

Observation 02be86c3-a349-4375-a67d-3d21bc6e14cf · outbound

This paper cites Qwen2 Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Qwen2 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.242615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.242615Z digest=sha256:4413e225444b219164e8c7de7ffff061d9f7289e6a197a61555d17ebd7288f51

Observation aa18151f-2968-4c77-9337-cd03dc8dcd71 · outbound

This paper cites write newline.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving write newline

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.245710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.245710Z digest=sha256:fa68176ad783d49a16b7950a3bc09e3599fbf402c62dad0a18752e726e4bdce2

Pith citing papers

No inbound Pith citation observations are available.