Pith. sign in

Paper Citation Record · LEDGER

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving

As of 13 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2506.08349.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08349 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:19:33.245710Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact1
  • verified fuzzy16
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a78a2827-c586-419e-b5d6-48029857868f · outbound

This paper cites Phi-4 Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Phi-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.126829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.126829Z digest=sha256:5f1299975e89f58ceff0c926d0aceb43c3481fc007f3fdc8fd8160d3132e01f2

Observation baaad06f-1a16-4614-83fb-429004cf21ea · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.130742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.130742Z digest=sha256:77808d9d9c2dd51f8af410261ca935065cfb072d61f670d9a31535214cfbfe0e

Observation e6fd8e85-c75c-4a90-b8dd-048bf80acbe6 · outbound

This paper cites GPT-4 Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.134332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.134332Z digest=sha256:64004642d102a48208a27520b5cfe4812a885dea04e155fffff0fc8925730ebe

Observation e1edc71c-ca9c-4b06-9973-ca92efbb74a8 · outbound

This paper cites an unresolved cited work.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:19:33.594125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.138904Z digest=sha256:06debbe3c93a390cf21e2232eb0570ae7cf0c87be065f037aaaf7a5c1ad3d676

Observation 73a37a47-5b85-4ec6-8b57-89151924ac1a · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Gemini: A Family of Highly Capable Multimodal Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.142732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.142732Z digest=sha256:14e14e4664ee776947fff334510605bfd302c6408de334fb808eb756f475c017

Observation b8962460-be4d-4715-a13a-e32a85b439d3 · outbound

This paper cites Qwen Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Qwen Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.146797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.146797Z digest=sha256:85d989200e0320906b988ff6a1582a1cfc7a4429db27ddbf0f1bbbbfabaf9118

Observation a1e12a8d-ce13-4f17-b6a5-19d07d4e3f56 · outbound

This paper cites Overview of the medical question answering task at trec 2017 liveqa.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Overview of the medical question answering task at trec 2017 liveqa

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.586684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.150105Z digest=sha256:a8227ecac50f34be8aac33d4bb9d9dd5e65d85a8cc48eaa8e85edf438e67019d

Observation e4be78c0-33bf-47be-a4e8-95f3a2c3a249 · outbound

This paper cites S., Englehart, M.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving S., Englehart, M

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.578962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.153164Z digest=sha256:3e528424417abe86e521486ca76da8598a8ff72f137bbddcadf38d17045a9924

Observation 5d9070cc-4f18-475a-a428-33118c443882 · outbound

This paper cites The unified medical language system (umls): integrating biomedical terminology.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving The unified medical language system (umls): integrating biomedical terminology

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.571091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.155904Z digest=sha256:672d06e9101ab89a8dd4bef181e7f007116a7203266c82404911535f20a2e693

Observation cbc39d3a-86ce-4bad-803a-46ab91b44a48 · outbound

This paper cites an unresolved cited work.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:19:33.562957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.158633Z digest=sha256:03b2fda2f59f3b1a38417fd660497995b882c9d21c429ea783847f63e42c6296

Observation 72be59a9-6e3a-4dc2-bef3-47e54e26ec34 · outbound

This paper cites Medbench: A large-scale chinese benchmark for evaluating medical large language models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Medbench: A large-scale chinese benchmark for evaluating medical large language models

Reference 11

Resolution
verified exact
doi, observed 2026-08-07T05:19:33.554631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.161403Z digest=sha256:e2224b95daa449948e7b250e4eb0c1439b610210a91adf1896a40b75994f9787

Observation 7b449a13-136e-4b1c-8e37-a8137733ce75 · outbound

This paper cites MEDITRON-70B: Scaling Medical Pretraining for Large Language Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving MEDITRON-70B: Scaling Medical Pretraining for Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.164270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.164270Z digest=sha256:e983f16609f26dea46fd775a3a9e2bad3ced81fdceddb3bf512f67c83ef1a677

Observation 8d05d94c-dc3b-458f-b289-ff7c5e008b1b · outbound

This paper cites U., Pimentel, M.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving U., Pimentel, M

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.546163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.167785Z digest=sha256:49343581219206bed17ab76d55bb8e9462f146e9d8e0c4909cd07549f5cc0b98

Observation 9a8cf196-9e57-404c-9580-351919d5c879 · outbound

This paper cites Med42-v2: A Suite of Clinical LLMs.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Med42-v2: A Suite of Clinical LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.170913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.170913Z digest=sha256:29d788e261d12125a20008941cd29659e5aac616e14ac9a4981c30991d3f0979

Observation 2f11382f-c86c-4fa6-a23b-6ceab1eb99e8 · outbound

This paper cites The Llama 3 Herd of Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving The Llama 3 Herd of Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.173687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.173687Z digest=sha256:d1b6c6ff2cba11038b7c668ba98e43f805f4de37118d628c779de9908cd70a13

Observation 105f22a7-a464-4273-91fb-2c2102d96403 · outbound

This paper cites Evaluation and mitigation of the limitations of large language models in clinical decision-making.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Evaluation and mitigation of the limitations of large language models in clinical decision-making

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.538983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.176259Z digest=sha256:d3b4a762cfaa9548cbe52210dc1b31bbe2d992fac67c9c70c11c32b6274d69be

Observation 0fcbfc62-bd5c-45bf-aff3-08c2a3a698a8 · outbound

This paper cites Qwen2.5-Coder Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Qwen2.5-Coder Technical Report

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.178782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.178782Z digest=sha256:4c8d4bc725df16e1681ff8faab1f9e2a65f9fd4a063a890c5e1673ca81c5e8bc

Observation f86ec334-4259-4823-86f7-b98d40e9386d · outbound

This paper cites What disease does this patient have? a large-scale open domain question answering dataset from medical exams.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving What disease does this patient have? a large-scale open domain question answering dataset from medical exams

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.531079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.181345Z digest=sha256:073f995516e3c0b44495c77e0629b1deef529ef44e7781645f88794962c703d6

Observation 4c0cd703-4887-4223-9258-4fd8c83a4fd0 · outbound

This paper cites P ub M ed QA : A dataset for biomedical research question answering.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving P ub M ed QA : A dataset for biomedical research question answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.184334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.184334Z digest=sha256:69fd0b0511a63e7607f7d0db70fd006be680bb9f029c75772422e9940e70b24b

Observation eaca3cc6-1e9f-43ff-a53f-6ab39b6068f4 · outbound

This paper cites E., Bulgarelli, L., Shen, L., Gayles, A., Shammout, A., Horng, S., Pollard, T.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving E., Bulgarelli, L., Shen, L., Gayles, A., Shammout, A., Horng, S., Pollard, T

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.522628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.187420Z digest=sha256:f166620cd7e47b01bb45c5cc8d5d8ff2d2fb0a2bc84275ded52bee3d17c6ef87

Observation b9c83da7-d8bd-4a6c-b933-c0c64d238864 · outbound

This paper cites A., Roberts, A., et al.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving A., Roberts, A., et al

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.514149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.190317Z digest=sha256:6d796b49deccaf3f95d959968b1969036d49a4679b9437befe0359635bcd0c91

Observation c46d6e30-cdb6-46b1-82cf-70978be20cd9 · outbound

This paper cites DeepSeek-V3 Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving DeepSeek-V3 Technical Report

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.193159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.193159Z digest=sha256:b0c9fd97823c7da56ca6b65efd1e132c9e84d73694fc1cf78e228e3091a05f0c

Observation 97da048d-0f1c-41db-bd60-efe06806f59d · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Gemma: Open Models Based on Gemini Research and Technology

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.196012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.196012Z digest=sha256:097c7fead7b075b4fcf66191f9b91775da053db316d7dd3a704cd05f41bd205c

Observation f6b9cd02-c8d5-4bcd-a317-cb386966799b · outbound

This paper cites Capabilities of GPT-4 on Medical Challenge Problems.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Capabilities of GPT-4 on Medical Challenge Problems

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.198772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.198772Z digest=sha256:d120ed4bd7ef962a6495bb67b3a610cc61f68c0ad2accb87c6bb7827943df364

Observation 3cd2eaad-b15f-42f2-bf48-c9c9468ef675 · outbound

This paper cites Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.202015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.202015Z digest=sha256:96042ac3d0c6ff3fabd7267c0fff61de66d9bf72a93e6a7bf509241be4fbea10

Observation a5c9893b-15d8-47a0-bc03-e45d70f23d47 · outbound

This paper cites Gpt-4o mini: advancing cost-efficient intelligence, 2024.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Gpt-4o mini: advancing cost-efficient intelligence, 2024

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.506444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.205196Z digest=sha256:706de39cf38dc97f73dae0c37605a98cb49840468276803d21eb950d8f3e5d8e

Observation c3795670-7509-4275-a135-76acf6b41bc6 · outbound

This paper cites L., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., Schulman, J., Hilton, J., Kelton, F., Miller, L., Simens, M., Askell, A., Welinder, P., Christiano, P.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving L., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., Schulman, J., Hilton, J., Kelton, F., Miller, L., Simens, M., Askell, A., Welinder, P., Christiano, P

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.498276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.207683Z digest=sha256:9079e2850c9010680a5abda8e50ebb379007c3089b11311c61e8726c1adbdf11

Observation bfc94c8e-6edd-48f9-8d71-382c70598620 · outbound

This paper cites Climedbench: A large-scale chinese benchmark for evaluating medical large language models in clinical scenarios.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Climedbench: A large-scale chinese benchmark for evaluating medical large language models in clinical scenarios

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.489846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.210803Z digest=sha256:62c848c68ec658b57f347a00f48104602b114bf19d295169c7466d593d4c73f7

Observation c8987b34-d0f1-4c58-9a8f-96e2a294843f · outbound

This paper cites K., and Sankarasubbu, M.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving K., and Sankarasubbu, M

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.480318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.213449Z digest=sha256:9bbc7c1fc86f474c348f7ff4430e3a6b6376dbaea5fde4a77273cbb54e7efe27

Observation 248fa337-d7e6-4676-bf79-384226bffc17 · outbound

This paper cites Towards building multilingual language model for medicine.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Towards building multilingual language model for medicine

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.472154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.216322Z digest=sha256:b5a509eb7482392c273844913aedae71b1d4d97424a769ba2426633fc745f3c9

Observation 5e131aa7-9c55-476c-9375-e08872162d29 · outbound

This paper cites S., Wei, J., Chung, H.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving S., Wei, J., Chung, H

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.463569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.218853Z digest=sha256:9f56ff914d9bf8d5e58b6f1bed6c7db63a1b78cda8786e5675ed940cd9eca8ce

Observation af6a1063-7661-448a-adea-d5ba1c302f63 · outbound

This paper cites Towards Expert-Level Medical Question Answering with Large Language Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Towards Expert-Level Medical Question Answering with Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.221544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.221544Z digest=sha256:a3a9442f18fb92f10a2d74c115b4f270fa35065a7ca38632dfd77055feacfa1f

Observation 3ffb2cb6-3710-49b2-8cef-d42da69b3dad · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Gemma 2: Improving Open Language Models at a Practical Size

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.224344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.224344Z digest=sha256:dfcb0498a69bb04f0f33cc934ddcfb04a5b40d8db9af204ce407e4c45f62d2b7

Observation 8336a65d-d5a4-4049-9261-69fb63164818 · outbound

This paper cites Clinical Camel: An Open Expert-Level Medical Language Model with Dialogue-Based Knowledge Encoding.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Clinical Camel: An Open Expert-Level Medical Language Model with Dialogue-Based Knowledge Encoding

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.227231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.227231Z digest=sha256:cd094b0f46a1c79e1bc9ca6118c40745067ea92d9272f3eb7244fa06988b84bc

Observation 0acfdcce-ad1c-497e-8b42-1212f8ceea78 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving LLaMA: Open and Efficient Foundation Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.231310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.231310Z digest=sha256:ae2468a831d57ec352a29421d25da4799e87c2386878f845746a823124cf5027

Observation d8504776-e4c1-4978-96c2-9dc051e9a5e8 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.234201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.234201Z digest=sha256:784f9c076395d6f4d0ebf3b852a0cbb3003544994532aac52e7f9818506cfcaa

Observation 17dcdb23-6cf0-47ba-9eb7-35b378b4663f · outbound

This paper cites CMB : A comprehensive medical benchmark in C hinese.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving CMB : A comprehensive medical benchmark in C hinese

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.454966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.236963Z digest=sha256:17a919645cef9d39fff586e960d2e4801b2181132802ae3165ebfe308fa5f007

Observation ef377819-65cf-4f80-9436-19ae876ac2df · outbound

This paper cites C., Wu, J., and Liu, X.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving C., Wu, J., and Liu, X

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.445187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.239860Z digest=sha256:25b995b68aead06d1234e3c660de7bde4a3cd0d1663601dd7569a833f3623210

Observation 02be86c3-a349-4375-a67d-3d21bc6e14cf · outbound

This paper cites Qwen2 Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Qwen2 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.242615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.242615Z digest=sha256:64bd3a520660a0e6447bccbe16c99e4d50ae442d68f5aff805f7a2206f18adaf

Observation aa18151f-2968-4c77-9337-cd03dc8dcd71 · outbound

This paper cites write newline.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving write newline

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.245710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.245710Z digest=sha256:fa68176ad783d49a16b7950a3bc09e3599fbf402c62dad0a18752e726e4bdce2

Pith citing papers

No inbound Pith citation observations are available.