Pith. sign in

Paper Citation Record · LEDGER

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving

As of 7 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2506.08349.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08349 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:19:33.245710Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact1
  • verified fuzzy16
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a78a2827-c586-419e-b5d6-48029857868f · outbound

This paper cites Phi-4 Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Phi-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.126829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.126829Z digest=sha256:6171f1c23395418cc4ec9e07d1a78cdcb18dded62d909bbc027d808eb2ccc8cd

Observation baaad06f-1a16-4614-83fb-429004cf21ea · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.130742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.130742Z digest=sha256:f9eda08a5a87356f8857bf9cf0e589cd6e07ed5488b3431989f68dcce20aae0a

Observation e6fd8e85-c75c-4a90-b8dd-048bf80acbe6 · outbound

This paper cites GPT-4 Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.134332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.134332Z digest=sha256:65f5cfff5da985d7f7e41680304572f5272257cbe8e8f2620eee89ead75589b5

Observation e1edc71c-ca9c-4b06-9973-ca92efbb74a8 · outbound

This paper cites an unresolved cited work.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:19:33.594125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.138904Z digest=sha256:9d9b08157e4908b1ec5787138342b0f761dc4f833fd1f202a0b12604236bd6ef

Observation 73a37a47-5b85-4ec6-8b57-89151924ac1a · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Gemini: A Family of Highly Capable Multimodal Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.142732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.142732Z digest=sha256:f28ee50897f1abfa7351bbceef1e94c2c5752e5a2a95fda8c616520e5a374b65

Observation b8962460-be4d-4715-a13a-e32a85b439d3 · outbound

This paper cites Qwen Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Qwen Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.146797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.146797Z digest=sha256:5046a399f95f95a71573fbbf7f47a695dc3bb60b912324076b7cdb9fb8f27d9f

Observation a1e12a8d-ce13-4f17-b6a5-19d07d4e3f56 · outbound

This paper cites Overview of the medical question answering task at trec 2017 liveqa.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Overview of the medical question answering task at trec 2017 liveqa

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.586684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.150105Z digest=sha256:066e0e994d903fd443da74b287377c44e206c039663f44c65ba6c4bac97a8181

Observation e4be78c0-33bf-47be-a4e8-95f3a2c3a249 · outbound

This paper cites S., Englehart, M.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving S., Englehart, M

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.578962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.153164Z digest=sha256:e2fcbf88f57e5906367bffabd91fde867c2ce7ca01d7c282f40d3f8e99f9e26e

Observation 5d9070cc-4f18-475a-a428-33118c443882 · outbound

This paper cites The unified medical language system (umls): integrating biomedical terminology.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving The unified medical language system (umls): integrating biomedical terminology

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.571091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.155904Z digest=sha256:24240b9c3470681ddffc58b154e20e0913d8e5cf5bfb24198b3bac6bfc406bf0

Observation cbc39d3a-86ce-4bad-803a-46ab91b44a48 · outbound

This paper cites an unresolved cited work.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:19:33.562957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.158633Z digest=sha256:9ac5595d858738a981bf130fb5ecc73b8c29fafbd13585a14c92935d7436d10f

Observation 72be59a9-6e3a-4dc2-bef3-47e54e26ec34 · outbound

This paper cites Medbench: A large-scale chinese benchmark for evaluating medical large language models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Medbench: A large-scale chinese benchmark for evaluating medical large language models

Reference 11

Resolution
verified exact
doi, observed 2026-08-07T05:19:33.554631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.161403Z digest=sha256:2cdd6dc213e56fc706ae9d027ca9d2a692d63a033d55a025ee8501ecc812e37d

Observation 7b449a13-136e-4b1c-8e37-a8137733ce75 · outbound

This paper cites MEDITRON-70B: Scaling Medical Pretraining for Large Language Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving MEDITRON-70B: Scaling Medical Pretraining for Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.164270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.164270Z digest=sha256:d8043fc02b9284489ae9500a39a8dced577cc21007afbaf429dc6b1d07defb92

Observation 8d05d94c-dc3b-458f-b289-ff7c5e008b1b · outbound

This paper cites U., Pimentel, M.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving U., Pimentel, M

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.546163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.167785Z digest=sha256:8a7ef0ff5d3bbf0c0b44e0b5d63d0ed550b77212fde83c7c40058ebcf70d3578

Observation 9a8cf196-9e57-404c-9580-351919d5c879 · outbound

This paper cites Med42-v2: A Suite of Clinical LLMs.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Med42-v2: A Suite of Clinical LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.170913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.170913Z digest=sha256:2de5e7cfb178d94771ebb1606c7a18c0e9bb7d396175a5198f2352183a122c2b

Observation 2f11382f-c86c-4fa6-a23b-6ceab1eb99e8 · outbound

This paper cites The Llama 3 Herd of Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving The Llama 3 Herd of Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.173687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.173687Z digest=sha256:29feec62eeec2ee3b30f9331d22ab64f7be8ce5d1798e81bcdb819a0f9c36fc8

Observation 105f22a7-a464-4273-91fb-2c2102d96403 · outbound

This paper cites Evaluation and mitigation of the limitations of large language models in clinical decision-making.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Evaluation and mitigation of the limitations of large language models in clinical decision-making

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.538983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.176259Z digest=sha256:af53f4247d256884c0a73ac34e8e3a535b5c77011348672475ea482c2667b7cc

Observation 0fcbfc62-bd5c-45bf-aff3-08c2a3a698a8 · outbound

This paper cites Qwen2.5-Coder Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Qwen2.5-Coder Technical Report

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.178782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.178782Z digest=sha256:5cd0e6501c8170bac50d9e995429fd0136da33df22aee51e9c03d9371f673633

Observation f86ec334-4259-4823-86f7-b98d40e9386d · outbound

This paper cites What disease does this patient have? a large-scale open domain question answering dataset from medical exams.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving What disease does this patient have? a large-scale open domain question answering dataset from medical exams

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.531079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.181345Z digest=sha256:71d51294ac9169f795517b70d9e2b817d13c23f6bf6a30d495a0b7a7bfb94c30

Observation 4c0cd703-4887-4223-9258-4fd8c83a4fd0 · outbound

This paper cites P ub M ed QA : A dataset for biomedical research question answering.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving P ub M ed QA : A dataset for biomedical research question answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.184334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.184334Z digest=sha256:a4086bab3c2b7d0bd995f0d305cee2a00a3571f7253fa027101eabce8c460277

Observation eaca3cc6-1e9f-43ff-a53f-6ab39b6068f4 · outbound

This paper cites E., Bulgarelli, L., Shen, L., Gayles, A., Shammout, A., Horng, S., Pollard, T.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving E., Bulgarelli, L., Shen, L., Gayles, A., Shammout, A., Horng, S., Pollard, T

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.522628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.187420Z digest=sha256:5daf39be09fda7a624e73f21dc17979a826feacf32ae4bd9c83f91fbed1a8777

Observation b9c83da7-d8bd-4a6c-b933-c0c64d238864 · outbound

This paper cites A., Roberts, A., et al.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving A., Roberts, A., et al

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.514149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.190317Z digest=sha256:765cbdd9a9725ed10a1ca7544146d85a08d7e582e14efb898f7aef4de5053ce6

Observation c46d6e30-cdb6-46b1-82cf-70978be20cd9 · outbound

This paper cites DeepSeek-V3 Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving DeepSeek-V3 Technical Report

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.193159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.193159Z digest=sha256:87501c9273051122a045c322e840c2de4a335aa33a51468fbefc5bce91734432

Observation 97da048d-0f1c-41db-bd60-efe06806f59d · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Gemma: Open Models Based on Gemini Research and Technology

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.196012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.196012Z digest=sha256:32450ffb0aba00b8144cdd0f1a77b2ef1c8d26ed118d5812957bb08293f5ea6e

Observation f6b9cd02-c8d5-4bcd-a317-cb386966799b · outbound

This paper cites Capabilities of GPT-4 on Medical Challenge Problems.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Capabilities of GPT-4 on Medical Challenge Problems

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.198772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.198772Z digest=sha256:cfa25b76794ed7e8be87da3a3d8f9e527c994703ff3083199ca905ae31066431

Observation 3cd2eaad-b15f-42f2-bf48-c9c9468ef675 · outbound

This paper cites Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.202015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.202015Z digest=sha256:a5c8afa1c671c2120b6a60a19459c284efc22e89bb852d108f8ce69c842a35c8

Observation a5c9893b-15d8-47a0-bc03-e45d70f23d47 · outbound

This paper cites Gpt-4o mini: advancing cost-efficient intelligence, 2024.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Gpt-4o mini: advancing cost-efficient intelligence, 2024

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.506444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.205196Z digest=sha256:198d661c25bd6f383bdaead3fcc201c3b681a3ab540e4408bad5e8f1e5052a14

Observation c3795670-7509-4275-a135-76acf6b41bc6 · outbound

This paper cites L., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., Schulman, J., Hilton, J., Kelton, F., Miller, L., Simens, M., Askell, A., Welinder, P., Christiano, P.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving L., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., Schulman, J., Hilton, J., Kelton, F., Miller, L., Simens, M., Askell, A., Welinder, P., Christiano, P

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.498276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.207683Z digest=sha256:51dd4cb5a4d9695296fcc6ee86e0ab8603de9d4665ad8967d031a2f30d20c455

Observation bfc94c8e-6edd-48f9-8d71-382c70598620 · outbound

This paper cites Climedbench: A large-scale chinese benchmark for evaluating medical large language models in clinical scenarios.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Climedbench: A large-scale chinese benchmark for evaluating medical large language models in clinical scenarios

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.489846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.210803Z digest=sha256:98cf7650dad115ad182d62cccdbf82e0b80149fd750c4e0bea84c52ab506fc30

Observation c8987b34-d0f1-4c58-9a8f-96e2a294843f · outbound

This paper cites K., and Sankarasubbu, M.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving K., and Sankarasubbu, M

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.480318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.213449Z digest=sha256:faa40a2ab4b57ce34d0cd3b34c974876f51cb056484848609c803f6801b6712d

Observation 248fa337-d7e6-4676-bf79-384226bffc17 · outbound

This paper cites Towards building multilingual language model for medicine.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Towards building multilingual language model for medicine

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.472154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.216322Z digest=sha256:2da4ce41dfa9dfd0138958146ede9bd2daa5c6787298d2acc013a566e279690e

Observation 5e131aa7-9c55-476c-9375-e08872162d29 · outbound

This paper cites S., Wei, J., Chung, H.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving S., Wei, J., Chung, H

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.463569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.218853Z digest=sha256:801151c7555a82a0941d5aa773ca7187d96d38689212d822c51818cc120941dd

Observation af6a1063-7661-448a-adea-d5ba1c302f63 · outbound

This paper cites Towards Expert-Level Medical Question Answering with Large Language Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Towards Expert-Level Medical Question Answering with Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.221544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.221544Z digest=sha256:022804e4987bc05f2b12447810a54485b20a4036fe1822f4b3d4412c01f58fa8

Observation 3ffb2cb6-3710-49b2-8cef-d42da69b3dad · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Gemma 2: Improving Open Language Models at a Practical Size

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.224344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.224344Z digest=sha256:336631be6ed1393cdc0ef56313700971942afc39f43399bace3785cb0c6e0e43

Observation 8336a65d-d5a4-4049-9261-69fb63164818 · outbound

This paper cites Clinical Camel: An Open Expert-Level Medical Language Model with Dialogue-Based Knowledge Encoding.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Clinical Camel: An Open Expert-Level Medical Language Model with Dialogue-Based Knowledge Encoding

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.227231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.227231Z digest=sha256:4a1894c98223d0b40d8a5b547622b1bdcbf85d366f157f7860efd94641817eb3

Observation 0acfdcce-ad1c-497e-8b42-1212f8ceea78 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving LLaMA: Open and Efficient Foundation Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.231310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.231310Z digest=sha256:202d5b7a7453ef707b8c9c3be41c6d2474fa9e1ee3ead10a690c17a66d462e4e

Observation d8504776-e4c1-4978-96c2-9dc051e9a5e8 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.234201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.234201Z digest=sha256:63e13622881da67113f37d8f595ceaeb12092f60d0fb18a7121eb6ec1fd91a7d

Observation 17dcdb23-6cf0-47ba-9eb7-35b378b4663f · outbound

This paper cites CMB : A comprehensive medical benchmark in C hinese.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving CMB : A comprehensive medical benchmark in C hinese

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.454966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.236963Z digest=sha256:1d5d4f19cd537766bbfa65fea93727d578b72399d4398f206fdf31d8c86d7a46

Observation ef377819-65cf-4f80-9436-19ae876ac2df · outbound

This paper cites C., Wu, J., and Liu, X.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving C., Wu, J., and Liu, X

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:19:33.445187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:19:33.239860Z digest=sha256:8bf9f3ac14db094d848a3cee485bd8d560771bacd2967fac6d22492c20433315

Observation 02be86c3-a349-4375-a67d-3d21bc6e14cf · outbound

This paper cites Qwen2 Technical Report.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Qwen2 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.242615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.242615Z digest=sha256:e49a02ad559d45675da79e5d8aa8ee09677cf3a45b123844116bd04f7a8d38df

Observation aa18151f-2968-4c77-9337-cd03dc8dcd71 · outbound

This paper cites write newline.

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving write newline

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:33.245710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:33.245710Z digest=sha256:e1b478d16697c7df04bbc9cc8b57edfac542ef7b45c0fd68cee22d1ec96b1df3

Pith citing papers

No inbound Pith citation observations are available.