Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:16:49.312244Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 0 inbound Pith citation observations for arXiv:2509.07135.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:16:49.312244Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
71 of 71 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e65df800-a62f-4fd7-bcf7-10705eceaee8 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Language Models are Few-Shot Learners
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e6010d7-479f-4ed2-b91b-2236d9014be6 · outbound
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a2e3a83b-9b98-486e-82e4-58a45ea40cc5 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Baidoo-Anu, L
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 326e9dee-dbb8-4493-b0df-d5d650d6b192 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 204fa921-5871-42bf-9b11-5d7a813520c1 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations SuperGLUE: A Stickier Benchmark for General-Purpose Language Understanding Systems
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6667e3b0-de83-4fce-af9c-15949911fea6 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Hendrycks, C
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c68c3aef-8b7e-4a85-bb0e-13549259717c · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Attanasio, P
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 419d4ade-231c-4b2d-a00d-bbd7f31016b2 · outbound
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 25e77af3-452b-4438-8f56-6fc147781370 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Attanasio, P
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b2f3cdf6-4b39-4373-b71a-8c7e00d947f8 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9001183f-ac9c-4fb2-8b4a-25112fd11c0a · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91c3b74b-c09a-46f5-a217-c44e7966071b · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations PubMedQA: A Dataset for Biomedical Research Question Answering
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f42af82-5b87-4985-ade6-f85c4f43ce9b · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ace3c07e-32a7-4e8d-b114-26ce30898445 · outbound
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 23bc6808-d427-46ac-8787-18ece3a7d9bb · outbound
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e89ac205-2c16-4b6f-913b-93f4e1898fb8 · outbound
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a6800399-d681-4e64-b362-6d9055d06838 · outbound
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation af934513-ee0b-4917-ba29-535199513247 · outbound
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8a896b5a-2d01-4cb1-b795-1bfb3a134c95 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Calibrate Before Use: Improving Few-Shot Performance of Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69b2e6f1-05b6-48eb-aa13-7ce947d1ea26 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Wei, et al., Chain-of-Thought Prompting Elicits Reasoning in Large Language Models,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d4e9987e-4e7b-403d-985c-f52aff34382e · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fbd38dc-624b-45bf-af5b-2ec106e795f7 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Yang, et al., Qwen2.5 technical report,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 54b2dff8-c75a-4ea7-a8b9-d40a4a847185 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Gemma: Open Models Based on Gemini Research and Technology
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a406451c-01f9-401a-a742-9e474aca0834 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Grattafiori, et al., The Llama 3 Herd of Models,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6020059f-a93d-45d9-a91f-736fc3a6a39d · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3bfd345-919a-43ef-a7bf-b711dfee23f8 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 70bba81e-5eb4-47ca-b173-59df023a23f9 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Groeneveld, et al., OLMo: Accelerating the sci- ence of language models, in: L.-W
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edd361d1-4dd9-4873-b94c-0b59175953ab · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Aya Model: An Instruction Finetuned Open-Access Multilingual Language Model
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b23dd12-0a49-42dc-802d-2af027aeef81 · outbound
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0b6a3a29-9f68-4616-a4b5-988c756e2d2a · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Tutti i bambini amano il gelato
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ff3760e2-5407-4bdb-9036-30622c17c823 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e278186-accb-4ab0-aded-4346cd79eeba · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations All children love ice cream
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e807ec9b-b50e-44e1-bb32-b151c542ae4e · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Logic Example Domanda:Se e solo se Giulia a luglio non va in vacanza in montagna, va poi in vacanza al mare ad agosto
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c801bf5b-9fc0-44a7-98a6-1562e4648bb8 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Carolina ha acquistato molte borse, dunque ha speso molti soldi
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d916dd1e-97df-414f-9e65-696f04cd43ef · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Stasera non ha piovuto, dunque è andata in motorino
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e193d783-787d-499f-88d0-07bbdd8f1c22 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Ha già man- giato albicocche a pranzo, dunque a cena non mangia le fragole
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 02ddee4a-1adf-4f59-bd04-7e046c8d81ac · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Clara ha superato gli esami, dunque ha studi- ato molto
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 89998033-3302-4271-87da-d6ce58e0dd1b · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 788372c4-1cc8-4420-8ddf-7e92d3bf499e · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Carolina bought many bags, therefore she spent a lot of money
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 79086ea3-cc52-4a2c-9711-0ce53e29b796 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Tonight it did not rain, there- fore she went on her scooter
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c37903f5-366b-4eb0-a422-0eeab25fa248 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations She already ate apricots for lunch, therefore she does not eat strawberries for dinner
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 01fc384f-0306-4105-b761-39bc7b6b7202 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Clara passed the exams, therefore she studied hard
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b955e2fb-3393-4b54-9464-e3e0decb00da · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Riccardo does not play ten- nis, therefore he did not play football (Correct Answer: 3) A.3
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation dc5abd4a-3dff-498b-8d32-e9f70ca1bce1 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d5b6dd9d-a3bc-4621-86ae-0004792b549b · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 06ee3f3e-d069-4cb8-9fb5-17392069896f · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e3ea7f55-0f8b-4746-b42e-23c6ee37c60e · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b12fa6f6-4463-4830-a9e2-31ee8344aa56 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 4d6b9822-df21-4664-aa4d-345e56eca906 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 52b337fb-7527-463e-9865-aeb06616114f · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0e9fc822-787c-4c40-a52d-759ec30c0132 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 1e68f2cf-7f91-4d6b-a60a-f724970bb7b3 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Chemistry Example Domanda:A quante moli corrispondono 5 mL (d=1,8 g ·cm−3) di un composto avente una massa molare di 450 g·mol−1? Possibili risposte:
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation dfa12a9b-b86c-4518-af17-7203273c4e4b · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f7dd84eb-8475-456b-97b8-c15416b26e80 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 4465b9ea-ee69-4b40-8315-7483de6bb1ed · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 4aa63aa9-1018-4f60-8830-53869bab582b · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d527f1af-176b-460f-a817-8d8c7490ab44 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 121ac5b2-ba64-47a5-84bc-0df35dfabf48 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Mathematics Example Domanda:Dati tre segmenti AA’, BB’ e CC’ tali che: AA’ = 2 cm, BB’ = 1,5 * AA’, CC’ = 2,0 * BB’
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation beb19267-748c-4cab-99ec-03486af1a569 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 98692b07-e034-4301-b79a-aca303a168cf · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 4d423c89-8307-4191-b0c9-6ecccd559b36 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0e97a9d2-5a33-4ca5-8f7f-f38a82323d18 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a02d19f7-f5bd-43fb-8159-821f7c13b79a · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Which triangle is possible to construct with these sides? Possible answers:
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2acda169-967c-40c5-b247-e08c418936e8 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 61a9a47e-e0f4-481b-8d34-bdc0bb827362 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation bd18b8fd-9f9f-4327-beaa-2e163ca0c557 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Unresolved cited work
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b8559297-e408-42e4-be3c-c5cf45afb821 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Per-Subject Model Performance Table 4 Per-subject accuracy (%) on MedBench-IT for Standard (Std.) and Reasoning (Reas.) prompts
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 612f47fa-c3d0-476a-80cc-d7b96b97530e · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Measuring Massive Multitask Language Understanding
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36afa985-006f-4174-a06b-ab01ab22e8ac · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b483a8fa-bc0c-4e15-bb33-0fc1cc6dc9c4 · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations The Llama 3 Herd of Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb6dd17c-2406-4603-9855-6aa9ff542f6c · outbound
MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations Qwen2.5 Technical Report
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.