Pith. sign in

Paper Citation Record · LEDGER

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark

As of 8 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 1 inbound Pith citation observation for arXiv:2506.00250.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00250 v4

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:13:33.624540Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T05:15:29.381495Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f5533fe7-f322-436d-8ea3-9668fb19fc61 · outbound

This paper cites Expert Evaluation: The model follows a Western protocol; how- ever, local clinical practice requires urgent biopsy due to high mor- tality risk.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Expert Evaluation: The model follows a Western protocol; how- ever, local clinical practice requires urgent biopsy due to high mor- tality risk

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:37.276094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:31.959523Z digest=sha256:ee43ec8eeb52c344ada2525105891f9392b5ff43fbf2f1bf05edb5607d62cc82

Observation d1b784b7-d21b-4b88-9533-7235fb2a6ea5 · outbound

This paper cites Expert Evaluation: The model selected a technically true but con- textually incorrect answer; expert notes ambiguity in phrasing and clinical intent.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Expert Evaluation: The model selected a technically true but con- textually incorrect answer; expert notes ambiguity in phrasing and clinical intent

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:37.055330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:32.056332Z digest=sha256:2475253c8fd676621035bb67c912b900addb4bd59ef42e4ac4656a53145010ce

Observation 87fca736-cd33-4b5b-aa1e-00d915312cd3 · outbound

This paper cites Expert Evaluation: The patient’s immunosuppression requires a different clinical approach, which the model failed to identify.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Expert Evaluation: The patient’s immunosuppression requires a different clinical approach, which the model failed to identify

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:36.843318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:32.168595Z digest=sha256:10c7c8cc33a1b5e7967830c4dbce6e8110e1b9aaeb162d4360ffee01cb9d52b9

Observation 79f3da1d-db18-4fdd-b1b9-06ab12612b71 · outbound

This paper cites ACM T ransactions on Computing for Healthcare, 3(1):1–23.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark ACM T ransactions on Computing for Healthcare, 3(1):1–23

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:38.508607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:31.354430Z digest=sha256:f89476488ff095544cace11e04f4dc72d243dd2c7453fda7c246561969cf3224

Observation 5d4f9807-8edd-46c2-92b5-05b617dcb905 · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:34.901288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:33.071965Z digest=sha256:2bbca5344a71c558611a4b941b5cda13c7c98be0ccf1464abb0c66b1b3de2d1c

Observation 11ae05f1-b7f8-4dca-aeb7-b7537db77fd5 · outbound

This paper cites CoT": "your step-by-step reasoning.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark CoT": "your step-by-step reasoning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:34.683065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:33.190736Z digest=sha256:c2436c9f8f194f5be18be6ca20bcf9b062b6fc0cf4cc577d3082f7564646b10a

Observation d09d3fb0-ba27-4ca5-b5a0-9ed9ae29be00 · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:38.016159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:31.582306Z digest=sha256:b394aff2850bb0ea2ab951c0c4a5c5bdca8cc9b4daaa28b6e9ebaf89858584fd

Observation 333772bc-b118-45bd-8463-5695d3ee4b87 · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:37.773092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:31.690237Z digest=sha256:f4072da3e945979f58c91d4305fb98484f60c60bf3b03096c31d68f027f87ea1

Observation 35c8c183-e759-4179-b9c4-3b6030641113 · outbound

This paper cites For each example, we highlight the clinical context, the correct answer, the model’s response, and a summary of the ex- pert’s evaluation.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark For each example, we highlight the clinical context, the correct answer, the model’s response, and a summary of the ex- pert’s evaluation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:37.523363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:31.818911Z digest=sha256:3d851f00f8bfef3cadd0a7d18e06e0dea9de4488d1bbec40eb6c2401b4573f63

Observation a570c20d-c27b-4821-bb8a-7f76be691e63 · outbound

This paper cites Few-shot Prompt You are a medical expert tasked with answering multiple-choice medical questions.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Few-shot Prompt You are a medical expert tasked with answering multiple-choice medical questions

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:36.331963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:32.435042Z digest=sha256:71516a0767a64bbce9a69387d7d2502493e8c6d1e1d2ba67107eaa4fb8ba2f13

Observation 57cdafd6-93cf-4792-b6e7-2bf5e66e8475 · outbound

This paper cites Expert Evaluation: Model lacks pharmacologic mechanism knowledge and defaults to common treatments.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Expert Evaluation: Model lacks pharmacologic mechanism knowledge and defaults to common treatments

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:36.621196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:32.260041Z digest=sha256:cc4c48f76089ef1acac1bc684b0b744a93109535250b11a46b96e839bb4dbf2b

Observation c2835af1-6185-4b20-8f16-497a74057d29 · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:34.442456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:33.289303Z digest=sha256:f9c2fde02fed5b11a4e2115ba722eee83dce82c3d0df308cf345998422f70b7e

Observation 28642780-5928-4511-9493-a6c717ab7fa6 · outbound

This paper cites For each question, please:.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark For each question, please:

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:36.110162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:32.575356Z digest=sha256:85101f7b9f3171a1313cfda2e551bd7dcd3158ca33879e35e1f0a52530462ba2

Observation 0043c02f-c76f-4440-bcf6-073b3d66d4e8 · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:35.842329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:32.708500Z digest=sha256:f8377ac2bde60b33e1584de45d6412263cd294d77247a36ca7d624ef0d54f591

Observation d0ddb1d1-3820-45b4-a157-8362ef4713ff · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:35.579474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:32.811868Z digest=sha256:2885fb10469a3c3937fa3fa92d0f955c971cf6a62b66d94ee216e8ec530f1654

Observation 1071749f-28db-419a-b6f7-11e5e9326468 · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:35.316209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:32.892532Z digest=sha256:868bbaee9f1e5e9b77f27fe65c937019f139513e143c8a620d2fa1be8f167ab4

Observation 1232599a-57fc-41f8-b0bf-ed2ef52e343c · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:35.138642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:32.979952Z digest=sha256:acba08832b0fbeb024ec4cb485c053a61df713d861ed50a2d2f61d77f072ea5f

Observation a81ee56f-5fd3-4150-ad55-fd6a8bfa56bc · outbound

This paper cites T able 3: Distribution of extracted Persian medical terms.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark T able 3: Distribution of extracted Persian medical terms

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:34.230043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:33.406130Z digest=sha256:d403e4f9551f0a8614cfc530a4e3b8d96dfc0a02f74051b1796e65e831f7daca

Observation d7154dfd-fa04-47ae-adcb-39a2e3df5def · outbound

This paper cites Models are sorted by average performance.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Models are sorted by average performance

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:34.042820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:33.511059Z digest=sha256:c27883f092ee23f5d09206a266a37470ac39b21a225d02d8281921f0dc7604c3

Observation 38f7dc6f-680e-4545-a1d3-3f463d2dab56 · outbound

This paper cites Repre- sentative examples for each category are provided below.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Repre- sentative examples for each category are provided below

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:33.827945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:33.624540Z digest=sha256:5c8e94f70f326af5c2e3ee62c7a14cb5181feb9acceecdf2f4c034996c28052a

Observation d700b77d-4915-46d4-bdf8-7d80c04c6ed6 · outbound

This paper cites Josepha Campinha-Bacote.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Josepha Campinha-Bacote

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:38.999641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:31.162016Z digest=sha256:c2bb637b3d361f16e0cf0bea4c285c92f6dc160590223ddbe031ad52f68aea4d

Observation 7dbe64bb-bf4e-4fbe-ab9e-bdc0994a8f12 · outbound

This paper cites Neural Pro- cessing Letters, 53(6):3831–3847.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Neural Pro- cessing Letters, 53(6):3831–3847

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:38.758968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:31.254863Z digest=sha256:f445090112228a0cb7886120b8f16a70fcc973d725f76cebeb1f74e431caa345

Observation 86658fb0-1204-4fff-8aa0-faac04349f9f · outbound

This paper cites Laurence J.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Laurence J

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:38.265643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:31.496739Z digest=sha256:6773f8cdee08b24102becd196fba2c1885c9ae335408f2cbc0647d265ac34198

Observation e9e32653-c1af-4426-9573-5e20002f85c0 · outbound

This paper cites Artificial Intelligence in Medicine , 155:102938.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Artificial Intelligence in Medicine , 155:102938

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:39.204341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:13:31.088923Z digest=sha256:48a1b244d248e0b432cc34599c92a654d23e1aafbbd7f0764741d4580799b8eb

Pith citing papers

Observation 2ccf1708-174f-45a7-95f7-37afece052a3 · inbound

Biomedical Machine Translation for Low-Resource Arabic-Script Languages via Cross-Lingual Transfer and LoRA Adapter Merging cites this paper.

Biomedical Machine Translation for Low-Resource Arabic-Script Languages via Cross-Lingual Transfer and LoRA Adapter Merging PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T05:15:29.381495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T05:15:29.381495Z digest=sha256:1936b9138db365d2bc5616af637d21a467432df452495c4b485bafe5533fd18c