Pith. sign in

Paper Citation Record · LEDGER

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark

As of 20 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 1 inbound Pith citation observation for arXiv:2506.00250.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00250 v4

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:13:33.624540Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T05:15:29.381495Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f5533fe7-f322-436d-8ea3-9668fb19fc61 · outbound

This paper cites Expert Evaluation: The model follows a Western protocol; how- ever, local clinical practice requires urgent biopsy due to high mor- tality risk.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Expert Evaluation: The model follows a Western protocol; how- ever, local clinical practice requires urgent biopsy due to high mor- tality risk

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:37.276094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:31.959523Z digest=sha256:f88b467c490d67774c6bc0d7ca3197072063a5c3dee5f41d9cdc84e713585413

Observation d1b784b7-d21b-4b88-9533-7235fb2a6ea5 · outbound

This paper cites Expert Evaluation: The model selected a technically true but con- textually incorrect answer; expert notes ambiguity in phrasing and clinical intent.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Expert Evaluation: The model selected a technically true but con- textually incorrect answer; expert notes ambiguity in phrasing and clinical intent

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:37.055330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:32.056332Z digest=sha256:91ed71d0896058184479961c0970660c9ed71fa4c1304d20f70a1d914a20fe96

Observation 87fca736-cd33-4b5b-aa1e-00d915312cd3 · outbound

This paper cites Expert Evaluation: The patient’s immunosuppression requires a different clinical approach, which the model failed to identify.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Expert Evaluation: The patient’s immunosuppression requires a different clinical approach, which the model failed to identify

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:36.843318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:32.168595Z digest=sha256:f7d261aec2506aca9cde61d5c89e6e2bc27bddd0cae308a4e8b00cf759c837f9

Observation 79f3da1d-db18-4fdd-b1b9-06ab12612b71 · outbound

This paper cites ACM T ransactions on Computing for Healthcare, 3(1):1–23.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark ACM T ransactions on Computing for Healthcare, 3(1):1–23

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:38.508607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:31.354430Z digest=sha256:b4b61ac304013209df5e4f01212c80219405ff4fdc79d9abe2c4fdf0ad5e2e84

Observation 5d4f9807-8edd-46c2-92b5-05b617dcb905 · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:34.901288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:33.071965Z digest=sha256:370f380f911a3404637b4f959b68bd1169494deb10a4dd95ca805eaee7281c8e

Observation 11ae05f1-b7f8-4dca-aeb7-b7537db77fd5 · outbound

This paper cites CoT": "your step-by-step reasoning.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark CoT": "your step-by-step reasoning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:34.683065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:33.190736Z digest=sha256:1cd6849e2e4263fa62c5640c552259ac251c35ca302f9828736b2423edcfc467

Observation d09d3fb0-ba27-4ca5-b5a0-9ed9ae29be00 · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:38.016159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:31.582306Z digest=sha256:e1c42ed7c1b922577d9479b7f455c03c1d0eb912cc6e2f828f067c19c53cbfca

Observation 333772bc-b118-45bd-8463-5695d3ee4b87 · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:37.773092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:31.690237Z digest=sha256:cfc7c4d78e0c8852af76fc08a8622cfc147de67c52d0ea66688ae948d1a6f890

Observation 35c8c183-e759-4179-b9c4-3b6030641113 · outbound

This paper cites For each example, we highlight the clinical context, the correct answer, the model’s response, and a summary of the ex- pert’s evaluation.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark For each example, we highlight the clinical context, the correct answer, the model’s response, and a summary of the ex- pert’s evaluation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:37.523363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:31.818911Z digest=sha256:4f9c661f003d82d917ec1f6001db7685de9f55675d279b6eb733415c7c5317a9

Observation a570c20d-c27b-4821-bb8a-7f76be691e63 · outbound

This paper cites Few-shot Prompt You are a medical expert tasked with answering multiple-choice medical questions.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Few-shot Prompt You are a medical expert tasked with answering multiple-choice medical questions

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:36.331963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:32.435042Z digest=sha256:fcbb7f1416d109b79f75cc682ec50f5bc1bb53168d961c476f60e6bf485f8bf8

Observation 57cdafd6-93cf-4792-b6e7-2bf5e66e8475 · outbound

This paper cites Expert Evaluation: Model lacks pharmacologic mechanism knowledge and defaults to common treatments.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Expert Evaluation: Model lacks pharmacologic mechanism knowledge and defaults to common treatments

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:36.621196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:32.260041Z digest=sha256:e2758ea92b058dcd574b380469144384f50bb2dbaa81859ce96785cfd63158e3

Observation c2835af1-6185-4b20-8f16-497a74057d29 · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:34.442456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:33.289303Z digest=sha256:95f803739cfdb801b4e2ccb65825ae9c4c1e8f25ed62e71b358418ef60671115

Observation 28642780-5928-4511-9493-a6c717ab7fa6 · outbound

This paper cites For each question, please:.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark For each question, please:

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:36.110162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:32.575356Z digest=sha256:49b48c1a7116b38b0bc8c83956550930340fd2214499585c939e6942fb796081

Observation 0043c02f-c76f-4440-bcf6-073b3d66d4e8 · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:35.842329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:32.708500Z digest=sha256:bf24206b94e48e43b73780834a8928370857496b6163b93c85e40b7946120eab

Observation d0ddb1d1-3820-45b4-a157-8362ef4713ff · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:35.579474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:32.811868Z digest=sha256:9ec4dd8a2d4dc1ab2b095b54e22ff2ab70d5c57e905c5296dd3fac8740a4bc8e

Observation 1071749f-28db-419a-b6f7-11e5e9326468 · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:35.316209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:32.892532Z digest=sha256:666da60421fb5518e3eacc7ca56cb03289e862c7b18e4813912d9d92748874b0

Observation 1232599a-57fc-41f8-b0bf-ed2ef52e343c · outbound

This paper cites an unresolved cited work.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:35.138642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:32.979952Z digest=sha256:e0d6725f1c88da159c95357a8184e085fc842e7bb50a42650181075ad5c877d7

Observation a81ee56f-5fd3-4150-ad55-fd6a8bfa56bc · outbound

This paper cites T able 3: Distribution of extracted Persian medical terms.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark T able 3: Distribution of extracted Persian medical terms

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:34.230043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:33.406130Z digest=sha256:392efccedc03a9939f58d2e452b64cb61b10d29a4d4efb181eec48f0338eeccc

Observation d7154dfd-fa04-47ae-adcb-39a2e3df5def · outbound

This paper cites Models are sorted by average performance.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Models are sorted by average performance

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:34.042820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:33.511059Z digest=sha256:a32ef3902ff1b9b589c3ff591768cf7741f2d723e8506f44855248433d95f234

Observation 38f7dc6f-680e-4545-a1d3-3f463d2dab56 · outbound

This paper cites Repre- sentative examples for each category are provided below.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Repre- sentative examples for each category are provided below

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:33.827945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:33.624540Z digest=sha256:38673ccef64b9432cd3503031b2f65493af759aa5815c56afaf2a3cf7c900d9d

Observation d700b77d-4915-46d4-bdf8-7d80c04c6ed6 · outbound

This paper cites Josepha Campinha-Bacote.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Josepha Campinha-Bacote

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:38.999641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:31.162016Z digest=sha256:c90728578b02bb33a3339b75763714dbe7e6e3d66141641b4a8f5e1663be237c

Observation 7dbe64bb-bf4e-4fbe-ab9e-bdc0994a8f12 · outbound

This paper cites Neural Pro- cessing Letters, 53(6):3831–3847.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Neural Pro- cessing Letters, 53(6):3831–3847

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:38.758968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:31.254863Z digest=sha256:825ae2338dcd80e4e9953e4ffed0db77da494d5038f3806b5506ffb57f39eddf

Observation 86658fb0-1204-4fff-8aa0-faac04349f9f · outbound

This paper cites Laurence J.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Laurence J

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:38.265643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:31.496739Z digest=sha256:58aeb1c9e7578a3288572ecb64caae3863e096f162fd637852c945dd59686014

Observation e9e32653-c1af-4426-9573-5e20002f85c0 · outbound

This paper cites Artificial Intelligence in Medicine , 155:102938.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark Artificial Intelligence in Medicine , 155:102938

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:39.204341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:13:31.088923Z digest=sha256:ac322e3ecdd21c32be910dc59a78f9e3bfec30383b5283fdf17b002f88119399

Pith citing papers

Observation 2ccf1708-174f-45a7-95f7-37afece052a3 · inbound

Biomedical Machine Translation for Low-Resource Arabic-Script Languages via Cross-Lingual Transfer and LoRA Adapter Merging cites this paper.

Biomedical Machine Translation for Low-Resource Arabic-Script Languages via Cross-Lingual Transfer and LoRA Adapter Merging PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T05:15:29.381495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T05:15:29.381495Z digest=sha256:8a8a4291dfa3f3353dec8c8b2c085318be6c11577d0831e254e3b0be5e5715c5