Pith. sign in

Paper Citation Record · LEDGER

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations

As of 7 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2506.11114.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11114 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:39:42.506184Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 060c174e-c991-4b8e-87d0-2019120b2618 · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:45.354572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:40.369014Z digest=sha256:b42c968e09d830c20e60ecbc53dbaf561bf12e2348f30e22a83c3083d59db04d

Observation 27856649-cae8-4e56-911f-f832df18cab1 · outbound

This paper cites Computer software (2024),https://acrobat.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Computer software (2024),https://acrobat

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:45.340053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:40.467025Z digest=sha256:886355c513c2113666732dbb9bcf80204a69dc9d368bcb2fb4e147509293694f

Observation 96fda058-952c-40bf-8a73-30b661501932 · outbound

This paper cites Clinical Anatomy38(2), 186–199 (2025).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Clinical Anatomy38(2), 186–199 (2025)

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:45.323735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:40.638781Z digest=sha256:bd973ad7435fe0cb80c203aa6930d5906079f74c4d432da6bf5bfd1546925e09

Observation 3795bcb1-f27c-41a8-91e8-033cfa954105 · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:45.307733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:40.809202Z digest=sha256:8d2a610772300168dace615d14d6ae3cac7f94c5cde58148ad1132908c649c3e

Observation b5464ab6-122e-485f-b741-99726d5abbfe · outbound

This paper cites ishiyaku-dental.jp/, accessed: 2025-06-05.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations ishiyaku-dental.jp/, accessed: 2025-06-05

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:45.292741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:41.025291Z digest=sha256:b761a1465a2eb4e3321c5a9fa5483012e2a1a3c2dd91084ec09fcebc1b016567

Observation 464a702b-addd-466d-9fb8-246070c98a74 · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:45.088649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:41.204662Z digest=sha256:139f26b3f6f0d4a67ce4e0fd6cd1a75a023182a1b129ad427c6dac86979f095f

Observation 94858ee3-b033-4a99-ad26-05bfe244faa4 · outbound

This paper cites mynavi.jp/conts/kokushi_kouryaku/, accessed: 2025-06-05 KokushiMD-10 9.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations mynavi.jp/conts/kokushi_kouryaku/, accessed: 2025-06-05 KokushiMD-10 9

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:44.891693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:41.344767Z digest=sha256:1ca431ffc2882c796e01a13954750b64c4697912a7831ec07ac1543710ece448

Observation 9b2af865-a51f-4637-9c95-60be196d77d3 · outbound

This paper cites Applied Sciences11(14), 6421 (2021).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Applied Sciences11(14), 6421 (2021)

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:39:41.416728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:39:41.416728Z digest=sha256:e3e323dc168c489badef2a2a543d81131f28ab7550a8e26d831d1e0611a97319

Observation 15ea280b-ca5b-4968-bfa2-489e37eabb39 · outbound

This paper cites Evaluating GPT-4 and ChatGPT on Japanese Medical Licensing Examinations.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Evaluating GPT-4 and ChatGPT on Japanese Medical Licensing Examinations

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:39:41.561909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:39:41.561909Z digest=sha256:dc5e3333bdc894fd803327f171bc4c21260d43908f757ba55600d596919ef96c

Observation 6eb9a93d-b5df-4203-8c67-3eb414012bdd · outbound

This paper cites Advances in Neural Information Processing Systems 36, 52430–52452 (2023).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Advances in Neural Information Processing Systems 36, 52430–52452 (2023)

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:44.516979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:41.728068Z digest=sha256:a957860bdce4867e7786d88d605c0c82d360a82b26dd7d47b9763ee6e1cee28c

Observation 25469ee5-777c-4eff-81d7-19685c55b6de · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:44.269056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:41.844522Z digest=sha256:1bd5132104c374579d8632c16272f7eb1768516c9c31c979e10b51af9a427869

Observation 71c0fcb2-ec5c-4f2f-a345-4377efcd3715 · outbound

This paper cites html, accessed 2025-06-07.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations html, accessed 2025-06-07

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:44.016572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:41.879923Z digest=sha256:2f22e12c4f1692bf5056c476c497d7d81286cf5da5ad8d0933c96916d5487016

Observation 88b648ca-0929-4668-a34f-202c24122b8a · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:43.756452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:41.927605Z digest=sha256:6355553725ced17db767228efe7478ea04ea3636de9d2c05e140f798090a70cb

Observation fed79d2e-52f3-4942-8150-f4e0437c148a · outbound

This paper cites JJDEA40, 3–10 (2024).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations JJDEA40, 3–10 (2024)

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:43.589654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:41.989459Z digest=sha256:1551c595e55b0d0c481a9f49cd9cb2bccbce031c373a715c4fc45590e8830458

Observation 7494be47-5966-49ea-8042-2d749708eb89 · outbound

This paper cites an unresolved cited work.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:39:43.479864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:42.034777Z digest=sha256:d933f0e2f62b2ee4424b9db6a6b3f9754b54d1b73f75ed31a399d5d6b1ba602b

Observation 13117dba-ec30-4cf4-aeca-4b095e4f3f0c · outbound

This paper cites In: Conference on health, inference, and learning.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations In: Conference on health, inference, and learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:43.328755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:42.087521Z digest=sha256:9d6025724efbbbec8558d087bd0afcd2ac7c829630268ff966520b0e2b52c437

Observation 46ab7c81-32c5-40c9-a13c-d56e4dcf76c1 · outbound

This paper cites guppy.jp/, accessed: 2025-06-05.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations guppy.jp/, accessed: 2025-06-05

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:43.154140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:42.148285Z digest=sha256:dc502bb596eed2080fe380cf6fc9ea609f007e34cb751ec7bbee3637656dea37

Observation 79da25e9-5225-4eaa-afb2-ec965edc9699 · outbound

This paper cites Bell System Technical Journal27(3), 379–423 (1948).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Bell System Technical Journal27(3), 379–423 (1948)

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:39:42.193816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:39:42.193816Z digest=sha256:bd5cae2bd22ee58823a353f0663c5d6c3e29c72b151d792661d7f8f6dda050ea

Observation fc12455e-3dac-4895-a750-c48eaca57604 · outbound

This paper cites Scientific Reports14(1), 9330 (2024).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Scientific Reports14(1), 9330 (2024)

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:43.040530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:42.253208Z digest=sha256:6592fa1d765a9e578e04e58db93ab2d16ec1c2d97a9bf23f1a096449df47f2a4

Observation 7798c562-97c6-46ae-a591-2a1f0aa362b5 · outbound

This paper cites A Japanese Language Model and Three New Evaluation Benchmarks for Pharmaceutical NLP.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations A Japanese Language Model and Three New Evaluation Benchmarks for Pharmaceutical NLP

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:39:42.647730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:42.317690Z digest=sha256:b38731939ed6fa821ea54e754717ef8d29182ba6ff46a9667ee40273a0b1ff4c

Observation f3a30e7f-1d35-4b32-93e0-0085db6287a8 · outbound

This paper cites JMIR medical education9(1), e48002 (2023).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations JMIR medical education9(1), e48002 (2023)

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:42.891688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:42.396218Z digest=sha256:a60f81e6d0d4a646e893f477e45100ed39f9c55aa79a4d524df0a47361677a53

Observation dceb8435-7d19-409c-bbb2-7283f6464edb · outbound

This paper cites Advances in Neural Information Processing Systems35, 24824–24837 (2022).

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations Advances in Neural Information Processing Systems35, 24824–24837 (2022)

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:39:42.767195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:39:42.448581Z digest=sha256:264211fda9d4042b40bd9cf55dde1e66f2cbb1344e0d431b6de64961874d3981

Observation 4c4cdb70-7b00-4997-831d-c99d919ca950 · outbound

This paper cites MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine.

KokushiMD-10: Benchmark for Evaluating Large Language Models on Ten Japanese National Healthcare Licensing Examinations MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:39:42.506184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:39:42.506184Z digest=sha256:7792d967382e2f5928c955eb85502266a0c1f199faec4f1de1692f535d5ef1b1

Pith citing papers

No inbound Pith citation observations are available.