Pith. sign in

Paper Citation Record · LEDGER

MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2407.10990.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.10990 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:47.005013Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T04:13:53.226694Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fe3c8a0f-4023-446d-9cd5-b0cd3398d2e1 · inbound

Data-Centric Foundation Models in Computational Healthcare: A Survey cites this paper.

Data-Centric Foundation Models in Computational Healthcare: A Survey MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 176

Resolution
verified exact
arxiv_id, observed 2026-05-24T04:13:53.229440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-24T04:13:05.328492Z digest=sha256:4321b3861ee1883932e8d184aa3feb67a79c7814557c0d77f10571565c7e9b01

Observation ce415c53-29d1-43eb-9314-0b3d91eec600 · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 215

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:47.005013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:47.005013Z digest=sha256:3efd36a4d15c4988556ff40e84db8f203f8c3da0a3bb28c1de32f5ff230c9116

Observation 1c088b28-fcb0-4393-bba6-5d3efbb4fcf1 · inbound

A Novel Evaluation Benchmark for Medical LLMs: Illuminating Safety and Effectiveness in Clinical Domains cites this paper.

A Novel Evaluation Benchmark for Medical LLMs: Illuminating Safety and Effectiveness in Clinical Domains MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T10:46:10.452156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:46:10.452156Z digest=sha256:0334f64cbe833a567871e5de8bbf344e4353ad81d48e60e9571b7cc98b82d60b

Observation 27d9ffa7-dd00-46aa-baa6-5383c9bec9e3 · inbound

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation cites this paper.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T05:07:42.040673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T05:07:42.040673Z digest=sha256:f9334b09f66d3d4730f2a97348e4feb43f9f78ddb3c575e4bcd310625e8de979

Observation cde0045c-0945-4411-bed9-9f84f6eb9ca6 · inbound

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation cites this paper.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:27.951292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:27.951292Z digest=sha256:7db07c5c60153ffe5d925667601ed8b09a2bf518f3b7863c9cac6969aef16565