Pith. sign in

Paper Citation Record · LEDGER

MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2412.18947.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.18947 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:19:11.288187Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:38:28.969559Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dd44847f-7dc3-4363-a261-9311cbbb39d0 · inbound

Loki's Dance of Illusions: A Comprehensive Survey of Hallucination in Large Language Models cites this paper.

Loki's Dance of Illusions: A Comprehensive Survey of Hallucination in Large Language Models MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 200

Resolution
unresolved
no resolver link, observed 2026-08-07T10:19:11.288187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:19:11.288187Z digest=sha256:f50e36f6801f1cf4f786e7e70ff33fb31a13280a9faa1b46ec55135002eeb3a7

Observation 04ce2e7e-d737-4d09-b44d-e3a85a5e8824 · inbound

A comprehensive taxonomy of hallucinations in Large Language Models cites this paper.

A comprehensive taxonomy of hallucinations in Large Language Models MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-06T05:29:21.056789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:29:21.056789Z digest=sha256:9abc8bb372adf9a523ed74aea065dcb6541677b435751e2c4ff205345a7a0e9f

Observation 038607a1-13e2-41aa-a690-ab5f2d7550b7 · inbound

TerraMAE: Learning Spatial-Spectral Representations from Hyperspectral Earth Observation Data via Adaptive Masked Autoencoders cites this paper.

TerraMAE: Learning Spatial-Spectral Representations from Hyperspectral Earth Observation Data via Adaptive Masked Autoencoders MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T22:26:40.829467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:26:40.829467Z digest=sha256:6db0dfb13a8c25fec8461eb3160273030cd54f2010121d0a88b8f3c26a327b4b

Observation baa9a7b1-57c2-4848-91ed-e9d0ce1365f3 · inbound

Trustworthy Medical Imaging with Large Language Models: A Study of Hallucinations Across Modalities cites this paper.

Trustworthy Medical Imaging with Large Language Models: A Study of Hallucinations Across Modalities MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T22:26:06.534145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:26:06.534145Z digest=sha256:f28bba1f63e67f8d812e297e68051e1a413919405bb9c3a6e1edce43b8b831d5

Observation 456e0b79-1c9f-4d9e-b161-5bfb1f2679aa · inbound

A Multi-Task Evaluation of LLMs' Processing of Academic Text Input cites this paper.

A Multi-Task Evaluation of LLMs' Processing of Academic Text Input MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T19:49:52.789854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:49:52.789854Z digest=sha256:0b4e71574114f96f394e7aac2a50d6bb97bc14c513f6cc10a429ca368e50d0f2

Observation f75f4bb8-a881-4007-bea9-ad58fdb4dc7d · inbound

PARALLAX: Separating Genuine Hallucination Detection from Benchmark Construction Artifacts cites this paper.

PARALLAX: Separating Genuine Hallucination Detection from Benchmark Construction Artifacts MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:12:44.896323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-19T20:09:47.750043Z digest=sha256:efbbfb1266563df75d1a9a99efe786c80e781e4324359db63b7be93209f7b702

Observation 1e92785f-c4af-4572-bda8-1688b7c71bfa · inbound

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models cites this paper.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:0e99e9f94d142bc4dd90acf90cf0bb3f0451e0e93322897dc588f0c4a736fa20

Observation be9bd705-9c52-4a50-8ad7-ff9c51334350 · inbound

Hallucination in Medical Imaging AI: A Cross-Modality Analytical Framework for Taxonomy, Detection, and Mitigation under Regulatory Constraints cites this paper.

Hallucination in Medical Imaging AI: A Cross-Modality Analytical Framework for Taxonomy, Detection, and Mitigation under Regulatory Constraints MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:38:28.970985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T07:00:54.990637Z digest=sha256:fc31ddad9f6786b0b535447b6b060032f50e3957d2d41466c5d71ba88c55ccd6

Observation eff9e784-fec0-4680-a789-70fef4a19c29 · inbound

KnowHal: A Knowledge-Driven Benchmark for Comprehensive Multimodal Hallucination Evaluation cites this paper.

KnowHal: A Knowledge-Driven Benchmark for Comprehensive Multimodal Hallucination Evaluation MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T12:40:15.241605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:40:15.241605Z digest=sha256:ea288c78d34ea8b15a4cf472072f459a3b4979eaa77a2b091ce519ee3014dd06