Pith. sign in

REVIEW 1 cited by

Revealing Treatment Non-Adherence Bias in Clinical Machine Learning Using Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2502.19625 v2 pith:XHWETILD submitted 2025-02-26 cs.LG

classification cs.LG
keywords non-adherencetreatmentclinicalbiaslearningmachinemodelpatients
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Machine learning systems trained on electronic health records (EHRs) increasingly guide treatment decisions, but their reliability depends on the critical assumption that patients follow the prescribed treatments recorded in EHRs. Using EHR data from 3,623 hypertension patients, we investigate how treatment non-adherence introduces implicit bias that can fundamentally distort both causal inference and predictive modeling. By extracting patient adherence information from clinical notes using a large language model (LLM), we identify 786 patients (21.7%) with medication non-adherence. We further uncover key demographic and clinical factors associated with non-adherence, as well as patient-reported reasons including side effects and difficulties obtaining refills. Our findings demonstrate that this implicit bias can not only reverse estimated treatment effects, but also degrade model performance by up to 5% while disproportionately affecting vulnerable populations by exacerbating disparities in decision outcomes and model error rates. This highlights the importance of accounting for treatment non-adherence in developing responsible and equitable clinical machine learning systems.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Local MDI+: Local Feature Importances for Tree-Based Models

    cs.LG 2025-06 conditional novelty 6.0 of 10

    Local MDI+ computes sample-specific feature importances for random forests and boosted trees by combining tree split structure with regularized linear models, and it outperforms LIME, TreeSHAP, and Local MDI at identi...

Pith tools