REVIEW 5 cited by
Difficult Lessons on Social Prediction from Wisconsin Public Schools
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Early warning systems (EWS) are predictive tools at the center of recent efforts to improve graduation rates in public schools across the United States. These systems assist in targeting interventions to individual students by predicting which students are at risk of dropping out. Despite significant investments in their widespread adoption, there remain large gaps in our understanding of the efficacy of EWS, and the role of statistical risk scores in education. In this work, we draw on nearly a decade's worth of data from a system used throughout Wisconsin to provide the first large-scale evaluation of the long-term impact of EWS on graduation outcomes. We present empirical evidence that the prediction system accurately sorts students by their dropout risk. We also find that it may have caused a single-digit percentage increase in graduation rates, though our empirical analyses cannot reliably rule out that there has been no positive treatment effect. Going beyond a retrospective evaluation of DEWS, we draw attention to a central question at the heart of the use of EWS: Are individual risk scores necessary for effectively targeting interventions? We propose a simple mechanism that only uses information about students' environments -- such as their schools, and districts -- and argue that this mechanism can target interventions just as efficiently as the individual risk score-based mechanism. Our argument holds even if individual predictions are highly accurate and effective interventions exist. In addition to motivating this simple targeting mechanism, our work provides a novel empirical backbone for the robust qualitative understanding among education researchers that dropout is structurally determined. Combined, our insights call into question the marginal value of individual predictions in settings where outcomes are driven by high levels of inequality.
Forward citations
Cited by 5 Pith papers
-
Reconsidering Fairness Through Unawareness From the Perspective of Model Multiplicity
Omitting protected attributes can cut disparate impact substantially with negligible accuracy loss, and new model-multiplicity bounds explain when fairer unaware models exist.
-
Three Types of Calibration with Properties and their Semantic and Formal Relationships
Three families of calibration definitions, distribution, property, and decision calibration, are unified: distribution calibration implies the other two, and for binary outcomes all three collapse.
-
The Value of Prediction in Identifying the Worst-Off
Expanding screening capacity often improves identification of the worst-off more than improving prediction accuracy, except when predictions are very bad or nearly perfect.
-
Trading off performance and human oversight in algorithmic policy: evidence from Danish college admissions
Simple machine learning models predict college completion better than GPA or human rankings in Danish admissions, and most of the benefit of AI admissions comes from models that remain interpretable.
-
Heterogeneous participation and allocation skews: when is choice "worth it"?
Participatory mechanisms such as budgeting, 311 requests, and school choice tend to skew public resources toward advantaged participants, so the paper argues designers should pair preference use with strong default al...
Discussion (0). Continue with ORCID to comment.