Pith. sign in

REVIEW 1 cited by

Towards Human-AI Complementarity with Prediction Sets

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2405.17544 v2 pith:V4MFIL5V submitted 2024-05-27 cs.LG cs.CYcs.HC

classification cs.LGcs.CYcs.HC
keywords predictionsetsconformalconstructedexpertshumanlabelpredictions
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Decision support systems based on prediction sets have proven to be effective at helping human experts solve classification tasks. Rather than providing single-label predictions, these systems provide sets of label predictions constructed using conformal prediction, namely prediction sets, and ask human experts to predict label values from these sets. In this paper, we first show that the prediction sets constructed using conformal prediction are, in general, suboptimal in terms of average accuracy. Then, we show that the problem of finding the optimal prediction sets under which the human experts achieve the highest average accuracy is NP-hard. More strongly, unless P = NP, we show that the problem is hard to approximate to any factor less than the size of the label set. However, we introduce a simple and efficient greedy algorithm that, for a large class of expert models and non-conformity scores, is guaranteed to find prediction sets that provably offer equal or greater performance than those constructed using conformal prediction. Further, using a simulation study with both synthetic and real expert predictions, we demonstrate that, in practice, our greedy algorithm finds near-optimal prediction sets offering greater performance than conformal prediction.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. A No Free Lunch Theorem for Human-AI Collaboration

    cs.AI 2024-11 conditional novelty 7.0 of 10

    For calibrated binary predictors, any collaboration rule that is guaranteed to be at least as accurate as the worst agent must essentially always defer to a single agent.

Pith tools