Pith. sign in

REVIEW 1 cited by

Error Rate Bounds and Iterative Weighted Majority Voting for Crowdsourcing

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1411.4086 v1 pith:62H36ZRT submitted 2014-11-15 stat.ML cs.HCcs.LGmath.PRmath.STstat.TH

classification stat.MLcs.HCcs.LGmath.PRmath.STstat.TH
keywords errormajorityratevotingcrowdsourcingweightedboundsiwmv
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Crowdsourcing has become an effective and popular tool for human-powered computation to label large datasets. Since the workers can be unreliable, it is common in crowdsourcing to assign multiple workers to one task, and to aggregate the labels in order to obtain results of high quality. In this paper, we provide finite-sample exponential bounds on the error rate (in probability and in expectation) of general aggregation rules under the Dawid-Skene crowdsourcing model. The bounds are derived for multi-class labeling, and can be used to analyze many aggregation methods, including majority voting, weighted majority voting and the oracle Maximum A Posteriori (MAP) rule. We show that the oracle MAP rule approximately optimizes our upper bound on the mean error rate of weighted majority voting in certain setting. We propose an iterative weighted majority voting (IWMV) method that optimizes the error rate bound and approximates the oracle MAP rule. Its one step version has a provable theoretical guarantee on the error rate. The IWMV method is intuitive and computationally simple. Experimental results on simulated and real data show that IWMV performs at least on par with the state-of-the-art methods, and it has a much lower computational cost (around one hundred times faster) than the state-of-the-art methods.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Meta-learning Representations for Learning from Multiple Annotators

    cs.LG 2025-06 accept novelty 6.0 of 10

    A meta-learned embedding with an EM-adapted Gaussian mixture and annotator confusion matrices improves few-shot classification from multiple noisy annotators.

Pith tools