Pith. sign in

REVIEW 2 cited by

Easy Learning from Label Proportions

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2302.03115 v2 pith:TJ5FNXJG submitted 2023-02-06 cs.LG stat.ML

classification cs.LGstat.ML
keywords learninglevellossapproacharbitraryindividualinstancelabel
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

We consider the problem of Learning from Label Proportions (LLP), a weakly supervised classification setup where instances are grouped into "bags", and only the frequency of class labels at each bag is available. Albeit, the objective of the learner is to achieve low task loss at an individual instance level. Here we propose Easyllp: a flexible and simple-to-implement debiasing approach based on aggregate labels, which operates on arbitrary loss functions. Our technique allows us to accurately estimate the expected loss of an arbitrary model at an individual level. We showcase the flexibility of our approach by applying it to popular learning frameworks, like Empirical Risk Minimization (ERM) and Stochastic Gradient Descent (SGD) with provable guarantees on instance level performance. More concretely, we exhibit a variance reduction technique that makes the quality of LLP learning deteriorate only by a factor of k (k being bag size) in both ERM and SGD setups, as compared to full supervision. Finally, we validate our theoretical results on multiple datasets demonstrating our algorithm performs as well or better than previous LLP approaches in spite of its simplicity.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Aggregating Data for Optimal and Private Learning

    cs.LG 2024-11 conditional novelty 6.0 of 10

    For multiple instance regression and learning from label proportions, optimal data bagging for linear regression reduces approximately to k-means clustering over labels or features, and the mechanisms can be made labe...

  2. Learning from Label Proportions and Covariate-shifted Instances

    cs.LG 2024-11 conditional novelty 6.0 of 10

    A new bag-level alignment loss (BagCSI) for covariate-shifted hybrid LLP, with a generalization error bound and consistent gains on large bags.

Pith tools