Pith. sign in

REVIEW

Cooperative Online Learning with Feedback Graphs

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2106.04982 v5 pith:R4TWK77I submitted 2021-06-09 cs.LG stat.ML

classification cs.LGstat.ML
keywords feedbackcooperativelearningnetworkonlineboundcasescommunication
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We study the interplay between communication and feedback in a cooperative online learning setting, where a network of communicating agents learn a common sequential decision-making task through a feedback graph. We bound the network regret in terms of the independence number of the strong product between the communication network and the feedback graph. Our analysis recovers as special cases many previously known bounds for cooperative online learning with expert or bandit feedback. We also prove an instance-based lower bound, demonstrating that our positive results are not improvable except in pathological cases. Experiments on synthetic data confirm our theoretical findings.

Discussion (0). Continue with ORCID to comment.

Pith tools