Pith. sign in

REVIEW 1 cited by

Efficient Maximal Coding Rate Reduction by Variational Forms

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2204.00077 v1 pith:TIOMBKC7 submitted 2022-03-31 cs.LG cs.CV

classification cs.LGcs.CV
keywords trainingformsobjectivesignificantbeencodinglog-determinantmaximal
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

The principle of Maximal Coding Rate Reduction (MCR$^2$) has recently been proposed as a training objective for learning discriminative low-dimensional structures intrinsic to high-dimensional data to allow for more robust training than standard approaches, such as cross-entropy minimization. However, despite the advantages that have been shown for MCR$^2$ training, MCR$^2$ suffers from a significant computational cost due to the need to evaluate and differentiate a significant number of log-determinant terms that grows linearly with the number of classes. By taking advantage of variational forms of spectral functions of a matrix, we reformulate the MCR$^2$ objective to a form that can scale significantly without compromising training accuracy. Experiments in image classification demonstrate that our proposed formulation results in a significant speed up over optimizing the original MCR$^2$ objective directly and often results in higher quality learned representations. Further, our approach may be of independent interest in other models that require computation of log-determinant forms, such as in system identification or normalizing flow models.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Token Statistics Transformer: Linear-Time Attention via Variational Rate Reduction

    cs.LG 2024-12 conditional novelty 6.0 of 10

    A variational reformulation of the MCR2 objective yields a linear-complexity attention operator, ToST, that matches transformer performance without computing pairwise token similarities.

Pith tools