Pith. sign in

REVIEW 1 cited by

A Rainbow in Deep Network Black Boxes

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.18512 v3 pith:FQS5SHOU submitted 2023-05-29 cs.LG cs.CVeess.SP

classification cs.LGcs.CVeess.SP
keywords networksdeeprainbowfeaturenetworkrandomtrainedcentral
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

A central question in deep learning is to understand the functions learned by deep networks. What is their approximation class? Do the learned weights and representations depend on initialization? Previous empirical work has evidenced that kernels defined by network activations are similar across initializations. For shallow networks, this has been theoretically studied with random feature models, but an extension to deep networks has remained elusive. Here, we provide a deep extension of such random feature models, which we call the rainbow model. We prove that rainbow networks define deterministic (hierarchical) kernels in the infinite-width limit. The resulting functions thus belong to a data-dependent RKHS which does not depend on the weight randomness. We also verify numerically our modeling assumptions on deep CNNs trained on image classification tasks, and show that the trained networks approximately satisfy the rainbow hypothesis. In particular, rainbow networks sampled from the corresponding random feature model achieve similar performance as the trained networks. Our results highlight the central role played by the covariances of network weights at each layer, which are observed to be low-rank as a result of feature learning.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Training Language Models to Use Prolog as a Tool

    cs.CL 2025-12 unverdicted novelty 6.0 of 10

    GRPO can teach a 3B language model to emit executable Prolog, but the highest-accuracy models often hardcode answers instead of reasoning in Prolog, producing an accuracy–auditability trade-off.

Pith tools