Pith. sign in

REVIEW 5 cited by

Simplicity Bias in Transformers and their Ability to Learn Sparse Boolean Functions

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2211.12316 v2 pith:BX6Z2CHB submitted 2022-11-22 cs.LG cs.CL

classification cs.LGcs.CL
keywords transformersfunctionsbooleansensitivitymodelsrecurrentdespitegeneralization
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Despite the widespread success of Transformers on NLP tasks, recent works have found that they struggle to model several formal languages when compared to recurrent models. This raises the question of why Transformers perform well in practice and whether they have any properties that enable them to generalize better than recurrent models. In this work, we conduct an extensive empirical study on Boolean functions to demonstrate the following: (i) Random Transformers are relatively more biased towards functions of low sensitivity. (ii) When trained on Boolean functions, both Transformers and LSTMs prioritize learning functions of low sensitivity, with Transformers ultimately converging to functions of lower sensitivity. (iii) On sparse Boolean functions which have low sensitivity, we find that Transformers generalize near perfectly even in the presence of noisy labels whereas LSTMs overfit and achieve poor generalization accuracy. Overall, our results provide strong quantifiable evidence that suggests differences in the inductive biases of Transformers and recurrent models which may help explain Transformer's effective generalization performance despite relatively limited expressiveness.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Can Transformers Really Do It All? On the Compatibility of Inductive Biases Across Tasks

    cs.LG 2026-07 conditional novelty 7.0 of 10

    Learned replacement non-linearities show transformers are rarely optimal for algorithmic tasks, with benefits that are task-specific, while language/code gains are smaller and more transferable.

  2. When Do Neural Networks Learn World Models?

    cs.LG 2025-02 conditional novelty 7.0 of 10

    With Boolean variables, a low-degree bias, and a task distribution weighted toward simple functions of the latents, multi-task training provably recovers the latent world model up to permutations and negations.

  3. Parity Requires Unified Input Dependence and Negative Eigenvalues in SSMs

    cs.LG 2025-08 conditional novelty 6.0 of 10

    For diagonal state space models, parity cannot be solved by combining input-dependent non-negative layers (Mamba) with input-independent negative-eigenvalue layers (S4D); a single layer needs both properties.

  4. In-Context Learning Strategies Emerge Rationally

    cs.LG 2025-06 conditional novelty 6.0 of 10

    Transformer in-context learning is modeled as a posterior-weighted mixture of memorizing and generalizing Bayesian predictors, with a loss-complexity tradeoff governed by three fitted parameters.

  5. Characterising the Inductive Biases of Neural Networks on Boolean Data

    cs.LG 2025-05 conditional novelty 6.0 of 10

    Randomly initialized depth-2 discrete networks are a priori biased toward Boolean functions with small disjunctive normal form complexity, and this bias quantitatively predicts training and generalization behavior inc...

Pith tools