Pith. sign in

REVIEW 2 cited by

Stabilizing Equilibrium Models by Jacobian Regularization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2106.14342 v1 pith:SJ7Q4MUM submitted 2021-06-28 cs.LG stat.ML

classification cs.LGstat.ML
keywords modelsdeepequilibriumnetworksregularizationarchitecturaldeqsfixed-point
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Deep equilibrium networks (DEQs) are a new class of models that eschews traditional depth in favor of finding the fixed point of a single nonlinear layer. These models have been shown to achieve performance competitive with the state-of-the-art deep networks while using significantly less memory. Yet they are also slower, brittle to architectural choices, and introduce potential instability to the model. In this paper, we propose a regularization scheme for DEQ models that explicitly regularizes the Jacobian of the fixed-point update equations to stabilize the learning of equilibrium models. We show that this regularization adds only minimal computational cost, significantly stabilizes the fixed-point convergence in both forward and backward passes, and scales well to high-dimensional, realistic domains (e.g., WikiText-103 language modeling and ImageNet classification). Using this method, we demonstrate, for the first time, an implicit-depth model that runs with approximately the same speed and level of performance as popular conventional deep networks such as ResNet-101, while still maintaining the constant memory footprint and architectural simplicity of DEQs. Code is available at https://github.com/locuslab/deq .

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SILVA Networks as Structured Implicit Layers and Vector Attractors via Dynamic Interaction Fields

    cs.LG 2026-07 conditional novelty 6.0 of 10

    A fixed-point layer whose update is explicitly split into stimulus, local, global, and damping terms runs across images and graphs; its global term is load-bearing only on the CLUSTER long-range benchmark.

  2. End-to-End Learning of Safe Optimal Feedback Control in High Dimensions with Control Barrier Function Layers

    cs.LG 2026-07 conditional novelty 6.0 of 10

    A scalable end-to-end training method for neural controllers with embedded control-barrier-function safety filters, demonstrated up to 1200 state dimensions and 400 control dimensions, with convergence guarantees unde...

Pith tools