REVIEW 2 cited by
Stabilizing Equilibrium Models by Jacobian Regularization
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Deep equilibrium networks (DEQs) are a new class of models that eschews traditional depth in favor of finding the fixed point of a single nonlinear layer. These models have been shown to achieve performance competitive with the state-of-the-art deep networks while using significantly less memory. Yet they are also slower, brittle to architectural choices, and introduce potential instability to the model. In this paper, we propose a regularization scheme for DEQ models that explicitly regularizes the Jacobian of the fixed-point update equations to stabilize the learning of equilibrium models. We show that this regularization adds only minimal computational cost, significantly stabilizes the fixed-point convergence in both forward and backward passes, and scales well to high-dimensional, realistic domains (e.g., WikiText-103 language modeling and ImageNet classification). Using this method, we demonstrate, for the first time, an implicit-depth model that runs with approximately the same speed and level of performance as popular conventional deep networks such as ResNet-101, while still maintaining the constant memory footprint and architectural simplicity of DEQs. Code is available at https://github.com/locuslab/deq .
Forward citations
Cited by 2 Pith papers
-
SILVA Networks as Structured Implicit Layers and Vector Attractors via Dynamic Interaction Fields
A fixed-point layer whose update is explicitly split into stimulus, local, global, and damping terms runs across images and graphs; its global term is load-bearing only on the CLUSTER long-range benchmark.
-
End-to-End Learning of Safe Optimal Feedback Control in High Dimensions with Control Barrier Function Layers
A scalable end-to-end training method for neural controllers with embedded control-barrier-function safety filters, demonstrated up to 1200 state dimensions and 400 control dimensions, with convergence guarantees unde...
Discussion (0). Continue with ORCID to comment.