Pith. sign in

REVIEW 3 cited by

When Deep Learning Meets Polyhedral Theory: A Survey

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.00241 v4 pith:EUVXLN2R submitted 2023-04-29 math.OC cs.LG

classification math.OCcs.LG
keywords linearnetworksneuraldeepbecamelearningnetworkpiecewise
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

In the past decade, deep learning became the prevalent methodology for predictive modeling thanks to the remarkable accuracy of deep neural networks in tasks such as computer vision and natural language processing. Meanwhile, the structure of neural networks converged back to simpler representations based on piecewise constant and piecewise linear functions such as the Rectified Linear Unit (ReLU), which became the most commonly used type of activation function in neural networks. That made certain types of network structure $\unicode{x2014}$such as the typical fully-connected feedforward neural network$\unicode{x2014}$ amenable to analysis through polyhedral theory and to the application of methodologies such as Linear Programming (LP) and Mixed-Integer Linear Programming (MILP) for a variety of purposes. In this paper, we survey the main topics emerging from this fast-paced area of work, which bring a fresh perspective to understanding neural networks in more detail as well as to applying linear optimization techniques to train, verify, and reduce the size of such networks.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Time to Spike? Understanding the Representational Power of Spiking Neural Networks in Discrete Time

    cs.LG 2025-05 conditional novelty 7.0 of 10

    Discrete-time LIF spiking networks realize piecewise constant functions on polyhedral regions, and each first-layer neuron generates only O(T^2) parallel hyperplanes over T time steps, not exponentially many.

  2. Conformal Mixed-Integer Constraint Learning with Feasibility Guarantees

    cs.LG 2025-06 reject novelty 6.0 of 10

    C-MICL embeds conformal prediction sets into mixed-integer constraint learning, claiming a 1-alpha probability that optimized solutions are feasible for the true unknown constraint.

  3. Global optimization of graph acquisition functions for neural architecture search

    cs.LG 2025-05 conditional novelty 6.0 of 10

    A mixed-integer programming formulation globally optimizes graph Bayesian optimization acquisition functions for neural architecture search, with a proved graph encoding.

Pith tools