REVIEW 2 cited by
Convex Potential Flows: Universal Probability Distributions with Optimal Transport and Convex Optimization
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Flow-based models are powerful tools for designing probabilistic models with tractable density. This paper introduces Convex Potential Flows (CP-Flow), a natural and efficient parameterization of invertible models inspired by the optimal transport (OT) theory. CP-Flows are the gradient map of a strongly convex neural potential function. The convexity implies invertibility and allows us to resort to convex optimization to solve the convex conjugate for efficient inversion. To enable maximum likelihood training, we derive a new gradient estimator of the log-determinant of the Jacobian, which involves solving an inverse-Hessian vector product using the conjugate gradient method. The gradient estimator has constant-memory cost, and can be made effectively unbiased by reducing the error tolerance level of the convex optimization routine. Theoretically, we prove that CP-Flows are universal density approximators and are optimal in the OT sense. Our empirical results show that CP-Flow performs competitively on standard benchmarks of density estimation and variational inference.
Forward citations
Cited by 2 Pith papers
-
On the Depth of Monotone ReLU Neural Networks and ICNNs
Monotone ReLU networks cannot compute or approximate the max function, input convex networks need depth n to compute the n-ary max, and some depth-2 ReLU networks beat every depth-k ICNN.
-
Profiling systematic uncertainties in Simulation-Based Inference with Factorizable Normalizing Flows
Systematic uncertainties can be profiled in unbinned likelihood fits by factorizing the normalizing-flow transformation into per-nuisance linear-plus-quadratic terms and training amortized over the nuisance space.
Discussion (0). Continue with ORCID to comment.