REVIEW 4 cited by
Ensemble and Mixture-of-Experts DeepONets For Operator Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
abstract
We present a novel deep operator network (DeepONet) architecture for operator learning, the ensemble DeepONet, that allows for enriching the trunk network of a single DeepONet with multiple distinct trunk networks. This trunk enrichment allows for greater expressivity and generalization capabilities over a range of operator learning problems. We also present a spatial mixture-of-experts (MoE) DeepONet trunk network architecture that utilizes a partition-of-unity (PoU) approximation to promote spatial locality and model sparsity in the operator learning problem. We first prove that both the ensemble and PoU-MoE DeepONets are universal approximators. We then demonstrate that ensemble DeepONets containing a trunk ensemble of a standard trunk, the PoU-MoE trunk, and/or a proper orthogonal decomposition (POD) trunk can achieve 2-4x lower relative $\ell_2$ errors than standard DeepONets and POD-DeepONets on both standard and challenging new operator learning problems involving partial differential equations (PDEs) in two and three dimensions. Our new PoU-MoE formulation provides a natural way to incorporate spatial locality and model sparsity into any neural network architecture, while our new ensemble DeepONet provides a powerful and general framework for incorporating basis enrichment in scientific machine learning architectures for operator learning.
Forward citations
Cited by 4 Pith papers
-
SPAMoE: Spectrum-Aware Hybrid Operator Framework for Full-Waveform Inversion
SPAMoE reduces average MAE by 44.4% on ten OpenFWI sub-datasets via a spectral-preserving DINO encoder plus frequency-routed MoE of FNO, MNO and LNO experts.
-
Time Resolution Independent Operator Learning
A DeepONet with a neural controlled differential equation branch and a trunk that takes space and time as inputs predicts transient mechanical fields from load histories at arbitrary spatiotemporal query points.
-
From Classification to Regression: Using a Fruitfly to Solve Equations
Nonlinear maps are approximated by softmax-weighted reconstruction from a finite library of representative patterns and their stored responses, with demos on Lotka–Volterra, Lorenz, 1D fits, and 1D Poisson.
-
Enhanced accuracy through ensembling of randomly initialized auto-regressive models for time-dependent PDEs
Deep ensembles of randomly initialized autoregressive models reduce long-horizon prediction error compared to any single model across three PDE-driven dynamical systems.
Discussion (0). Continue with ORCID to comment.