Error estimates for DeepONets: a deep learning framework in infinite dimensions
Open access
Datum
2022-01Typ
- Journal Article
ETH Bibliographie
yes
Altmetrics
Abstract
DeepONets have recently been proposed as a framework for learning nonlinear operators mapping between infinite-dimensional Banach spaces. We analyze DeepONets and prove estimates on the resulting approximation and generalization errors. In particular, we extend the universal approximation property of DeepONets to include measurable mappings in non-compact spaces. By a decomposition of the error into encoding, approximation and reconstruction errors, we prove both lower and upper bounds on the total error, relating it to the spectral decay properties of the covariance operators, associated with the underlying measures. We derive almost optimal error bounds with very general affine reconstructors and with random sensor locations as well as bounds on the generalization error, using covering number arguments. We illustrate our general framework with four prototypical examples of nonlinear operators, namely those arising in a nonlinear forced ordinary differential equation, an elliptic partial differential equation (PDE) with variable coefficients and nonlinear parabolic and hyperbolic PDEs. While the approximation of arbitrary Lipschitz operators by DeepONets to accuracy ϵϵ is argued to suffer from a ‘curse of dimensionality’ (requiring a neural networks of exponential size in 1/ϵ1/ϵ), in contrast, for all the above concrete examples of interest, we rigorously prove that DeepONets can break this curse of dimensionality (achieving accuracy ϵϵ with neural networks of size that can grow algebraically in 1/ϵ1/ϵ).Thus, we demonstrate the efficient approximation of a potentially large class of operators with this machine learning framework. Mehr anzeigen
Persistenter Link
https://doi.org/10.3929/ethz-b-000558811Publikationsstatus
publishedExterne Links
Zeitschrift / Serie
Transactions of Mathematics and Its ApplicationsBand
Seiten / Artikelnummer
Verlag
Oxford University PressThema
operator learning; deep neural networks; PDEsOrganisationseinheit
03851 - Mishra, Siddhartha / Mishra, Siddhartha
Förderung
770880 - Computation and analysis of statistical solutions of fluid flow (EC)
Zugehörige Publikationen und Daten
Is new version of: http://hdl.handle.net/20.500.11850/491123
ETH Bibliographie
yes
Altmetrics