1
Return

An optimal Petrov–Galerkin framework for operator networks

delete2026-05-08
delete0
delete
OA
AI
P
Philip Charles
D
Deep Ray *
Y
Yue Yu
J
Joost Prins
H
Hugo Melchers
M
Michael Abdelmalik
J
Jeffrey Cochran
A
Assad A. Oberai
T
Thomas J.R. Hughes
M
Mats G. Larson
DOI:10.1016/j.cma.2026.119046delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
The optimal Petrov–Galerkin formulation to solve partial differential equations (PDEs) recovers the best approximation in a specified finite-dimensional (trial) space with respect to a suitable norm. However, the recovery of this optimal solution is contingent on being able to construct the optimal weighting functions associated with the trial basis. While explicit constructions are available for simple one- and two-dimensional problems, such constructions for a general multidimensional problem remain elusive. In the present work, we revisit the optimal Petrov–Galerkin formulation through the lens of deep learning. We propose an operator network framework called Petrov–Galerkin Variationally Mimetic Operator Network (PG-VarMiON), which emulates the optimal Petrov–Galerkin weak form of the underlying PDE. The PG-VarMiON is trained in a supervised manner using a labeled dataset comprising the PDE data and the corresponding PDE solution, with the training loss depending on the choice of the optimal norm. The special architecture of the PG-VarMiON allows it to implicitly learn the optimal weighting functions, thus endowing the proposed operator network with the ability to generalize well beyond the training set. We derive approximation error estimates for PG-VarMiON, highlighting the contributions of various error sources, particularly the error in learning the true weighting functions. Several numerical results are presented for the advection-diffusion equation to demonstrate the efficacy of the proposed method. By embedding the Petrov–Galerkin structure into the network architecture, PG-VarMiON exhibits greater robustness and improved generalization compared to other popular deep operator frameworks, particularly when the training data is limited.
Keywords:
Operator learning
Optimal Petrov–Galerkin
Deep learning
Generalization error
Advection–diffusion equation
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Computer Methods in Applied Mechanics and Engineering cover
Computer Methods in Applied Mechanics and Engineering
IF:
7.3
Papers:
1.3W
Citations:
5.6W

Organization

U
University of Texas
Scholars:
780
Papers: 401
Citations: 9.5K
U
university of southern california
Scholars:
4.5W
Papers: 3.8W
Citations: 51
U
umea university
Scholars:
1.1K
Papers: 499
Citations: 0
L
Lehigh University
Scholars:
4.8K
Papers: 5.1K
Citations: 6.3K
U
university of maryland
Scholars:
4.0K
Papers: 1.9K
Citations: 1
E
Eindhoven University of Technology
Scholars:
1.6W
Papers: 1.5W
Citations: 2.2W
Cited Papers

Cited Papers

Citing Papers

Citing Papers