Skip to content
Preprint

FAST-DeepONet: Factor-Augmented Branch Representations for High-Dimensional PDE Inputs in the Small-Sample Regime

Aug 2026 · 0 citations · 25 references
Computer Science

TL;DR

FAST-DeepONet, a branch representation combining a fixed spectral path with a regularized projection of the orthogonal residual, in which the directional penalty acts on the effective residual map after each of its rows is normalized, is introduced.

Abstract

Deep operator networks can become statistically unstable when partial differential equation inputs are observed at thousands of strongly correlated sensors but only a small number of operator samples is available. We introduce FAST-DeepONet, a branch representation combining a fixed spectral path with a regularized projection of the orthogonal residual, in which the directional penalty acts on the effective residual map after each of its rows is normalized. On Navier--Stokes flow a plain DeepONet degrades from $0.0394$ to $0.1556$ mean relative $L_2$ error as the branch grows from $129$ to $8193$ coordinates, while FAST-DeepONet stays near $0.04$, so the sensor grid can be refined without a statistical penalty. Across independent test sets for Navier--Stokes flow, Darcy flow, and signed terminal wavefield prediction it lowers mean relative $L_2$ error by $4.7\%$ to $37.0\%$ with three to seven times fewer trainable parameters. A spectral-only branch sharing the same basis separates the two paths: the fixed spectral path carries the improvement on Navier--Stokes and Darcy, while terminal wave prediction requires the residual path together with its directional penalty. FAST-DeepONet targets coordinate-query architectures and trains on solution values alone.

View source

Similar papers

Preprint Sep 2026

Deep Koopman Sensing

Real-time reconstruction of fluid flows from sparse sensor measurements is important for both physical understanding and flow control. When first-principles models are too expensive for online data assimilation (DA), learned reduced-order models provide an efficient alternative, but are commonly optimized for forward p...

Nithin Somasekharan, Yadi Cao, Shao-Wu Pan · 0 citations
Preprint Aug 2026

A Generalized Ridge Regression and Convolutional LASSO

It is proved that the extracted trend is invariant across representations, and a closed-form Bregman-type divergence quantifying their disagreement under a shared coefficient vector is obtained, and this divergence vanishes as $\lambda\to\infty$.

Shintaro Yoshizawa · 0 citations
#machine learning Preprint Sep 2026

Disentangling Attention in Deep Operator Learning: A Controlled Study of Data-Driven and Physics-Informed Architectures

Overall, query-dependent cross-attention is the most reliable mechanism, whereas branch self-attention is most useful for large, spatially complex functional inputs, whereas branch self-attention is most useful for large, spatially complex functional inputs.

Amar Alem Koric, Qi-Bang Liu, S. Koric · 0 citations
Aug 2026

Physics-informed feature decomposition in residual dense block neural networks for incompressible viscous flow

The growing application of physics-informed neural networks (PINNs) for solving parametric partial differential equations (PDEs) in fluid dynamics has demonstrated their potential for modeling complex multiscale flows; however, conventional PINNs often exhibit spectral bias and slow, unstable convergence, limiting accu...

Sarmad Iftikhar, Ishfaq Ahmad, Diltaj Ali et al. · 0 citations
#machine learning Preprint Sep 2026

Local gradient neural operator

Field temporal prediction and source identification constitute canonical problems in dynamical systems. Conventional approaches to these problems depend on a thorough understanding of the governing partial differential equations (PDEs). Recently, deep learning, as represented by neural operators, has provided a data-dr...

Bai-Ming Zhang, Jin-Song Tang, Ying Xu et al. · 0 citations
#machine learning Preprint Aug 2026

Sensitivity-Constrained Neural Operators for Data-Efficient Forward and Inverse Modeling of Partial Differential Equation Systems

Results support sampled sensitivity supervision as a practical way to improve neural PDE surrogates when forward accuracy, inverse stability, robustness, and computational cost must be considered together.

Abdolmehdi Behroozi, Chao-Peng Shen, Daniel Kifer et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.