CORTEXA
← Browse
arxivcs.LG2026-07-24

Energy Manifold Natural Gradient Descent: Riemannian Optimization for Neural PDE Solvers

Zhangyong Liang, Huanhuan Gao

Energy natural gradient descent (ENGD) aligns parameter updates with the curvature of an underlying function-space energy, but existing formulations assume an unconstrained Euclidean parameter domain. We introduce \EMNGDfull{}, a manifold optimization framework for physics-informed and variational neural PDE solvers whose parameters lie on a Riemannian manifold. EMNGD restricts the energy-induced quadratic model to feasible tangent directions and uses retractions to preserve parameter constraints throughout optimization. Under coercivity, we prove that the push-forward of the undamped EMNGD direction is the best feasible approximation to the function-space Newton vector in the energy metric. We establish coordinate invariance, exact reduction to ENGD in Euclidean space, global first-order convergence with Armijo backtracking, and robustness to inexact tangent solves. For quadratic residual energies and generalized Gauss--Newton pullbacks, the Woodbury identity transfers the tangent system to sample space without changing the direction. Nyström approximation provides scalable sample-space solves with controlled direction error and recovers the exact direction after iterative convergence. On the evaluated neural PDE benchmarks, EMNGD achieves higher accuracy and faster convergence than the compared state-of-the-art baselines. Woodbury preserves the EMNGD direction, while scalable-solver diagnostics quantify the accuracy--cost trade-off of preconditioning and residual subsampling.

View free PDFSource page

Related papers

arxivmath.OCcs.LG2026-07-05

Unified convergence analysis for gradient descent optimization methods in the training of deep neural networks

Shokhrukh Ibragimov, Arnulf Jentzen

Gradient based optimization methods are nowadays the methods of choice for training deep neural networks (DNNs) in artificial intelligence (AI) systems. In practically relevant DNN training problems, one does usually not apply the standard gradient descent (GD) optimization metho…

View free PDFSource page
arxivcs.LGcs.MS2026-07-20

FlashPDE: A Drop-In Fused Triton Operator Library for Neural PDE Solvers

Peiyu Zang, Bosen Xie, Ruoxiang Xu, Yongqiang Cai

Physics-Informed Neural Networks (PINNs) solve PDEs by incorporating physical constraints into neural-network training, but large-scale problems are limited by automatic-differentiation memory overhead and inefficient execution of grid-based PDE operators. We present FlashPDE, a…

View free PDFSource page
arxivquant-phcs.LGphysics.chem-phphysics.comp-ph2026-07-21

Enhanced Neural Quantum State via Annealed Gradient Descent

Shiwei Zhou, Yiming Huang, Xiao Yuan, Xiaoxia Cai

Neural quantum states offer expressive representations of quantum many-body wave functions, yet their practical accuracy can be limited by stochastic optimization rather than representational capacity. Here we identify a finite-sample instability, termed subspace trapping, in whi…

View free PDFSource page
arxivmath.OCcs.LGcs.MA2026-07-22

Decentralized Online Riemannian Optimization for Strongly Geodesically Convex Functions

Zhanyuan Cai, Emre Sahinoglu, Shahin Shahrampour

We study decentralized online optimization for strongly geodesically convex (strongly g-convex) losses on Riemannian manifolds with bounded sectional curvature, including positively curved manifolds. In centralized Riemannian optimization, strong g-convexity tightens the optimal…

View free PDFSource page