CORTEXA
← Browse
arxivq-fin.CPecon.EMmath.OCstat.ML2026-07-10

Deep Learning for Dynamic Programming with Recursive Utility Using First-order Conditions

Xianhua Peng, Wu Guo, Songyan Wang, Jianfei Zhu

This paper proposes the certainty-equivalent first-order learning (CEFOL) algorithm, a deep learning algorithm for solving discrete-time dynamic programming problems with recursive utility. Dynamic programming with recursive utility is challenging because nonlinear certainty equivalent appears in the Bellman equation and the first-order optimality conditions but is difficult to evaluate. By introducing a separate neural network to represent the certainty equivalent, CEFOL enables the exploitation of the Bellman and model-specific first-order optimality conditions. In addition to certainty equivalent, CEFOL also uses neural networks to learn the value functions, policy functions, and Lagrange multipliers by using model-specific first-order conditions to construct residuals for minimization. By using first-order and KKT residuals to learn the policy, CEFOL directly accommodates general equality and inequality constraints on the controls, including occasionally binding constraints, without requiring penalty functions or problem-specific reformulations. We apply the algorithm to risk-sensitive and Epstein--Zin consumption-saving problems, a small-noise robust-control problem, and a DSGE model with recursive preferences and stochastic volatility. Across these applications, out-of-sample Bellman diagnostics and model-specific optimality residuals, including Euler or first-order residuals where applicable, are generally of order 1.0e-4 to 1.0e-3 over the relevant state regions, with larger values mainly near binding constraints, and the learned value and policy functions closely match VFI benchmarks when available. The CEFOL algorithm also works for dynamic programming problems with expected utility, as expected utility is a special case of recursive utility.

View free PDFSource page

Related papers

arxivecon.EMcs.LGq-fin.CPq-fin.RMstat.ML2026-06-27

Liquidity-Based Audit of Algorithmic Trading Strategies

Irene Aldridge

We show that net demand for liquidity by algo strategies is identifiable from its trade and price history alone, with no knowledge of its signal or optimization problem. An exact multi-period regret decomposition implies that the sign of this statistic classifies a linear strateg…

View free PDFSource page
arxivmath.OCcs.LGstat.ML2026-07-09

Nonconvex Composite Functional Constraints via First-Order Augmented Lagrangian Methods under Local Regularity

Linglingzhi Zhu, Jiajin Li

We study nonasymptotic convergence of primal-dual methods for a class of nonconvex constrained optimization problems with a convex-composite structure. In this class, both the objective and the functional inequality constraints are given by convex Lipschitz outer functions compos…

View free PDFSource page
arxivcs.LGmath.OCstat.ML2026-07-04

A Structural Interpretation of GELU and Threshold-Transmission Activations via the First-Order Loss Function

Roberto Rossi

The Gaussian Error Linear Unit is usually motivated as the expected output of an input-dependent Bernoulli gate. This work gives an alternative interpretation: GELU is the expected output of a hard linear gate with a Gaussian random threshold. This view provides a generative inte…

View free PDFSource page