EconBase
← All papers

Deep Learning for Dynamic Programming with Recursive Utility

Xianhua Peng, Wu Guo

arXiv 5 Jul 2026 · Finance — Computational

arXiv:2607.04278 · PDF · DOI · OpenAlex · Extracted main text

Abstract

We propose the first deep learning algorithm, the Certainty Equivalent Learning (CEL) algorithm, for solving high-dimensional discrete-time dynamic programming problems with recursive utility. Dynamic programming with recursive utility is numerically challenging because the recursive utility does not have an explicit representation and the Bellman equation contains a certainty equivalent that is difficult to evaluate. The CEL algorithm learns this certainty-equivalent value directly with neural networks and jointly approximates value functions, policy functions, and certainty-equivalent functions. The CEL algorithm is mesh-free and simulation-based, allowing high-dimensional state and control spaces, and does not rely on Euler equations, first-order conditions, or differentiability of the state transition function. The CEL algorithm also works for dynamic programming problems with expected utility as expected utility is a special case of recursive utility. We apply the CEL to discounted linear exponential quadratic Gaussian control, small-noise robust control, Epstein-Zin DSGE, and multivariate strategic asset allocation problems. Compared with closed-form and VFI-based benchmarks, the CEL delivers accurate value and policy approximations, remains effective in high-dimensional problems, achieves accuracy comparable to VFI in the small-noise robust-control case, and produces out-of-sample Bellman errors and Euler or first-order residuals that are in the range from 1.0e-4 to 1.0e-3 for most problems.

Citation extraction

79
references
112
in-text mentions
80
distinct cited
0
self-citations
22,733
main-text words

appendix boundary found by appendix_command · 93% of the source is main text. Read the extracted text to check this.

Most heavily cited references

The works this paper leans on most, across its whole bibliography — not restricted to papers in our corpus. Ranked by composite intensity, which combines how often a work is mentioned, how many sections mention it, and how much of that falls in the main text rather than the appendix.

ReferenceIntensityMentionsSectionsMain text
1Hansen \ Sargent (1995) Discounted linear exponential quadratic gaussian control, IEEE Transactions on Automatic control 40(5): 968–9710.92843100%
2Epstein \ Zin (1989) Substitution, risk aversion, and the temporal behavior of consumption and asset returns: a theoretical framework, Econometrica 5…0.84333100%
3Hansen \ Sargent (2013) Recursive Models of Dynamic Linear Economies, Princeton University Press0.73732100%
4Judd (1998) Numerical Methods in Economics, MIT Press, Cambridge, MA0.73732100%
5Friedl, Kübler, Scheidegger \ Usui (2023) Deep uncertainty quantification: with an application to integrated assessment models, Technical report, Working Paper University…0.69391100%
6Duffie \ Epstein (1992) Stochastic differential utility, Econometrica 60: 353–3940.64422100%
7Weil (1989) The equity premium puzzle and the risk-free rate puzzle, Journal of Monetary Economics 24(3): 401–4210.64422100%
8Weil (1990) Nonexpected utility in macroeconomics, The Quarterly Journal of Economics 105(1): 29–420.64422100%
9Campbell, Chan \ Viceira (2003) A multivariate model of strategic asset allocation, Journal of financial economics 67(1): 41–800.64422100%
10Hansen \ Sargent (2008) Robustness, Princeton university press0.64422100%

Showing the top 10 of 80 scored citations.

Cited by, within the corpus

arXiv econ.EM papers that cite this one, ranked by how heavily they lean on it.

Citing paperIntensityMentionsSections
1Deep Learning for Dynamic Programming with Recursive Utility Using First-order Conditions0.51121