IDEAS home Printed from https://ideas.repec.org/p/arx/papers/2607.04278.html

Deep Learning for Dynamic Programming with Recursive Utility

Author

Listed:
  • Xianhua Peng
  • Wu Guo

Abstract

We propose the first deep learning algorithm, the Certainty Equivalent Learning (CEL) algorithm, for solving high-dimensional discrete-time dynamic programming problems with recursive utility. Dynamic programming with recursive utility is numerically challenging because the recursive utility does not have an explicit representation and the Bellman equation contains a certainty equivalent that is difficult to evaluate. The CEL algorithm learns this certainty-equivalent value directly with neural networks and jointly approximates value functions, policy functions, and certainty-equivalent functions. The CEL algorithm is mesh-free and simulation-based, allowing high-dimensional state and control spaces, and does not rely on Euler equations, first-order conditions, or differentiability of the state transition function. The CEL algorithm also works for dynamic programming problems with expected utility as expected utility is a special case of recursive utility. We apply the CEL to discounted linear exponential quadratic Gaussian control, small-noise robust control, Epstein-Zin DSGE, and multivariate strategic asset allocation problems. Compared with closed-form and VFI-based benchmarks, the CEL delivers accurate value and policy approximations, remains effective in high-dimensional problems, achieves accuracy comparable to VFI in the small-noise robust-control case, and produces out-of-sample Bellman errors and Euler or first-order residuals that are in the range from 1.0e-4 to 1.0e-3 for most problems.

Suggested Citation

  • Xianhua Peng & Wu Guo, 2026. "Deep Learning for Dynamic Programming with Recursive Utility," Papers 2607.04278, arXiv.org.
  • Handle: RePEc:arx:papers:2607.04278
    as

    Download full text from publisher

    File URL: https://arxiv.org/pdf/2607.04278
    File Function: Latest version
    Download Restriction: no
    ---><---

    References listed on IDEAS

    as
    1. Ma, Qingyin & Stachurski, John & Toda, Alexis Akira, 2022. "Unbounded dynamic programming via the Q-transform," Journal of Mathematical Economics, Elsevier, vol. 100(C).
    2. Jiequn Han & Yucheng Yang & Weinan E, 2026. "DeepHAM: A global solution method for heterogeneous agent models with aggregate shocks," Quantitative Economics, Econometric Society, vol. 17(2), pages 297-341, May.
    3. Skavysh, Vladimir & Priazhkina, Sofia & Guala, Diego & Bromley, Thomas R., 2023. "Quantum monte carlo for economics: Stress testing and macroeconomic deep learning," Journal of Economic Dynamics and Control, Elsevier, vol. 153(C).
    4. Duffie, Darrel & Lions, Pierre-Louis, 1992. "PDE solutions of stochastic differential utility," Journal of Mathematical Economics, Elsevier, vol. 21(6), pages 577-606.
    5. Marlon Azinovic & Jan v{Z}emliv{c}ka, 2023. "Economics-Inspired Neural Networks with Stabilizing Homotopies," Papers 2303.14802, arXiv.org.
    6. Vytautas Valaitis & Alessandro T. Villa, 2024. "A machine learning projection method for macro‐finance models," Quantitative Economics, Econometric Society, vol. 15(1), pages 145-173, January.
    7. Yongyang Cai & Kenneth L. Judd & Thomas S. Lontzek, 2013. "The Social Cost of Stochastic and Irreversible Climate Change," NBER Working Papers 18704, National Bureau of Economic Research, Inc.
    8. John Y. Campbell & Yeung Lewis Chanb & M. Viceira, 2013. "A multivariate model of strategic asset allocation," World Scientific Book Chapters, in: Leonard C MacLean & William T Ziemba (ed.), HANDBOOK OF THE FUNDAMENTALS OF FINANCIAL DECISION MAKING Part II, chapter 39, pages 809-848, World Scientific Publishing Co. Pte. Ltd..
    9. Coeurdacier, Nicolas & Rey, Hélène & Winant, Pablo, 2020. "Financial integration and growth in a risky world," Journal of Monetary Economics, Elsevier, vol. 112(C), pages 1-21.
    10. Lars Peter Hansen & Thomas J. Sargent, 2013. "Recursive Models of Dynamic Linear Economies," Economics Books, Princeton University Press, edition 1, number 10141, December.
    11. Emmet Hall-Hoffarth, 2023. "Non-linear approximations of DSGE models with neural-networks and hard-constraints," Papers 2310.13436, arXiv.org.
    12. Weil, Philippe, 1989. "The equity premium puzzle and the risk-free rate puzzle," Journal of Monetary Economics, Elsevier, vol. 24(3), pages 401-421, November.
    13. TallariniJr., Thomas D., 2000. "Risk-sensitive real business cycles," Journal of Monetary Economics, Elsevier, vol. 45(3), pages 507-532, June.
    14. Victor Duarte & Diogo Duarte & Dejanir H Silva, 2024. "Machine Learning for Continuous-Time Finance," The Review of Financial Studies, Society for Financial Studies, vol. 37(11), pages 3217-3271.
    15. Victor Duarte & Julia Fonseca & Aaron S. Goodman & Jonathan A. Parker, 2021. "Simple Allocation Rules and Optimal Portfolio Choice Over the Lifecycle," NBER Working Papers 29559, National Bureau of Economic Research, Inc.
    16. Amine Mohamed Aboussalah & Ziyun Xu & Chi-Guhn Lee, 2022. "What is the value of the cross-sectional approach to deep reinforcement learning?," Quantitative Finance, Taylor & Francis Journals, vol. 22(6), pages 1091-1111, June.
    17. Walter Pohl & Karl Schmedders & Ole Wilms, 2024. "Existence of the Wealth-Consumption Ratio in Asset Pricing Models with Recursive Preferences," The Review of Financial Studies, Society for Financial Studies, vol. 37(3), pages 989-1028.
    18. Pascal, Julien, 2024. "Artificial neural networks to solve dynamic programming problems: A bias-corrected Monte Carlo operator," Journal of Economic Dynamics and Control, Elsevier, vol. 162(C).
    19. Achref Bachouch & Côme Huré & Nicolas Langrené & Huyên Pham, 2022. "Deep Neural Networks Algorithms for Stochastic Control Problems on Finite Horizon: Numerical Applications," Methodology and Computing in Applied Probability, Springer, vol. 24(1), pages 143-178, March.
    20. Kenneth L. Judd, 1998. "Numerical Methods in Economics," MIT Press Books, The MIT Press, edition 1, volume 1, number 0262100711, December.
    21. Qingyin Ma & John Stachurski, 2021. "Dynamic Programming Deconstructed: Transformations of the Bellman Equation and Computational Efficiency," Operations Research, INFORMS, vol. 69(5), pages 1591-1607, September.
    22. Marlon Azinovic & Luca Gaegauf & Simon Scheidegger, 2022. "Deep Equilibrium Nets," International Economic Review, Department of Economics, University of Pennsylvania and Osaka University Institute of Social and Economic Research Association, vol. 63(4), pages 1471-1525, November.
    23. Anis Matoussi & Hao Xing, 2018. "Convex duality for Epstein–Zin stochastic differential utility," Mathematical Finance, Wiley Blackwell, vol. 28(4), pages 991-1019, October.
    24. Christensen, Timothy M., 2022. "Existence and uniqueness of recursive utilities without boundedness," Journal of Economic Theory, Elsevier, vol. 200(C).
    25. Yifan Zhao & Arnab Basu & Thomas S. Lontzek & Karl Schmedders, 2023. "The Social Cost of Carbon When We Wish for Full-Path Robustness," Management Science, INFORMS, vol. 69(12), pages 7585-7606, December.
    26. Holger Kraft & Thomas Seiferling & Frank Thomas Seifried, 2017. "Optimal consumption and investment with Epstein–Zin recursive utility," Finance and Stochastics, Springer, vol. 21(1), pages 187-226, January.
    27. Duffie, Darrell & Epstein, Larry G, 1992. "Stochastic Differential Utility," Econometrica, Econometric Society, vol. 60(2), pages 353-394, March.
    28. Holger Kraft & Frank Seifried & Mogens Steffensen, 2013. "Consumption-portfolio optimization with recursive utility in incomplete markets," Finance and Stochastics, Springer, vol. 17(1), pages 161-196, January.
    29. Hao Xing, 2017. "Consumption–investment optimization with Epstein–Zin utility in incomplete markets," Finance and Stochastics, Springer, vol. 21(1), pages 227-262, January.
    30. Anderson, Evan W. & Hansen, Lars Peter & Sargent, Thomas J., 2012. "Small noise methods for risk-sensitive/robust economies," Journal of Economic Dynamics and Control, Elsevier, vol. 36(4), pages 468-500.
    31. Manuel S. Santos, 2000. "Accuracy of Numerical Solutions using the Euler Equation Residuals," Econometrica, Econometric Society, vol. 68(6), pages 1377-1402, November.
    32. Ji Huang, 2023. "A Probabilistic Solution to High-Dimensional Continuous-Time Macro and Finance Models," CESifo Working Paper Series 10600, CESifo.
    33. Victor Duarte & Diogo Duarte & Dejanir H. Silva, 2024. "Machine Learning for Continuous-Time Finance," CESifo Working Paper Series 10909, CESifo.
    34. Dan Cao & Wenlan Luo & Guangyu Nie, 2023. "Uncovering the Effects of the Zero Lower Bound with an Endogenous Financial Wedge," American Economic Journal: Macroeconomics, American Economic Association, vol. 15(1), pages 135-172, January.
    35. John Y. Campbell & Luis M. Viceira, 1999. "Consumption and Portfolio Decisions when Expected Returns are Time Varying," The Quarterly Journal of Economics, President and Fellows of Harvard College, vol. 114(2), pages 433-495.
    36. repec:bla:jfinan:v:59:y:2004:i:4:p:1481-1509 is not listed on IDEAS
    37. Philippe Weil, 1990. "Nonexpected Utility in Macroeconomics," The Quarterly Journal of Economics, President and Fellows of Harvard College, vol. 105(1), pages 29-42.
    38. Greg Kaplan & Giovanni L. Violante, 2014. "A Model of the Consumption Response to Fiscal Stimulus Payments," Econometrica, Econometric Society, vol. 82(4), pages 1199-1239, July.
    39. Kreps, David M & Porteus, Evan L, 1978. "Temporal Resolution of Uncertainty and Dynamic Choice Theory," Econometrica, Econometric Society, vol. 46(1), pages 185-200, January.
    40. Stachurski, John & Wilms, Ole & Zhang, Junnan, 2024. "Asset pricing with time preference shocks: Existence and uniqueness," Journal of Economic Theory, Elsevier, vol. 216(C).
    41. Stachurski, John & Zhang, Junnan, 2021. "Dynamic programming with state-dependent discounting," Journal of Economic Theory, Elsevier, vol. 192(C).
    42. Martin Andreasen, 2012. "On the Effects of Rare Disasters and Uncertainty Shocks for Risk Premia in Non-Linear DSGE Models," Review of Economic Dynamics, Elsevier for the Society for Economic Dynamics, vol. 15(3), pages 295-316, July.
    43. Duffie, Darrell & Epstein, Larry G, 1992. "Asset Pricing with Stochastic Differential Utility," The Review of Financial Studies, Society for Financial Studies, vol. 5(3), pages 411-436.
    44. repec:spo:wpmain:info:hdl:2441/8686 is not listed on IDEAS
    45. Simon Scheidegger, 2026. "Deep Learning for Solving and Estimating Dynamic Models in Economics and Finance," Papers 2605.14493, arXiv.org.
    46. Stachurski, John & Wilms, Ole & Zhang, Junnan, 2024. "Asset pricing with time preference shocks: Existence and uniqueness," Other publications TiSEM 29da00af-3cca-4717-aa55-a, Tilburg University, School of Economics and Management.
    47. Schroder, Mark & Skiadas, Costis, 1999. "Optimal Consumption and Portfolio Selection with Stochastic Differential Utility," Journal of Economic Theory, Elsevier, vol. 89(1), pages 68-126, November.
    48. Maliar, Lilia & Maliar, Serguei & Winant, Pablo, 2021. "Deep learning for solving dynamic economic models," Journal of Monetary Economics, Elsevier, vol. 122(C), pages 76-101.
    49. Yongyang Cai & Kenneth L. Judd & Rong Xu, 2013. "Numerical Solution of Dynamic Portfolio Optimization with Transaction Costs," NBER Working Papers 18709, National Bureau of Economic Research, Inc.
    50. Carroll, Christopher D., 2006. "The method of endogenous gridpoints for solving dynamic stochastic optimization problems," Economics Letters, Elsevier, vol. 91(3), pages 312-320, June.
    51. Lorenzo Garlappi & Georgios Skoulakis, 2010. "Solving Consumption and Portfolio Choice Problems: The State Variable Decomposition Method," The Review of Financial Studies, Society for Financial Studies, vol. 23(9), pages 3346-3400.
    52. Matoussi, Anis & Xing, Hao, 2018. "Convex duality for Epstein-Zin stochastic differential utility," LSE Research Online Documents on Economics 82519, London School of Economics and Political Science, LSE Library.
    53. A. Max Reppen & H. Mete Soner & Valentin Tissot-Daguette, 2023. "Deep stochastic optimization in finance," Digital Finance, Springer, vol. 5(1), pages 91-111, March.
    Full references (including those not matched with items on IDEAS)

    Citations

    Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.
    as


    Cited by:

    1. Xianhua Peng & Wu Guo & Songyan Wang & Jianfei Zhu, 2026. "Deep Learning for Dynamic Programming with Recursive Utility Using First-order Conditions," Papers 2607.09461, arXiv.org.

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Xianhua Peng & Wu Guo & Songyan Wang & Jianfei Zhu, 2026. "Deep Learning for Dynamic Programming with Recursive Utility Using First-order Conditions," Papers 2607.09461, arXiv.org.
    2. Dirk Becherer & Wilfried Kuissi-Kamdem & Olivier Menoukeu-Pamen, 2023. "Optimal consumption with labor income and borrowing constraints for recursive preferences," Working Papers hal-04017143, HAL.
    3. Dejian Tian & Weidong Tian & Jianjun Zhou & Zimu Zhu, 2025. "Optimal Consumption-Investment with Epstein-Zin Utility under Leverage Constraint," Papers 2509.21929, arXiv.org, revised Oct 2025.
    4. Li, Hanwu & Riedel, Frank & Yang, Shuzhen, 2024. "Optimal consumption for recursive preferences with local substitution — the case of certainty," Journal of Mathematical Economics, Elsevier, vol. 110(C).
    5. Zixin Feng & Dejian Tian, 2021. "Optimal consumption and portfolio selection with Epstein-Zin utility under general constraints," Papers 2111.09032, arXiv.org, revised May 2023.
    6. Joshua Aurand & Yu-Jui Huang, 2019. "Epstein-Zin Utility Maximization on a Random Horizon," Papers 1903.08782, arXiv.org, revised May 2023.
    7. Joshua Aurand & Yu‐Jui Huang, 2023. "Epstein‐Zin utility maximization on a random horizon," Mathematical Finance, Wiley Blackwell, vol. 33(4), pages 1370-1411, October.
    8. Feng, Zixin & Tian, Dejian & Zheng, Harry, 2026. "Consumption–investment optimization with Epstein–Zin utility in unbounded non-Markovian markets," Stochastic Processes and their Applications, Elsevier, vol. 192(C).
    9. Xianhua Peng & Steven Kou & Lekang Zhang, 2024. "A Machine Learning Algorithm for Finite-Horizon Stochastic Control Problems in Economics," Papers 2411.08668, arXiv.org, revised Dec 2024.
    10. Yifan Zhao & Arnab Basu & Thomas S. Lontzek & Karl Schmedders, 2023. "The Social Cost of Carbon When We Wish for Full-Path Robustness," Management Science, INFORMS, vol. 69(12), pages 7585-7606, December.
    11. Campani, Carlos Heitor & Garcia, René, 2019. "Approximate analytical solutions for consumption/investment problems under recursive utility and finite horizon," The North American Journal of Economics and Finance, Elsevier, vol. 48(C), pages 364-384.
    12. Chen, Xingjiang & Ruan, Xinfeng & Zhang, Wenjun, 2021. "Dynamic portfolio choice and information trading with recursive utility," Economic Modelling, Elsevier, vol. 98(C), pages 154-167.
    13. Dariusz Zawisza, 2020. "On the parabolic equation for portfolio problems," Papers 2003.13317, arXiv.org, revised Oct 2020.
    14. Anis Matoussi & Hao Xing, 2016. "Convex duality for stochastic differential utility," Papers 1601.03562, arXiv.org.
    15. Guanxing Fu & Ulrich Horst, 2025. "Mean Field Portfolio Games with Epstein-Zin Preferences," Papers 2505.07231, arXiv.org, revised Jun 2026.
    16. Erhan Bayraktar & Emmet Lawless, 2026. "Infinite Horizon Optimal Consumption: Intertemporal Hedging under Epstein-Zin Preferences," Papers 2606.02945, arXiv.org, revised Aug 2026.
    17. Shigeta, Yuki, 2020. "Gain/loss asymmetric stochastic differential utility," Journal of Economic Dynamics and Control, Elsevier, vol. 118(C).
    18. Immacolata Oliva & Ilaria Stefani, 2023. "Co-jumps and recursive preferences in portfolio choices," Annals of Finance, Springer, vol. 19(3), pages 291-324, September.
    19. Kraft, Holger & Seifried, Frank Thomas, 2014. "Stochastic differential utility as the continuous-time limit of recursive utility," Journal of Economic Theory, Elsevier, vol. 151(C), pages 528-550.
    20. Kraft, Holger & Munk, Claus & Weiss, Farina, 2022. "Bequest motives in consumption-portfolio decisions with recursive utility," Journal of Banking & Finance, Elsevier, vol. 138(C).

    More about this item

    Statistics

    Access and download statistics

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:arx:papers:2607.04278. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: arXiv administrators (email available below). General contact details of provider: https://arxiv.org/ .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.