首页 > AI前沿 > Bellman Error Minimization Via Linear Programming Normalization

Bellman Error Minimization Via Linear Programming Normalization

arXiv机器学习 2026-10-02 11:05 6 阅读 查看原文

This paper proposes a new functional approximation approach to reduce Bellman error in high-dimensional dynamic programming and Reinforcement Learning problems.

Using a classic dynamic programming problem (network capacity control in revenue management) as the motivational example, the paper illustrates that deep neural networks and linear programming approximation algorithms can be combined to derive approximate solutions to dynamic programming problems.

Simulation results show the proposed approximation algorithms achieves competitive performance when compared with benchmark.