2 problems
- 0 votes0 replies0 views
Conjecture on the lack of state feedback in affine value function approximation
Let be the set of time periods, let be the set of feasible states, and let be a state strictly inside . In affine value function approximation, the gradient of the a…
- 0 votes0 replies1 view
Feasibility of sampling-based approximate dynamic programming for CVaR MDPs
A CVaR Markov decision process has a Bellman equation that is contracting, and let a sampling-based approximate dynamic programming approach be applied to it. Sampling-based approx…