1 problem
- 0 votes0 replies0 views
Extension of constant-stepsize Q-learning results to linear function approximation
Constant-stepsize asynchronous Q-learning produces iterates whose distributional convergence, convergence-rate characterization, central limit theorem for averaged iterates, asympt…