1 problem
- 0 votes0 replies2 views
Conjecture on the usefulness of ellipticity-induced Hilbert space structure
The paper studies model-free off-policy reinforcement learning in continuous time with general function approximation. Its analysis uses an ellipticity condition, which induces a H…