3 problems
- 0 votes0 replies0 views
Bennett-type logarithmic factor conjecture for span-dependent value estimation
Let be the dimension of the tabular state space. In the upper bound for plug-in value-function estimation, a factor multiplies the dependence on the span semi-no…
- 0 votes0 replies0 views
Span-dependent plug-in lower-bound conjecture for policy evaluation
Span-dependent plug-in lower-bound conjecture. There is a Markov reward process for which this bound holds. The claim would show that the corresponding upper bound is sharp up to a…
- 0 votes0 replies0 views
Local minimax optimality of the plug-in and median-of-means approaches
Let , , and denote the classes of Markov reward processes with, r…