4 problems
- 0 votes0 replies1 view
The minimax regret conjecture for stochastic bandit convex optimization
Let be the dimension, let denote the action space, and let be the minimax expected regret after rounds in stochastic bandit convex op…
- 0 votes0 replies0 views
The wider-regime conjecture for minimax lower bounds in off-policy evaluation
Wider-regime conjecture. A similar lower bound should hold in the wider regime .
- 0 votes0 replies0 views
Generalization of the single-trajectory lower-bound construction for ordinary differential equations
Single-trajectory generalization conjecture. Similar arguments should establish the same lower bounds for arbitrary and , wit…
- 0 votes0 replies0 views
The factor- minimax lower-bound conjecture for non-stationary finite-horizon MDPs
The factor- minimax lower-bound conjecture. The minimax lower bound for the non-stationary finite-horizon MDP setting should be a factor of larger than the corresponding low…