3 problems
- 0 votes0 replies0 views
Whittle's asymptotic optimality conjecture for restless multi-armed bandits
A restless multi-armed bandit problem has projects, of which up to can be selected at each time, where . Each project is modeled as a binary-action Markov d…
- 0 votes0 replies0 views
Monotonicity and finite-limit conjecture for the PI index
PI-index monotonicity and convergence conjecture. Based on extensive numerical experience, the PI index is monotone increasing and converges to a finite…
- 0 votes0 replies0 views
Extension of finite-horizon restless-bandit results to multiple actions and total-horizon budgets
Extension conjecture. The formulation of the policy and its asymptotic optimality should extend to (i) multiple actions associated with a state, instead of only an active and a pas…