13 problems
- 0 votes0 replies0 views
Whittle's asymptotic optimality conjecture for restless multi-armed bandits
A restless multi-armed bandit problem has projects, of which up to can be selected at each time, where . Each project is modeled as a binary-action Markov d…
- 0 votes0 replies0 views
Order-optimality conjecture for RCA-M with rested-bandit index policies
Order-optimality conjecture. The order optimality of RCA-M should hold when it is used with any index policy that is order optimal for the rested bandit problem.
- 0 votes0 replies1 view
PCLI3 threshold and marginal PCL identities for the Kalman-filter bandit
PCLI3 threshold and marginal PCL identities. The functions and admit càdlàg locally bounded-variation extensions to , with nonincreasing and nonconstant. For…
- 0 votes0 replies0 views
PCLI2 completion conjecture for the Kalman-filter bandit
PCLI2 completion conjecture. The function admits a finite-valued, nondecreasing, continuous extension , with as…
- 0 votes0 replies0 views
PCLI1 completion conjecture for the Kalman-filter bandit
PCLI1 completion conjecture. For every and , the limit exists and satisfies . Consequently, the Abelian limit in the prefix-surplus formula…
- 0 votes0 replies1 view
Average-criterion PCL-indexability for restless bandits with imperfect feedback
Let and denote the average-criterion limiting marginal metrics, and define the diagonal MP index by … The average-criterion analogues of and…
- 0 votes0 replies0 views
Global PCL-indexability of discounted single-project problems
Let , , , and be parameters satisfying … A discounted single-project problem is PCL-indexable when the conditions hold on…
- 0 votes0 replies1 view
Uniform-horizon conjecture for heterogeneous restless multi-armed bandits
Uniform-horizon conjecture. The function is independent of and depends only on
- 0 votes0 replies0 views
Whittle's indexability conjecture for asymptotic optimality
The restless bandit problem consists of multiple arms whose states evolve whether or not they are activated, subject to a resource constraint that allows a constant fraction of the…
- 0 votes0 replies0 views
Non-degeneracy as a necessary condition for exponentially fast asymptotic optimality in restless bandits
Non-degeneracy conjecture. Non-degeneracy is also a necessary condition for the existence of a policy satisfying
- 0 votes0 replies0 views
Indexability conjecture for lazy restless bandits
A lazy restless bandit has a belief state and subsidy ; let denote the relevant state-transition parameters, let denote the rewards associated…
- 0 votes0 replies0 views
Constraint-specific explanation for the optimality gap in weakly coupled dynamic programs
Constraint-specific explanation conjecture. The difference between the results obtained for the paper's RMAB constraints and those obtained for the more general WCDP constraints is…
- 0 votes0 replies0 views
Extension of finite-horizon restless-bandit results to multiple actions and total-horizon budgets
Extension conjecture. The formulation of the policy and its asymptotic optimality should extend to (i) multiple actions associated with a state, instead of only an active and a pas…