RandomBernoulliBandit#
- statrl.settings.bandits.stochastic.anytime.envs.parametric.RandomBernoulliBandit(Delta, K, name='MAB-RandomBernoulli')[source]#
Draw a random Bernoulli instance with a prescribed optimality gap.
Useful for studying how regret scales with the gap: the difficulty of a bandit instance is governed by \(\Delta\), so sweeping it while holding
Kfixed isolates that dependence.- Parameters:
- Returns:
A Bernoulli bandit whose two leading means differ by exactly
Delta.- Return type: