New algorithm tackles fairness in multi-armed bandit problem.
problem Fairness in stochastic multi-armed bandit problem.
method Characterized a class of Fair-SMAB algorithms with two parameters.
result Achieves O(log(T)) r-Regret with UCB1 learning algorithm.