Skip to content
arXiv stat.ML · Papers

The Greedy Advantage in Finite-Horizon Bandits

arXiv:2607.29375v1 Announce Type: new Abstract: Organizations increasingly rely on sequential experimentation to improve decision-making. While the multi-armed bandit literature has developed algorithms with strong asymptotic regret guarantees, many practical applications operate over finite and externally imposed hori