arXiv stat.ML
· Papers
The Greedy Advantage in Finite-Horizon Bandits
arXiv:2607.29375v1 Announce Type: new Abstract: Organizations increasingly rely on sequential experimentation to improve decision-making. While the multi-armed bandit literature has developed algorithms with strong asymptotic regret guarantees, many practical applications operate over finite and externally imposed hori