Skip to yearly menu bar Skip to main content


Non-Asymptotic Best Policy Identification Guarantees in Online Reinforcement Learning

Joseph Lazzaro ⋅ Alessio Russo ⋅ Aldo Pacchiano

Abstract

Chat is not available.