Fetching the paper…

Learning Linear-Quadratic Regulators Efficiently with only $\sqrt{T}$ Regret · Around