Fetching the paper…

Instance-Dependent Near-Optimal Policy Identification in Linear MDPs via Online Experiment Design · Around