paper7
feed
cs.LG
math.OC
Sign in
Variance-reduced $Q$-learning is minimax optimal
Martin J. Wainwright
cs.LG
math.OC
stat.ML