paper7feed
cs.LGmath.OC

Variance-reduced $Q$-learning is minimax optimal

Martin J. Wainwright

cs.LGmath.OCstat.ML

Annotations (0)

No annotations yet.

Select any passage to add the first one.

paper7papersfollowingopen source
GitHub