Terra Cognita

Policy Mirror Descent for Regularized Reinforcement Learning: A Generalized Framework with Linear Convergence

arxiv-nonexclusive link only — licence forbids redistribution arxiv 2021

Wenhao Zhan, Shicong Cen, Baihe Huang, Yuxin Chen, Jason D. Lee, Yuejie Chi

arXiv preprint (cs.LG, cs.IT, math.IT).

Source ↗

ABSTRACT · 0.0 MB · sha256 5a19dbd8ad32…

Concepts this teaches

Not yet mapped to any concept.

A resource only becomes useful here once a curator has anchored it to concepts at specific pages. Until then it is a book on a shelf.