首页 正文

Generalized Policy Improvement Algorithms with Theoretically Supported Sample Reuse

{{output}}
We develop a new class of model-free deep reinforcement learning algorithms for data-driven, learning-based control. Our Generalized Policy Improvement algorithms combine the policy improvement guarantees of on-policy methods with the efficiency of sample reus... ...