Reinforcement LearningREINFORCE with Baseline: Reducing Variance in Policy Gradient Methods

ARTICLE

Understanding the Policy Gradient Theorem with a Baseline

Loading lesson…