Reinforcement LearningREINFORCE with Baseline: Reducing Variance in Policy Gradient Methods

ARTICLE

How a Baseline Changes the REINFORCE Update

Loading lesson…