Reinforcement LearningOff-policy Learning with n-step Tree Backup

ARTICLE

How the n-step Target Fits into Tree Backup

Loading lesson…