Courses
Concepts
Explore
Playground
Practice
Log in
Course contents
← Back to course
Reinforcement Learning
You are in chapter 75
0% of course complete
75
Off-policy Learning with n-step Tree Backup
Module 1 · Mathematical Formulation of n-step Tree Backup
→
ARTICLE
How Expected Action Value and TD Error Build Tree-Backup Returns
›
Reinforcement Learning
›
Off-policy Learning with n-step Tree Backup
ARTICLE
How Expected Action Value and TD Error Build Tree-Backup Returns
Loading lesson…
← PREVIOUS
How n-step Tree Backup Weights Branches
MARK COMPLETE ✓
NEXT →
How the n-step Target Fits into Tree Backup