Courses
Concepts
Explore
Playground
Practice
Log in
Course contents
← Back to course
Reinforcement Learning
You are in chapter 85
0% of course complete
85
Unifying n-step Action-Value Backups with Q(σ)
Module 1 · Off-policy n-step Q(σ) Algorithm
→
ARTICLE
Estimating Q-values with Off-policy n-step Q(σ)
›
Reinforcement Learning
›
Unifying n-step Action-Value Backups with Q(σ)
ARTICLE
Estimating Q-values with Off-policy n-step Q(σ)
Loading lesson…
← PREVIOUS
n-step Q(σ) Basics Quiz
MARK COMPLETE ✓
NEXT →
How ε-greedy Policies Guide Off-policy n-step Q(σ)