Courses
Concepts
Explore
Playground
Practice
Log in
Course contents
← Back to course
Reinforcement Learning
You are in chapter 8
0% of course complete
8
Gradient Bandit Algorithms for Action Selection
Module 0 · Introduction to Gradient Bandit Algorithms
→
ARTICLE
How Numerical Preferences Guide Action Selection
›
Reinforcement Learning
›
Gradient Bandit Algorithms for Action Selection
ARTICLE
How Numerical Preferences Guide Action Selection
Loading lesson…
← PREVIOUS
Where UCB Action Selection Becomes Difficult
MARK COMPLETE ✓
NEXT →
How Stochastic Gradient Ascent Adjusts Action Preferences