Courses
Concepts
Explore
Playground
Practice
Log in
Course contents
← Back to course
Reinforcement Learning
You are in chapter 3
0% of course complete
3
The k-Armed Bandit Problem: Balancing Exploration and Exploitation
Module 1 · Exploration vs Exploitation
→
ARTICLE
Greedy Action Selection
›
Reinforcement Learning
›
The k-Armed Bandit Problem: Balancing Exploration and Exploitation
ARTICLE
Greedy Action Selection
Loading lesson…
← PREVIOUS
The Exploration-Exploitation Trade-off
MARK COMPLETE ✓
NEXT →
Why Action-Value Averages Need Incremental Computation