Learning by trial and error
Some problems have no answer key, only feedback. Balance exploring against exploiting, watch an agent learn to cross a gridworld from rewards alone, and see how the same ideas fine-tune language models and why they can be gamed.
- Level
- Intermediate
- Length
- 0 lessons, 0 min
- Assumes
- The machine learning and neural networks tracks.
- Builds on
- How machines learn and Neural networks
Lessons for this track are being written. In the meantime, try the labs.