Learning by trial and error

Some problems have no answer key, only feedback. Balance exploring against exploiting, watch an agent learn to cross a gridworld from rewards alone, and see how the same ideas fine-tune language models and why they can be gamed.

Level
Intermediate
Length
0 lessons, 0 min
Assumes
The machine learning and neural networks tracks.
Builds on
How machines learn and Neural networks

Lessons for this track are being written. In the meantime, try the labs.

Try "embedding", "softmax", "overfitting", or "backpropagation".