Inside the model

Open up a real language model running in your browser. Read its predictions layer by layer, switch off attention heads to see what they do, watch an induction circuit form during training, and steer the model's output. Then see why understanding and aligning models is hard.

Level
Advanced
Length
0 lessons, 0 min
Assumes
The transformers track.
Builds on
Transformers and LLMs

Lessons for this track are being written. In the meantime, try the labs.

Try "embedding", "softmax", "overfitting", or "backpropagation".