Chapter 0.1 - Course Index
Welcome to the complete, step-by-step mathematical breakdown of Large Language Models (Decoder-only Transformer Architecture)! Below is the structured index of all course modules:
Chapter 0.2 - Architecture Ingredients
Below are all the static weight matrices and parameters assumed to be learned during model training. These exact values are referenced and multiplied throughout the interactive visualizations in this course: