Menu
Course
How LLMs Actually Work
0/4 modules complete0%
Modules
4 total
Model Creation Lifecycle
ApplyTrace raw data through pretraining and post-training to a deployed assistant checkpoint.
Continue
2How LLMs Evolved
Explain why one API endpoint now handles classification, summarisation, and generation.
Open
3Parallel prefill
Explain Transformers as parallel sequence processors during prefill.
Open
4Quantization: running LLMs cheaper
I can explain how quantization shrinks an LLM's memory and cost by storing weights at lower precision, and what quality it trades away.
Open