5 modules · 40 lessons · 50 hours

All tracks

Advanced · every module free · 5 challenges

Tokenization & Embeddings
8h · 7 lessons

Byte-pair encoding, vocabulary design, and what an embedding space actually contains.

Attention & the Transformer
14h · 10 lessons

Self-attention, multi-head attention and a complete transformer implemented from the paper, with every shape written out.

Training Transformers at Scale
10h · 8 lessons

Data pipelines, curricula, stability tricks and the failure modes of long training runs.

Fine-tuning & Parameter-Efficient Methods
10h · 8 lessons

Full fine-tuning, LoRA, QLoRA and adapters — adapting a pretrained model on a budget you actually have.

NLP Tasks & Production Pipelines
8h · 7 lessons

Classification, extraction, summarization and the evaluation each one needs.

The other tracks

Level
Advanced
Size
7 modules · 62h
Level
Advanced
Size
7 modules · 62h
Level
All levels
Size
5 modules · 42h
Level
Beginner
Size
5 modules · 52h
Level
Beginner
Size
5 modules · 46h
Level
Intermediate
Size
7 modules · 64h
Level
Intermediate
Size
6 modules · 60h