Pretraining, Then Fine-Tuning
Next-token prediction on unlabelled text to obtain a foundation model, then a small labelled stage that adapts it - and what cross-entropy is really measuring.
AdvancedModule 335 min · 140 XP
This is a premium lesson
Sign in and enrol to read the full lesson, run the code, take the quiz, and earn XP toward the path badge.