#fine-tuning
-
Deep Dive into LLMs like ChatGPT
Karpathy's 3.5-hour general-audience deep dive (9.7M views, Feb 2025): the full ChatGPT pipeline โ pretraining (internet โ tokens โ next-token prediction), supervised fine-tuning (conversations โ an assistant persona), and reinforcement learning (practice problems โ emergent reasoning) โ mapped throughout onto how children learn from textbooks: exposition, worked examples, practice problems.