Description
In this three-and-a-half-hour lecture, Andrej Karpathy—former Tesla AI lead and OpenAI founding member—walks a general audience through the inner workings of large language models like ChatGPT. He delves into the full training stack: from preprocessing massive internet text corpora and tokenization to model structure, inference dynamics, fine-tuning techniques, hallucinations, reinforcement learning, memory systems, and model alignment strategies. With clarity and insight, Karpathy shares mental models to understand how LLMs think, their limitations, and how to use them most effectively—making complex AI concepts accessible and actionable.