Friday, April 18, 2025

How LLMs Work, by Andrej Karpathy

Andrej Karpathy, Eureka Labs founder and computer scientist (Tesla, OpenAI), explains how language models work, and are built. You'll need about 3.5 hours to view the whole video, but it covers transformer networks, training (human and algorithmic; labeling) and reinforcement learning. 

No comments:

Does Generative AI Use Stunt Cognitive Skill Development in Children?

We might still not know whether using generative artificial language models has any negative effect on cognitive skills, but Norway believes...