Friday, April 18, 2025

How LLMs Work, by Andrej Karpathy

Andrej Karpathy, Eureka Labs founder and computer scientist (Tesla, OpenAI), explains how language models work, and are built. You'll need about 3.5 hours to view the whole video, but it covers transformer networks, training (human and algorithmic; labeling) and reinforcement learning. 

No comments:

Lots of AI Regulations are Conceivable; Few Will Address Existential Threats

One problem with calls for “regulating artificial intelligence” is that it is not entirely clear what should be done, especially on the core...