Ad
Skip to content

Want to understand ChatGPT? Watch Andrej Karpathy's explanation of how LLMs work

If you want to understand AI, there's one video you need to watch this week. Andrej Karpathy, formerly of OpenAI and Tesla, has released what might be the clearest explanation yet of how Large Language Models actually work. The video breaks down the entire training process of these AI systems and provides mental models for understanding their "psychology" - essentially, how they process and respond to information. Karpathy, who recently co-founded the AI education company Eureka Labs, also includes practical tips for getting the most out of these tools in real-world applications. What makes this explanation special is how Karpathy brings technical depth without sacrificing clarity.

AI News Without the Hype – Curated by Humans

Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.

Read on for the full picture.
Subscribe for hype-free coverage.

  • Full access to every article on THE DECODER
  • No ads
  • Join the comments and community discussions
  • A weekly AI news recap via mail
  • 6x/year: "AI Radar" — deep dives on the AI topics that matter most
  • Daily AI news, always up to date
  • Our full ten-year archive
  • Covered by a team with 10+ years in AI
Subscribe to The Decoder