LLM101n is an educational repository that walks you through building and understanding large language models from first principles. It emphasizes intuition and hands-on implementation, guiding you from tokenization and embeddings to attention, transformer blocks, and sampling. The materials favor compact, readable code and incremental steps, so learners can verify each concept before moving on. You’ll see how data pipelines, batching, masking, and positional encodings fit together to train a small GPT-style model end to end. The repo often complements explanations with runnable notebooks or scripts, encouraging experimentation and modification. By the end, the focus is less on polishing a production system and more on internalizing how LLM components interact to produce coherent text.

Features

  • Step-by-step build of a GPT-style transformer from scratch
  • Clear coverage of tokenization, embeddings, attention, and MLP blocks
  • Runnable code and exercises for experiential learning
  • Demonstrations of batching, masking, and positional encodings
  • Training and sampling loops you can inspect and modify
  • Emphasis on readability and conceptual understanding over framework magic

Project Samples

Project Activity

See All Activity >

Categories

Education

Follow LLM101n

LLM101n Web Site

Other Useful Business Software
MongoDB Atlas runs apps anywhere Icon
MongoDB Atlas runs apps anywhere

Deploy in 115+ regions with the modern database for every enterprise.

MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
Start Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of LLM101n!

Additional Project Details

Registered

2025-10-15