Start Here

Find the most relevant content based on what you're working on.

Curated Reading Paths

Reinforcement Learning for LLMs

From RL foundations through PPO, GRPO, and GDPO to the GRPO family map — the complete policy optimization series.

View series landing page →

Recommendation Systems

From contextual bandits for personalization to retrieval-augmented generation.

LLM Evaluation

Moving beyond vibes to systematic evaluation.

Foundations

Reference explainers covering the building blocks of modern ML. Start with the Transformer Internals series — written alongside RLVR from Scratch, where every component is implemented and tested from raw tensors.

View all foundations →