- What Are We Actually Doing?
- The Tensor: What Is Actually Flowing Through the Loop?
- Autograd: What Did PyTorch Record, and Where Does the Gradient Stop?
- The Network: What Is It Without nn.Module?
- nn.Module: What Does PyTorch Think Belongs to Your Model?
- DataLoader: Where Is the Training Loop Actually Waiting?
- Transforms: What Does the Model Actually See?
- CNN Geometry: What Shape Reaches the Next Layer?
- Feature Space: What Does a Linear Model Actually See?
- Attention: Which Position Is Comparing With Which?
- Training: Which Link in the Learning Chain Is Broken?
- Performance: What Is the Machine Waiting For?
- Compilation: Which Assumption Stopped Holding?
- Regressions: Did the Model Change, or the Measurement?
- Assembly: A Language Model You Can Interrogate
- Appendix A: PyTorch Diagnostic Field Guide
- Make the Automaton Differentiable
- Learn the Local Update Rule
- Hidden Cell Channels and Local Memory
- Grow a Target From One Seed
- Randomize the Update Schedule
- Train for Persistence
- Regenerate After Damage
- Test Generalization Beyond Training
- Neural Cellular Automata for Pathfinding
- Learn to Solve Mazes
- Generalize to Harder and Larger Mazes
- Inspect Hidden-State Propagation
- What Did the Neural CA Actually Learn?
- Run Cellular Automata on the GPU
- Preference Rankers — Learning Which Answer Is Better
- Which Model Should You Use? MR.Q, EBT, SICQL, HRM, Tiny and PACS Compared
- PACS — Building an Optimizer From Gradient Statistics
- Inside Tiny — Residual Blocks, Attention and Sparse Autoencoders
- Tiny — Recursive Reasoning With a Small Neural Network
- HRM — Hierarchical Reasoning With Fast and Slow Recurrent State
- SICQL — Building a Model From Q, V and Policy Networks
- EBT — From One Score to Q, V, Policy and Advantage
- MR.Q — Building a Neural Quality Model From Two Embeddings
- The Model Inside the Model
- PyTorch Zero to Hero 10: Build a Small GPT-Style Language Model From Scratch
- PyTorch Performance Debugging: CUDA OOM, Slow Training, GPU Utilization and torch.compile
- PyTorch Model Not Learning? A Systematic Debugging Guide
- PyTorch Attention Shapes: Q, K, V, Multi-Head Attention Masks and Transformer Dimension Errors
- PyTorch CNN Shape Errors: Conv2d Output Sizes, Channels, Flatten Bugs and How to Debug Them
- PyTorch DataLoader Performance: num_workers, pin_memory, Prefetching and Why Your GPU Is Waiting
- PyTorch nn.Module Explained: Missing Parameters, state_dict, Buffers and Registration Bugs
- Build a Neural Network From Scratch in PyTorch Without nn.Module
- PyTorch Autograd Debugging: requires_grad, detach, backward() and NaN Gradients
- PyTorch Tensor Shapes: Broadcasting, Reshape, View, Permute and the Errors That Waste Your Time
- PyTorch Zero to Hero 00: What Are We Actually Doing?
- MR.Q: Model-Based Representations for Model-Free Trading
- Writing Neural Networks with PyTorch