AIBussinBuild useful systems with modern AI
  • Home
  • Books
  • Solutions
  • Toolkit
  • Prompts
  • Articles
  • About

Pytorch

  • What Are We Actually Doing?
  • The Tensor: What Is Actually Flowing Through the Loop?
  • Autograd: What Did PyTorch Record, and Where Does the Gradient Stop?
  • The Network: What Is It Without nn.Module?
  • nn.Module: What Does PyTorch Think Belongs to Your Model?
  • DataLoader: Where Is the Training Loop Actually Waiting?
  • Transforms: What Does the Model Actually See?
  • CNN Geometry: What Shape Reaches the Next Layer?
  • Feature Space: What Does a Linear Model Actually See?
  • Attention: Which Position Is Comparing With Which?
  • Training: Which Link in the Learning Chain Is Broken?
  • Performance: What Is the Machine Waiting For?
  • Compilation: Which Assumption Stopped Holding?
  • Regressions: Did the Model Change, or the Measurement?
  • Assembly: A Language Model You Can Interrogate
  • Appendix A: PyTorch Diagnostic Field Guide
  • Make the Automaton Differentiable
  • Learn the Local Update Rule
  • Hidden Cell Channels and Local Memory
  • Grow a Target From One Seed
  • Randomize the Update Schedule
  • Train for Persistence
  • Regenerate After Damage
  • Test Generalization Beyond Training
  • Neural Cellular Automata for Pathfinding
  • Learn to Solve Mazes
  • Generalize to Harder and Larger Mazes
  • Inspect Hidden-State Propagation
  • What Did the Neural CA Actually Learn?
  • Run Cellular Automata on the GPU
  • Preference Rankers — Learning Which Answer Is Better
  • Which Model Should You Use? MR.Q, EBT, SICQL, HRM, Tiny and PACS Compared
  • PACS — Building an Optimizer From Gradient Statistics
  • Inside Tiny — Residual Blocks, Attention and Sparse Autoencoders
  • Tiny — Recursive Reasoning With a Small Neural Network
  • HRM — Hierarchical Reasoning With Fast and Slow Recurrent State
  • SICQL — Building a Model From Q, V and Policy Networks
  • EBT — From One Score to Q, V, Policy and Advantage
  • MR.Q — Building a Neural Quality Model From Two Embeddings
  • The Model Inside the Model
  • PyTorch Zero to Hero 10: Build a Small GPT-Style Language Model From Scratch
  • PyTorch Performance Debugging: CUDA OOM, Slow Training, GPU Utilization and torch.compile
  • PyTorch Model Not Learning? A Systematic Debugging Guide
  • PyTorch Attention Shapes: Q, K, V, Multi-Head Attention Masks and Transformer Dimension Errors
  • PyTorch CNN Shape Errors: Conv2d Output Sizes, Channels, Flatten Bugs and How to Debug Them
  • PyTorch DataLoader Performance: num_workers, pin_memory, Prefetching and Why Your GPU Is Waiting
  • PyTorch nn.Module Explained: Missing Parameters, state_dict, Buffers and Registration Bugs
  • Build a Neural Network From Scratch in PyTorch Without nn.Module
  • PyTorch Autograd Debugging: requires_grad, detach, backward() and NaN Gradients
  • PyTorch Tensor Shapes: Broadcasting, Reshape, View, Permute and the Errors That Waste Your Time
  • PyTorch Zero to Hero 00: What Are We Actually Doing?
  • MR.Q: Model-Based Representations for Model-Free Trading
  • Writing Neural Networks with PyTorch
AIBussin

Build useful systems with modern AI.

Books, solutions, experiments and working tools for practical AI.

© 2026 Ernan Hughes
  • Books
  • Solutions
  • Toolkit
  • Prompts
  • Articles
  • About
  • Programmer.ie
  • ZeroModel.org
  • GitHub