← Deep Learning Essentials
Free
Neural Nets & Backprop Intuition
Layers, activations, and how gradients flow.
Cheatsheet — Neural Nets
- Init + normalization matter as much as architecture
- Watch gradient norms
- Learning rate is the first knob to tune
- Batch size interacts with LR (linear scaling rules)