EVERYTHING AIAI engineering, made visual
Visual edition planned
Phase 07 · Lesson 13LearnPython45 min16 lessons in phase

Scaling Laws

The 2020 Kaplan paper said: bigger model, lower loss. The 2022 Hoffmann paper said: you were under-training. Compute goes into two buckets — parameters and tokens — and the split is not obvious.

Visual edition planned

This lesson isn’t interactive yet.

It is part of the curriculum and will get the same treatment as Phase 1 — a visual cover, hands-on labs, derivations with numeric checks and a quiz. Until then, the original lesson is the best place to read it:

Part of Phase 07Transformers Deep Dive. Use the previous / next cards below to keep browsing the phase.