AI Research Scientist · Generalization Theory & Phenomena
Grokking mechanistic explanations
Grokking mechanistic explanations
Weight-norm growth dynamics hypothesis
Grokking’s mechanistic explanations are competing causal hypotheses, not settled facts. This one claims the weight norm’s trajectory itself drives delayed generalization: it grows during memorization, plateaus, then its shrinkage triggers the switch to generalizing.Weight decay: delay then trigger
- 1Memorization dominates earlyWeight decay is weak early, so the large memorizing circuit fits training data fast while the smaller generalizing circuit stays underdeveloped.
- 2Decay erodes itOnce training loss is near zero, decay keeps shrinking the memorizing circuit’s weights every step, slowly eroding its dominance.
- 3Threshold flips controlWhen decay pushes memorizing weights below the generalizing circuit’s effective size, the cheaper circuit wins — test accuracy jumps abruptly.
- Signatureweight-norm and loss oscillate in spikes before the transition, rather than declining smoothly.
- Triggergeneralization follows a spike-and-recovery cycle in the training dynamics, not a steady descent.
- Signatureweight norm falls in one smooth, monotonic decline once decay dominates the loss landscape.
- Triggergeneralization appears once the shrinking norm crosses a threshold, with no spikes in the curve.
Recall check from the same lesson
If a generalizing circuit is already measurably more weight-efficient than the memorizing circuit from early in training, the circuit-efficiency hypothesis alone fully explains why the generalization transition doesn't fire until much later.
Review the explanation
Answer: False. Circuit efficiency is a causal claim about which circuit ultimately wins, not about when the switch happens — it says nothing about the delay itself. The timing is driven by cumulative weight-decay pressure (and possibly slingshot-style dynamics) eroding the memorizing circuit's dominance over many steps, so efficiency must combine with a decay-based account rather than substitute for it.
Sources
One sitting · 20–30 minutes
A focused session on your AI Research Scientist interview
LearnBench starts from what you already know — skip what you have, master what you’re missing.
Start now