As models grow, some abilities (multi-step arithmetic, in-context learning, instruction following) seem to switch on past a size threshold rather than improving gradually. These emergent behaviors surprised researchers and are part of why scaling was pursued so aggressively. There's debate about how "sharp" emergence really is (it can depend on the metric), but the practical point stands: scale unlocked qualitatively new capabilities.