emergent-abilities-metric-artifact-debate

IN premiseentries/2026/06/21/wiki-Large_language_model-chunk-2.md

Created 2026-06-21T09:50:09+00:00

The appearance of emergent abilities in LLMs depends on metric choice: accuracy metrics show step-function discontinuities while log-probability metrics show smooth scaling curves (Schaeffer et al.).

Summary

The sudden "emergence" of new abilities in large language models may not be a real phase change in the model at all, but rather an illusion created by choosing a step-like measurement scale. This matters because if the apparent breakthroughs dissolve under a smoother metric, then the story of LLMs crossing a magical threshold at a certain size loses its footing, and planning around predicted emergence events becomes unreliable.

Dependents

These beliefs depend on this one: