Kyle Harrison
research-paper

Broken Neural Scaling Laws

Ethan Caballero, Kshitij Gupta, Irina Rish and David Krueger 2023 View original ↗

TL;DR — Proposes a single functional form that fits how AI model performance changes with scale, including the bends and breaks plain power laws miss.

How much weight it carries: A peer-reviewed ICLR 2023 paper; useful, though the fits are descriptive rather than explanatory.

Where this came from

38 pages. A copy is archived locally against link rot; the header links the original source.