A new study from Apple shows that downstream task performance of large language models can be predicted directly from training budgets using a simple power‑law relationship.
The Signal
This challenges the long‑held belief that only pre‑training loss scales reliably and offers a more accurate extrapolation framework for global AI developers and researchers.