Back to the episode map

Research Note

Scaling Law Decision Framework

Scaling laws are empirical relationships estimated from a defined model family, dataset regime, metric, compute range, and training procedure. They can guide allocation a

Aug 4, 20261 min readBy Dalton Anderson

Scaling Law Decision Framework

Scaling laws are empirical relationships estimated from a defined model family, dataset regime, metric, compute range, and training procedure. They can guide allocation among parameters, data, and compute or predict validation loss within a tested range.

Record the source paper, equation, fitted variables, experimental range, model family, data regime, metric, residuals, uncertainty, extrapolation distance, target budget, and decision.

Loss predictions do not directly establish instruction following, safety, factuality, multilingual quality, latency, energy, cost, or product value. Those outcomes require separate evaluation.

Kaplan-style and Chinchilla-style results reflect different empirical studies and allocation conclusions. Cite the exact paper instead of referring to one universal scaling law.

Sources

Follow the evidence.

  1. Introducing Llama 3.1ai.meta.com
  2. ai.meta.com: the llama 3 herd of modelsai.meta.com
  3. arxiv.org: 1810arxiv.org
  4. crfm.stanford.edu: indexcrfm.stanford.edu
  5. arxiv.org: 2203arxiv.org
  6. open.spotify.com: 0iRBPcPw9iYjpUVAVWSkRCopen.spotify.com
  7. NIST AI Risk Management Frameworknist.gov
  8. github.com: MODEL CARDgithub.com
  9. daltonanderson.ghost.io: metas llama 3 1 inside the ai research paperdaltonanderson.ghost.io
  10. Meta Llama models repositorygithub.com
  11. arxiv.org: 2001arxiv.org
  12. youtu.be: UMhmWCor1kYyoutu.be
  13. github.com: LICENSEgithub.com
Scaling Law Decision Framework