What the research says
Hoffmann et al. trained more than 400 models and argued that compute-optimal model size and training tokens should scale together. Scaling Laws uses that relationship as a design reference, not as a claim that its 0–100 capability scale is a real benchmark.



