scalingopinionneutralWhen test loss flatlines while training loss drops with scale, the model is fundamentally limited by information content in the data, not parameters or computeswyx & Alessio28 Jul 2026https://www.latent.space/p/xaira