scalingfactneutralThe optimal compute split between pretraining and RL changes with scale, shifting from roughly 20% to 28% RL as budget growsLewis Tunstall08 Aug 2026https://x.com/_lewtun/status/2084700690820321319