Machine Learning Street Talk
The intelligence explosion assumes intelligence is a single number you can buy with compute
Kurzweil's law of accelerating returns rests on cherry-picked data, and every exponential ends
Open-weight models still trail the frontier on size, active parameters and data
Predicting latent representations rather than raw tokens could make learning far more sample-efficient