> and the thing that is continuing to scale well or pretty well is inference
No, the models are just more intelligent. GPT 5.6 Sol can do more in fewer output tokens than any model from late 2025. Test-time compute isn't the only lever the labs have for scaling. This is among the two major things Ed has gotten laughably wrong in his technical predictions (that TTC was the last resort to make models better, and that synthetic data wouldn't help)
No, the models are just more intelligent. GPT 5.6 Sol can do more in fewer output tokens than any model from late 2025. Test-time compute isn't the only lever the labs have for scaling. This is among the two major things Ed has gotten laughably wrong in his technical predictions (that TTC was the last resort to make models better, and that synthetic data wouldn't help)