Pre-training scaling laws have been widely validated and continue to yield gains, reinforcing confidence in the approach.
It’s continuing to give us gains. What has changed is that now we're also seeing the same thing for RL. We're seeing a pre-training phase and then an RL phase on top of that.