We were able to significantly improve the model quality and cost to serve. These quality improvements come from our first continued pretraining run, providing a far stronger base to scale our reinforcement learning.
Continued Pretraining Boosts Model Quality and Cuts Serving Costs
By
–
