Interesting that they mentioned faster & cheaper compared to OpenAI’s latest models not “customizable”. That makes me think they are specifically referring to gpt-oss, This in turn means they are using the small, dense Qwen3 models, maybe 0.6 to 4B range. And this is
Qwen3 Small Models: Faster and Cheaper Alternative to GPT
By
–