Enable "Think" to use our reasoning model. It's best for math, science and coding. You can also ask Grok to “Think harder” about any question that might need a little more brain power
LLMS
-

Grok 3 Crushes Benchmarks in Reasoning Math and Coding
By
–
Grok 3 is the most powerful model ever: crushing it in reasoning, math, coding, world knowledge, and instruction-following tasks and showing remarkable performance across a range of benchmarks
-

Grok 3 Introduces DeepSearch and Think Features
By
–
With Grok 3, we introduce two new features: DeepSearch and Think. DeepSearch is a powerful agent that can rapidly synthesize key information, reason about conflicting facts & opinions, and distill clarity from complexity
-
OpenAI Strategy: Model Releases and Open Source o3-mini Alternative
By
–
they don't usually release models unless they've got a strong lead in some dimension. but would rather see oss o3-mini if they can't do that.
-
Phone AI Model Performance Threshold for Product Innovation
By
–
A 10x better phone AI model completely changes what people can build. If it's just 20% better it won't matter at all and they should just oss o3-mini instead. So the specifics matter here.
-
Tiny AI Models Need Quantum Leaps Not Incremental Gains
By
–
the base itself needs to be a leap ahead though or there's no point. a 2-10X better tiny model would be amazing. a 10% better tiny model would be pointless
-
Building the First Tiny Smart Model Not Another Generic One
By
–
no point in doing it if it's the 101st tiny dumb model, it needs to be the 1st tiny smart model
-
Small Base Models Beyond Distillation: Seeking Better Alternatives
By
–
id want to see a great small base model. i don't think distilling alone will yield something a few multiples better than what's out there today
-
O3-Mini distillation challenges depend on student model architecture
By
–
a distillation being good will still depend on the student model architecture. not convinced distilling o3-mini to existing models will yield anything significantly better than what's out there today
-

DeepSeek-R1 Performance Evaluation: Benchmarks and Enterprise Use Cases
By
–
Does DeepSeek-R1 live up to the hype? We put DeepSeek to the test, evaluating its performance on key benchmarks, comparing it to other leading models, and exploring its best use cases for the enterprise. See the results and key takeaways: https://
bit.ly/4b6iGYa #deepseek
