I think o1 is smarter than R1 too; I just mean because you can see the CoT, R1 is better for showing someone how o1 works
LLMS
-
O3-Mini vs Gemini 2.0 Pro: Battle of AI Models
By
–
O3-mini vs Gemini 2.0 Pro will be interesting. Lets see what will come on top
-

French Ministry AI for Students Called Catastrophic Failure
By
–
Les spécialistes numériques du Ministère de l’éducation se donnent du mal et sont BIEN INTENTIONNÉS Mais cette IA destinée aux jeunes Français est CATASTROPHIQUE Le niveau est nul : c’est moins bon que GPT2 C’est pathétique ! Sortir cette horreur en janvier 2025 est fou
-
Explaining o1 with R1 for a quick demo
By
–
Yesterday I had to explain to someone how o1 works, so I opened R1 for a quick demo.
-

Perplexity Adds Reasoning Toggle
By
–
Perplexity released a reasoning toggle to let users select between normal Pro Search and Pro Search with Reasoning. Now you are in control
-
O1 Pro Better Writing Quality: Training Data or Reasoning Engine?
By
–
O1 Pro generates better writing, none of the AI slop of “seamlessly, effortlessly, …”. Good, on point, communication. But I wonder if this is just the training data, and less of the reasoning engine. Are we being upsold a less sloppy model?
-
Reasoning Less Competitive Moat Than Expected in 2025
By
–
2025's biggest surprise so far: Reasoning is less of a moat than anyone thought.
-
Early AI Adoption and Knowledge Retention Beyond the LLM Era
By
–
i got into AI in 2017 thank you very much also "all the work went to the garbage after the llm era, but all the grinding and learning still stayed within me" lol
-
LLMs overhype: breakthrough claims versus realistic applications
By
–
You mean it’s not an AMAZING BREAKTHROUGH and here are 50 USE CASES THAT WILL BLOW YOUR MIND? (LLMs have a lot to answer for.)
-

Rigorous AI Benchmarks Advance Field Standards
By
–
Awesome work from @hendrycks @alexandr_wang and teams, the field urgently needs much harder and scientifically rigorous benchmarks like this – congrats, and we look forward to testing on it!