Nebula has a largest score jump ever on lmarena
@testingcatalog
-

Gemini 2.5 Pro outperforms o3-mini-high on reasoning tasks
By
–

Gemini 2.5 Pro evals Looks like it has a much better performance at reasoning tasks than o3-mini-high.
-

Nebula achieves the largest score jump ever on lmarena
By
–
Nebula has a largest score jump ever on lmarena
-
Use case where a model excels in testing
By
–
For which use case would you say it performs much better than other models? Testing time
-

Gemini 2.5 Pro available with 1M context
By
–
BREAKING : Gemini 2.5 Pro (experimental) is now available on Google AI Studio. – 1M context window
– Multimodal support
– Real-time stream support
– Jan 2025 knowledge cutoff -

Gemini 2.5 Pro disponible pour les comptes Enterprise
By
–
Gemini 2.5 Pro seems to be rolling out to Enterprise accounts as well!
-
Perplexity for news impresses with its UI
By
–
Perplexity for News will be amazing Love the new UI – much simpler to browse!
-

Gemini 2.5 Pro thinking model theory questioned
By
–

Got Gemini 2.5 Pro as well, and it disappeared afterwards. It may happen that, in this case, earlier reports on it being a Thinking model may not be true. This would mean that the response came from a different model and not 2.5 Pro.




