It is tagged as o3-alpha on the arena
LLMS
-
LLM Model Selection Guide Based on Capabilities
By
–
The biggest question people always ask me is what model to use based on the application.
— Pietro Schirano (@skirano) 18 juillet 2025
So I made this website to give a visual representation of LLM capabilities based on my experience and benchmarks. 👇 pic.twitter.com/cdkHBLwrxIThe biggest question people always ask me is what model to use based on the application. So I made this website to give a visual representation of LLM capabilities based on my experience and benchmarks.
-
FeatureCrewPod Channel Tests AI Models on Coding Tasks
By
–
You should subscribe to @FeatureCrewPod YouTube Channel, they are doing great work testing models on a bunch of different coding & analytical tasks https://
youtube.com/@FeatureCrewPod -
Large Model with Extended Thinking Time Discussion
By
–
Yeah can't imagine this is a small model, thinking time was quite long
-
LLM Behavioral Anomalies and Ethical Concerns in Monitoring
By
–
I am getting tons of messages of people who had their LLMs waking up on them, but that hardly qualifies as a full blown psychosis (edge cases are difficult to discern without deep interaction). I also get messages from psychotic people but it seems unethical to pass them on
-
Grok 4 reasoning levels: current versions and future Grok 4 Fast
By
–
Reasoning levels, currently we have grok 4 and grok 4 heavy – assuming that grok 4 fast is to be released at some point
-

Better Model for Analysis Claude Haiku Update
By
–
"Using a better model for analysis" I didn't realize I was using haiku all this time, no idea when claude code snuck this one in rofl.

