Un autre gros problème de ChatGPT-5 qui vient d’être révélé. La version disponible dans le Chat est classée seulement 5ᵉ sur LM Arena, loin derrière Gemini et Claude. La véritable version top 1, c’est celle de l’API. Mais OpenAI n’avait rien dit jusqu’à aujourd’hui, ce qui
LLMS
-

GPT-5 System Card Details Discussed in Post
By
–
Yeah there was a section in the GPT-5 system card about that
-
Grok Demonstrates Advanced Mathematical Problem-Solving Capabilities
By
–
I asked Grok to show me the math @BasedBeffJezos is working on right now. Nailed it: pic.twitter.com/gycvGMh8Cl
— Bob Gourley – e/acc (@bobgourley) 16 août 2025I asked Grok to show me the math @beffjezos is working on right now. Nailed it:
-

Access Available Models through Abacus.ai LLM APIs Route
By
–
Click on available models – https://
abacus.ai/app/route-llm-
apis
… -
Gemma 3 270M: Efficient Open Model for Edge Devices
By
–
New hyper-efficient addition to our amazing Gemma open models: Gemma 3 270M packs a real punch for its tiny size! It’s super compact and power efficient, so you can easily run your own task-specific fine-tuned systems on edge devices. Enjoy building with it!
-
Lawyer Liability: The Key Differentiator From LLM Services
By
–
Having the lawyer competently sign off on correctness and be punished if they are not correct is the only reason why clients are not using an LLM instead of a lawyer
-

GPT-5 Hidden System Prompt Transparency Request
By
–
Wrote up some notes on the API version of GPT-5's hidden system prompt – it definitely adds today's date and appears to add other stuff too I'd really like to see this documented by @OpenAIDevs – as an API user I want visibility into the whole prompt! https://
simonwillison.net/2025/Aug/15/gp
t-5-has-a-hidden-system-prompt/
… -
Grok 4 Performance: Code vs Information Gathering Capabilities
By
–
Grok 4 is pretty good. When it comes to code, it underperforms Claude and GPT-5 for me. But for information gathering, it’s really solid.
-

OpenAI Reasoners vs Non-Reasoners: Performance Comparison Analysis
By
–
How do OpenAI reasoners compare to non-reasoners? Looking at the @lmarena_ai rankings, we can compare models’ overall ratings with their category ratings, and clear patterns emerge. Reasoners (o3, o4-mini, GPT-5-Thinking) are much better at Maths and Hard Problems, while
