Shipments have already started to first customers. Knowing how Nvidia works, there will be at least 6-9 months worth of bugs to fix first. My realistic expectation is the end of this year at the earliest for the inference part. Some time next year the next generation models will
LLMS
-
WordPress Categories for AI News Aggregator
By
–
This is my first attempt at subscribers thread. 1/2
-

GPT-5.4 Thinking now available on Replicate with extended context
By
–
GPT-5.4 Thinking is live on Replicate Handles up to 1M tokens of context, best for agentic coding for complex tasks.
-
Grok not as good, Levangielabs way ahead of hyperscalers, says Scobleizer
By
–
Grok isn't even close to as good. I tried. https://
levangielabs.com is way ahead of any of the hyperscalers. Seriously. You don't know, because you don't have access (I'm the first that isn't a big company). -

Switching AI LLM Tools: Ollama to llama.cpp, OpenClaw to Hermes
By
–
looks like we’re slowly improving (learning more about tikz now)
-
GPT-4 Mentioned in Argument, Highlighting AI Relevance
By
–
although the form of your argument has some merit, the *abstract* mentions GPT-4. hence “I only check the title, but i am sure …. “ applies to your counterargument
-

GPT-5.4 scores 74% on ARC-AGI-2
By
–

GPT-5.4 scored 74.0% on ARC-AGI-2 GPT-5.4 Pro got 83.3%, getting close to a level of Gemini 3 Deep Think.
-
GPT-5.4 Thinking CoT Controllability and Safety Monitoring
By
–
We're publishing a new evaluation suite and research paper on Chain-of-Thought (CoT) Controllability. We find that GPT-5.4 Thinking shows low ability to obscure its reasoning—suggesting CoT monitoring remains a useful safety tool.
-
OpenAI GPT 5.4 Launched: Major Update for Builders and Developers
By
–
Breaking News: @OpenAI GPT 5.4 launched. Just posted a MAJOR update to the report. Now heavily focused on what builders/developers need to know.
