Composer 2.5 is built on top of Kimi K2.5 Also interesting > Together with SpaceXAI, we're training a significantly larger model from scratch, using 10x more total compute. With Colossus 2's million H100-equivalents and our combined data and training techniques, we expect
GENERATIVE AI
-

Composer 2.5 Model Performance and Efficiency
By
–
Composer 2.5 is exceptionally intelligent and up to 10x more efficient than similarly capable models.
-
Technical implementation of AI integrity gates and hallucination checks
By
–
You got it right! both happens. The integrity gates at 2.5 and 4.5 run a 7-mode checklist and hard-block on serious stuff like hallucinations or unsupported claims. Milder flags just warn and let you decide at the human checkpoints.
-
User experience report on Claude model performance and responsiveness
By
–
Claude recently feels even more lazy than usual. I have to force him to do his job properly and correctly. At some point, I'm just writing in all caps and yelling at Claude in my head. Dont know if it helps.
-
PolyAI Launches New Platform for Rapid Voice Agent Development
By
–
This is insane… PolyAI just shipped the Open Platform and it is producing production voice agents, knowledge bases, and integrations from a single prompt.
— AI Highlight (@AIHighlight) 18 mai 2026
A year ago this required a telephony engineer, a conversation designer, a vendor team, and a $150K contract. https://t.co/Na4hlvzyIGThis is insane… PolyAI just shipped the Open Platform and it is producing production voice agents, knowledge bases, and integrations from a single prompt. A year ago this required a telephony engineer, a conversation designer, a vendor team, and a $150K contract.
-
AI agents approaching autonomous full analyses
By
–
Cada vez queda menos para que los agentes de IA hagan análisis completos solos
-
Addressing out-of-distribution detection in LLMs
By
–
Why don’t LLM’s just tell you when you are asking a question / doing something that is out of distribution?
-
Models show consistent theory-of-mind failures
By
–
Its a consistent theory-of-mind failure in models that are otherwise suprisingly good at theory-of-mind
-
MiniMax Releases M2.7 Open-Weight LLM with Self-Evolution Capabilities
By
–
💡 @MiniMax_AI M2.7 is an open-weight LLM built for serious dev work.
— SambaNova (@SambaNovaAI) 18 mai 2026
It’s the first in MiniMax’s M-series to “self-evolve” via its own training + eval loop (agent harness optimization). Designed for complex coding, multi-agent systems, and pro-grade workflows.
Learn more:… pic.twitter.com/BZJp8zgCLg@MiniMax_AI M2.7 is an open-weight LLM built for serious dev work. It’s the first in MiniMax’s M-series to “self-evolve” via its own training + eval loop (agent harness optimization). Designed for complex coding, multi-agent systems, and pro-grade workflows. Learn more:
-
Gemini Live Implementation Found in Desktop Client
By
–
That's the official (unreleased) implementation of Gemini Live inside the Gemini desktop. However, there is no guarantee that it won't change before it is released.