xAI keeps preparing Grok 3.5 for the upcoming release. A new model reference "grok-3-5-api-2k-p2" has been spotted in the latest web build, along with a mention of "grok 3.5 flexible input". Let's see if the API will be released as well
@testingcatalog
-
Feeling ‘Timescale Compression’ from Daily AI News
By
–
If you follow AI news, you are feeling "Timescale compression" every day
-

Claude works 7 hours without supervision
By
–
Claude can work autonomously for hours (e.g., 7 hours)
-

Mistral Enhances OCR with New Features
By
–

Mistral AI has released an update to the Mistral OCR! – New OCR model with overall improvements
– BBox Annotations: Structured outputs for extracted images
– Document Annotations: Structured outputs for documents -
Claude 4 Multimodal and Voice Mode Updates Expected
By
–
Will Claude 4 be multimodal? Claude voice mode is almost ready too!
— 🚨 AI News | TestingCatalog (@testingcatalog) 22 mai 2025
Now there is a high chance we will hear about it today as well. Currently it uses tts but with Claude 4 it likely to change!
Expected:
– Claude 4 Sonnet and Opus
– Voice mode 🔥
– SWE Agent updates https://t.co/yLhKmBmHJh pic.twitter.com/RkoIsqkvLlWill Claude 4 be multimodal? Claude voice mode is almost ready too! Now there is a high chance we will hear about it today as well. Currently it uses tts but with Claude 4 it likely to change! Expected: – Claude 4 Sonnet and Opus – Voice mode – SWE Agent updates
-

Mistral unveils Devstral, a SOTA open-source AI agent model
By
–

BREAKING : Mistral announced Devstral, a SOTA open-source model for building AI agents!
-

Native Multi-Speaker Voice Generation
By
–
Native speech generation can produce audio with multiple speakers from text. Includes:
– Gemini 2.5 Flash Preview TTS
– Gemini 2.5 Pro Preview TTS -

Gemini 2.5 Flash: New Audio Models
By
–

BREAKING : Native Speech Generation and Live Audio Generation with Gemini 2.5 Flash is now available on AI Studio! 4 new models from the Gemini 2.5 family Confirmed



