You can now talk to a Copilot Portrait in real-time! We’ve heard from some users that they'd feel more comfortable talking to a face when using voice. So in the US, UK + Canada, we’re rolling out a new Copilot Labs experiment where you can talk with animated portraits.
MULTIMODAL AI
-

Computer Use Performance Gains in Mathematics and Finance
By
–
Aparte veo interesante la subida en Computer Use (OS World) frente a Opus 4.1, en matemáticas y finanzas.
-

LongLive: Real-time Interactive Long Video Generation Technology
By
–
LongLive
— AK (@_akhaliq) 29 septembre 2025
Real-time Interactive Long Video Generation pic.twitter.com/oSPtiqZoGWLongLive Real-time Interactive Long Generation
-
GPT-5-Codex outperforms in SVG bicycle drawing capabilities
By
–
It's SVG of a pelican riding a bicycle wasn't as good as GPT-5-Codex though, Codex is much better at drawing bicycles
-

Anthropic to release Claude Sonnet 4.5 soon
By
–

BREAKING : Anthropic is about to release Claude Sonnet 4.5 soon! SOTA on SWE bench
-
Hume AI Octave 2 Multilingual Model Announced
By
–
BREAKING 🚨: Hume AI is preparing to release Octave 2 Multilingual model! Here is a sample dialogue between a Robot and a Russian hacker.
— 🚨 AI News | TestingCatalog (@testingcatalog) 29 septembre 2025
"Expressive, natural-sounding voices in 10+ languages, low latency, perfect for real-time transition and conversational use cases" pic.twitter.com/IUIZ8gb8WUBREAKING : Hume AI is preparing to release Octave 2 Multilingual model! Here is a sample dialogue between a Robot and a Russian hacker. "Expressive, natural-sounding voices in 10+ languages, low latency, perfect for real-time transition and conversational use cases"
-
GATO Project: Building a Generalist Agent Model with GPT
By
–
Thank you … this certainly was and still is one of my favourite projects. The moment @scott_e_reed and I saw GPT we started working on Generalist AgenT One (GATO). We strongly believed a model could do anything a human could do from motor control, to perception, generation,
-

From Few-Shot to Zero-Shot: The Next AI Frontier
By
–
In 2020, OpenAI showed that language models are few-shot learners. In 2022, ChatGPT was released. By 2025, Google DeepMind demonstrated that video models are zero-shot learners and reasoners. What’s next?
-
a16z Games Backs AI 3D Creation Tools Startups
By
–
https://t.co/waqIw0EsGX https://t.co/Ci8SFhtKfh
— Bilawal Sidhu (@bilawalsidhu) 28 septembre 2025Big news, folks! I'm thrilled to join forces with @a16zgames for their upcoming @Speedrun class, backing visionary founders at the frontier of AI, 3D & immersive computing. They're looking for startups building the next-generation of: Visual 3D Creation Tools
-
Video Model Limitations and LLM Improvement Potential
By
–
In both cases you can come up with edge-cases that expose their limitations – with LLMs may or those limitations see overcome as the models improve, will be interesting to see if that happens for video models
