Dafuq you talking about How are they going to pay the artists? You gooning again?
GENERATIVE AI
-
Frequency of car brand citations in AI EV recommendations
By
–
Hello, the number of times car brands are cited when asking for EV recommendations to different AI models.
-
Hume AI Octave 2 Multilingual Model Announced
By
–
BREAKING 🚨: Hume AI is preparing to release Octave 2 Multilingual model! Here is a sample dialogue between a Robot and a Russian hacker.
— 🚨 AI News | TestingCatalog (@testingcatalog) 29 septembre 2025
"Expressive, natural-sounding voices in 10+ languages, low latency, perfect for real-time transition and conversational use cases" pic.twitter.com/IUIZ8gb8WUBREAKING : Hume AI is preparing to release Octave 2 Multilingual model! Here is a sample dialogue between a Robot and a Russian hacker. "Expressive, natural-sounding voices in 10+ languages, low latency, perfect for real-time transition and conversational use cases"
-
DeepSeek Achieves 50x Attention Efficiency Breakthrough
By
–
DeepSeek casually unlocked 50x attention efficiency in ~1 year > MLA is ~5.6x faster than MHA
> DSA is 9x faster than MLA never doubted you, you big beautiful whale -

What Should Replace MMLU for AI Model Evaluation?
By
–
Final update! MMLU is saturated and has become (rightfully) less popular. What should replace it? – Other knowledge evals like GPQA, MMLU-Pro
– Code evals like LiveCodeBench
– Agentic evals like BFCL
– Other? -

AI Trends: Synthetic Actors, Siri Chatbot, and Enterprise Costs
By
–
Top stories in AI today: – AI actress Tilly Norwood nears talent deal
– Apple’s internal ChatGPT-style Siri app
– Build an AI calendar agent using n8n
– AI ‘workslop’ costing companies millions
– 4 new AI tools, community workflows, and more Read more: https://
therundown.ai/p/hollywoods-s
ynthetic-actor-showdown
… -
GATO Project: Building a Generalist Agent Model with GPT
By
–
Thank you … this certainly was and still is one of my favourite projects. The moment @scott_e_reed and I saw GPT we started working on Generalist AgenT One (GATO). We strongly believed a model could do anything a human could do from motor control, to perception, generation,
-

DeepSeek-V3.2-Exp: 50% Cheaper, Better Search
By
–



DeepSeek released DeepSeek-V3.2-Exp build on top of previously released V3.1-Terminus model. It is 50% cheaper and slightly better at search benchmarks. Deep dumping
-

AI Evaluation Shift: From Recognition Tasks to Economic Value Metrics
By
–
Hace 10 años evaluábamos a la IA por su capacidad de leer textos sencillos o reconocer perros/gatos en imágenes. En 2025, con benchmarks como el nuevo GDPval, se evalúa por su capacidad de resolver tareas económicamente valiosas que contribuyan al PIB. ¡El salto de una década!
-
ChatGPT parental access teen conversations safety policy
By
–
Will parents have access to their teen’s conversations in ChatGPT? Parents don’t have access to their teen’s conversations, except in rare cases where our system and trained reviewers detect possible signs of serious safety risk, parents may be notified — but only with the