Civilizational collapse awaits. (same in Europe my friend)
SAFETY
-

Generative AI Clinical Reasoning Models Fail Under Scrutiny
By
–
The performance of generative A.I. models for clinical reasoning are not holding up to increased scrutiny https://
arxiv.org/abs/2509.18234 @MSFTResearch https://
ai.nejm.org/doi/full/10.10
56/AIdbp2500120
… @NEJM_AI @AdamRodmanMD @LiamGMcCoy -
AI Error Rate 30% on Trust-Critical Client Work
By
–
Wow (and google translate is ai, but point still stands – wow). Error rate of 30% on trust-earning components of client work (like name) is a red flag.
-
Generative AI’s Failure to Build Robust World Models
By
–
Generative AI’s crippling and widespread failure to induce robust models of the world: https://
open.substack.com/pub/garymarcus
/p/generative-ais-crippling-and-widespread?r=8tdk6&utm_campaign=post&utm_medium=web&showWelcomeOnShare=false
… -
Securing AI Agents: Preventing Malicious Tool Exploitation
By
–
The easiest one is to make sure you don't expose all three legs of the lethal trifecta at the same time – and also that you design things to assume that anyone who gets malicious content into your agent can take full control any of the tools it's allowed to execute
-

OpenAI Tests GPT-5 Safety Routing for Sensitive Topics
By
–

OpenAI has begun testing safety routing to GPT-5 for 4o responses on sensitive and emotional topics. Loads of users are unhappy about that. GPT-5 as a guardrail
-

Jules by Google to get Memory feature soon
By
–
Jules by Google is about to get Memory soon! As well as a new file selector in the prompt composer. "Enable memories to let Jules use context from your past tasks to improve its responses"
-
Why Today’s Humanoid Robots Won’t Learn Dexterity
By
–
I have just finished and just published some weekend reading for you. 9,600 words of not easy reading, on why today's humanoid robots won't learn to be dexterous. https://
rodneybrooks.com/why-todays-hum
anoids-wont-learn-dexterity/
… -
Perplexity Tests New ‘Sonar Testing’ Reasoning Model
By
–
BREAKING 🚨: Perplexity is testing a new "Sonar Testing" reasoning model internally. Potentially, it will arrive as a reasoning upgrade to the existing Sonar model.
— 🚨 AI News | TestingCatalog (@testingcatalog) 26 septembre 2025
Imo, "Sonar Testing" is the best name ever 👀 pic.twitter.com/a1ChGPDGLhBREAKING : Perplexity is testing a new "Sonar Testing" reasoning model internally. Potentially, it will arrive as a reasoning upgrade to the existing Sonar model. Imo, "Sonar Testing" is the best name ever