Elon Musk félicite ChatGPT (et donc OpenAI) pour le fait que l’IA réponde parfois « je ne sais pas » lorsqu’elle ne sait pas. Il déclare : « C’est une réponse impressionnante. » Si vous connaissez un tant soit peu l’IA, vous savez à quel point c’est remarquable. Car, par
SAFETY
-
AI Misinformation Spreads Rapidly During Spain Wildfires Crisis
By
–
Cuidado!
— Juan Merodio (@juanmerodio) 18 août 2025
En momentos críticos, como es el caso de los terribles incendios que están arrasando España, la desinformación corre como la pólvora en redes sociales. Además, la aparición de la ia no ayuda… pic.twitter.com/guvsgNq8KVCuidado! En momentos críticos, como es el caso de los terribles incendios que están arrasando España, la desinformación corre como la pólvora en redes sociales. Además, la aparición de la ia no ayuda…
-

Verena Rieser challenges single gold standard for AI alignment
By
–


A truly thought-provoking keynote at #DLI2025 by @verena_rieser ! Her talk, "Whose Gold?", challenged the idea of a single “gold standard” for AI alignment. It was a powerful call to action: moving beyond narrow safety metrics to embrace a pluralistic approach that accounts for our diverse, and sometimes conflicting, human values. #Indaba2025 #Urunana
→ View original post on X — @shakir_za, 2025-08-18 13:56 UTC
-
AI Model Evolution Safety and Training Infrastructure Challenges
By
–
Other AI research in last week's AI Lab newsletter: The evolutionary tree of AI models. How to stop chatbots learning to be unsafe. And the insane power needed for future AI training runs. …
-
Oxford researchers demonstrate AI safety training without output blocking
By
–
Researchers from the University of Oxford, the UK AI Safety Institute, and EleutherAI demonstrate a way to get AI models to avoid learning unsafe knowledge, as an alternative to trying to block them from outputting it post training.
-
AI Automation Risks Human Cognitive Atrophy Without Mental Exercise
By
–
Yes, AI gives everyone an iron man suit, but our bodies are wasting away underneath. I truly think we will start writing and thinking manually – with weekly recommended ranges to avoid brain rot. The cognitive equivalent of us working out our bodies, even if machines do most of
-
Grok AI Delays Politically Sensitive Responses to China Questions
By
–
Since it seems @grok often delays responses to public questions that may be deemed politically sensitive to China, here’s what it said privately.
-
LLMs accountability reliability hype limitations challenges
By
–
don’t insult my new friend @Lauren_79
;
it is not a skill issue. it is a hype and reliability issue, exacerbated by the fact that LLMs don’t know what they don’t. better prompts might sometimes yield better results. sometimes. but they should be held responsible for the output -

Chatbots and Information Manipulation: Ethical Concerns in AI Era
By
–
Build your own chatbot, and rewrite history in your favor! Happy Orwell Day!* *Every day is Orwell Day, in the ChatGPT era.
-
Nick Cave Warns: Technology Being Misused in Worst Ways
By
–
Nick Cave was right; this technology is being used in all the worst ways. it didn’t have to be this way.
