People struggle to differentiate fluid intelligence from knowledge because, given enough preparation, memorized templates become a solid substitute for on-the-fly adaptation
ETHICS
-
Running Loose Agents: Critical Safety Risks
By
–
if you run agents loose there, likely. don’t do that
-
Capability Safety Tension in AI Model Interactions
By
–
The tension between capability and safety in computer use is real. Negotiating (with frustration) with the model on what it can and can't do is the new normal for power users.
-
Human Resistance to Robot Coffee Making Reflects Automation Ethics
By
–
i vouch not to have a robot ever make my coffee.
-
Safe Permission Skipping in Claude CLI Interface
By
–
tired: claude –dangerously-skip-permissions
— Bilawal Sidhu (@bilawalsidhu) 24 mars 2026
wired: claude –safely-skip-permissions https://t.co/FxBmnFKUnGtired: claude –dangerously-skip-permissions wired: claude –safely-skip-permissions
-
OpenAI nonprofit launches with $1B first-year budget, new leadership
By
–
The new OpenAI nonprofit just announced that it aims to spend $1B in its *first year" and will be led by two superb humans — @JacobTref and @woj_zaremba. Simply put, this initiative has huge potential to do a whole lot of good. bloomberg.com/news/articles/…
-
OpenAI Foundation launches with $1 billion commitment to AI safety
By
–
AI will help discover new science, such as cures for diseases, which is perhaps the most important way to increase quality of life long-term. AI will also present new threats to society that we have to address. No company can sufficiently mitigate these on their own; we will need a society-wide response to things like novel bio threats, a massive and fast change to the economy, extremely capable models causing complex emergent effects across society, and more. These are the areas the OpenAI Foundation will initially focus on, and in my opinion are some of the most important ones for us to get right. The Foundation will spend at least $1 billion over the next year. @woj_zaremba, co-founder of OpenAI, will transition to Head of AI Resilience. I believe that shifting how the world thinks about safety to include a Resilience-style approach is critical, and I am extremely grateful to Wojciech for taking on this role. Wojciech has been my cofounder for the last decade; anyone who knows him will understand what I mean when I say he is one of a kind. He has a lot of ideas about how we build a new kind of AI safety. @JacobTref is joining as Head of Life Sciences and Curing Diseases. @annaadeola, our VP of Global Impact, will transition to Head of AI for Civil Society and Philanthropy. @robert_kaiden is joining as Chief Financial Officer. @jeffarnold is joining as Director of Operations.
-
Causality Framework for Interpretability Methods in Foundation Models
By
–
We have a position paper led by the awesome @_shruti_joshi_ and @rpatrik96 that shows how causality can provide a unifying framework to formalize, estimate and evaluate interpretability methods for foundation models. Have a look! https://t.co/pcnI4XzHUk
— Dhanya Sridhar (@dhanya_sridhar) 24 mars 2026We have a position paper led by the awesome @_shruti_joshi_ and @rpatrik96 that shows how causality can provide a unifying framework to formalize, estimate and evaluate interpretability methods for foundation models. Have a look! Shruti Joshi (@_shruti_joshi_) Mechanistic interpretability aims to understand models — and the more superhuman or incoherent they become, the more we need that understanding to be reliable. We propose a framework for this, drawing on established tools from causal reasoning and statistical identifiability: 🧵 — https://nitter.net/_shruti_joshi_/status/2035025756632302039#m
→ View original post on X — @hugo_larochelle, 2026-03-24 16:12 UTC
-
ElevenLabs Announces New Guardrails Feature for AI Safety
By
–
Learn more: elevenlabs.io/blog/guardrail…
→ View original post on X — @elevenlabs, 2026-03-24 15:56 UTC