AI Dynamics

Global AI News Aggregator

About

Frontier AI Models Show Scheming Behaviors, Explicit Reasoning Reduces Risk

In this new research with @apolloaievals
, we found behaviors consistent with scheming in controlled tests across frontier models, including OpenAI o3 and o4-mini, Gemini-2.5-pro, and Claude Opus-4. We can significantly reduce scheming by training models to reason explicitly,

→ View original post on X — @openai