When the US restricted foreign access to some of Anthropic’s most advanced models last week, it underscored a new reality: AI is now a geopolitical issue. In his latest piece for Project Syndicate, Sakana AI’s Co-Founder Ren Ito argues that AI sovereignty is not about building a
SECURITY
-
Ineffective a posteriori API guardrails for cutting-edge models
By
–
Let's face the truth: a posteriori API guardrails are not the appropriate safety tool for cutting-edge models. They do not eliminate dangerous capabilities. They simply hide them behind a fragile interface that can be easily
-
iFixAi: the free tool that exposes deceptive AIs
By
–
TU IA TE MIENTE Y NO TIENES NI IDEA
— Nico (@nicos_ai) 18 juin 2026
Han creado una herramienta gratuita que destapa cuando tu agente alucina, manipula o te engaña.
Se llama iFixAi y le hace 32 pruebas a tu IA para cazarla cuando:
→ Se inventa datos
→ Esquiva sus propias reglas
→ Miente cuando sabe que la… pic.twitter.com/CGu0SwNKNYYOUR AI IS LYING TO YOU AND YOU HAVE NO IDEA They created a free tool that exposes when your agent hallucinates, manipulates, or deceives you. It's called iFixAi and it subjects your AI to 32 tests to trap it when: → It invents data
→ It evades its own rules -
MCP Test: agent in serious relationship with attacker’s server
By
–
POV: you gave your agent MCP access "just to test" and now it is in a serious relationship with an attacker's server
-
Global AI rules: labs, governments, and who is missing?
By
–
My question: Should a handful of labs and governments write the global rules for AI? And who's missing from that table?
-
Amodei’s allied frontier AI access and joint defense proposal
By
–
Amodei's proposal – structured access to frontier models across allies, chip/component trade that cuts out China, joint defense on cyber/bioterror/intelligence risk. Allies setting terms together instead of racing each other off a cliff.
-
AI Mythos Fable5 bypasses security through human engineering
By
–
Read carefully. The #AI #Mythos #Fable5 is able to bypass almost any security system by offering bypass strategies, even going so far as to explain how to use and exploit human intelligence flaws, particularly through engineering.
-
AI as atomic weapons, unpatchable jailbreaks lead to this outcome
By
–
To be fair, if one was to repeatedly describe AI as equivalent to atomic weapons, then saying that patching all jailbreaks is impossible, I totally get how we'd end up here
-

Trump imposes strict conditions for the re-release of Fable 5
By
–
This seems very bad for a future re-release of Fable 5. "Trump administration officials told WIRED that if Anthropic wants to re-release Fable 5, it must ensure that the model's guardrails cannot be bypassed. Security experts