The US government forces Anthropic to block global access to its new AI module https://
youtu.be/rGFyB0jPwkU?si
=0RDQENU9NHl5AKhi
… via @YouTube #anthropic #mythos #fable5 #cybersecurity #vulnerability #infosec #security #AI @sonu_monika @enilev @Jagersbergknut @TysonLester
SAFETY
-
US Government forces Anthropic to block global access to its AI
By
–
-
Gary Marcus notes rediscovery of his 2022 AI insights
By
–
So funny how so many people are rediscovering so many things I pointed out in late 2022. (In this case in my essay “AI’s Jurassic Park Moment”)
-
Gary Marcus questions if Karpathy was hired for recursive self improvement
By
–
is that actually true that @karpathy was hired to “run recursive self improvement”? https://t.co/nDmB4P99Z3
— Gary Marcus (@GaryMarcus) 13 juin 2026is that actually true that @karpathy was hired to “run recursive self improvement”?
-
Fable’s broken guardrails surprise only Anthropic
By
–
No one was surprised that Fable’s guardrails were broken. Except Anthropic.
-
Guardrails of cutting-edge APIs: superficial and bypassable
By
–
Many people have known for a while that the guardrails for cutting-edge model APIs are very easily bypassed, quite superficial, and impossible to fix. They are for the most part just a smokescreen and a distraction, in my opinion. We need a
-
Studying and regulating the dangerous use of AI
By
–
We should continue to study and regulate the dangerous use of AI in cyberattacks, financial services risks, fraud, biological warfare and other areas. AI safety is incredibly important, but slowing down progress
-

Thought-Aligner corrects AIs’ dangerous thoughts in real time
By
–
What if your AI agent could think twice before acting dangerously? Researchers from Fudan University present Thought-Aligner, a plug-in safety model that detects and corrects dangerous thoughts in real time, before they become
-
Jailbroken models require better tech and fair enforcement
By
–
Every model has been jailbroken. We need better tech. But not selective enforcement.
-

Amazon CEO warned Trump officials about Claude security risks
By
–

It was in fact Amazon (CEO Andy Jassy) who reportedly helped trigger the Claude shutdown. Via The Information Amazon CEO Andy Jassy reportedly warned senior Trump administration officials about security risks in Anthropic’s newest Claude models, helping trigger late-night export
-
The smartest AIs fail a simple color test
By
–
The smartest AI models on Earth just failed a color test that a 6-year-old can pass. The longer the test went, the worse they performed. One went from 91% to 15%. Scientists have discovered a flaw in the way AI