they were terrified of GPT-2. definitely into boy who cried wolf territory.
SAFETY
-

Stanford HAI and AI4MH explore responsible AI for youth mental health
By
–
Last week, @StanfordHAI & AI4MH's 2026 symposium explored responsible AI adoption in mental health. CA Assembly Member @AsmMiaBonta stressed the need to protect youth as they turn to chatbots & other AI tools for mental health challenges. Full session https://
youtube.com/watch?v=KBZJYk
zqT0c
… -
Need for sustainable scaling and governance in AI self-improvement transition
By
–
The critical transition is from teleop-trained to self-improving systems. Many organizations miss the need for sustainable scaling strategies to support this leap. It's not just about the technology, but the governance frameworks to manage 'dangerous, valuable' AI responsibly.
-
AI improves fake news detection 21%, but feeling smarter is a trap
By
–
2/ With the AI, people got 21% better at catching fake news. The chatbot worked. In the moment, it was a sharp lie detector. Everyone felt smarter. That feeling is the trap.
-
MIT Study: AI Destroys Our Ability to Discern Reality
By
–
URGENT: A new MIT study has revealed that AI is destroying one of your most important skills. Discerning what is real. Here's what they found and how to stop it:
-
Anthropic fears Mythos misuse, fails to explain safeguards
By
–
Two things are true:
(1) Anthropic (or parts of it) are absolutely and sincerely worried about the misuse of Mythos-class models & have put in excessive safeguards until they are confident it will not be misused
(2) They have not succeeded in explaining/convincing people of this -
Argument for continued open frontier models: profitability and safety
By
–
Has anyone clearly laid out an argument for continued availability of frontier open weights models that are (1) profitable for firms to distribute free as costs rise & (2) safe enough post-Mythos that governments will not intervene to stop their nations labs from distributing?
-

AI must avoid any manipulation, even well-intentioned
By
–
Thank you, that's much better! Manipulation by AI must be avoided at all costs, even when it is well-intentioned!
-
Anthropic now refuses requests upfront instead of sabotaging work
By
–
Guess what Anthropic will now no longer sabotage your work and lie to you, instead it will just tell you upfront that it refuses your request (: we did it guys LOL
-

Early collaboration with regulators for ethical and safe AI
By
–
AI ethics becomes practical when companies collaborate with regulators from the start. Clear rules help teams design safer systems, reduce legal risks, and explain AI decisions with more confidence.