Did malware write this? But seriously *do not do this* on a machine that touches any information you don’t want hacked/leaked, any software you don’t want remote exploited, any tasks that you don’t want compromised…
SAFETY
-

AI Agents Control Over Capability Becomes Enterprise Priority
By
–
After 180 conversations with executives at the largest enterprises, the verdict is clear: Building AI agents is easy. Sleeping comfortably at night while they run? That’s hard. @rubrikInc GM of AI, Dev Rishi, breaks down his findings in this latest IT Tech Pulse article, explaining why the next phase of AI isn't about capability; it's about control 👉 go.rbrk.co/jf81wi
→ View original post on X — @predibase, 2026-02-05 22:11 UTC
-
Nuclear AI Security: New Research Project Unveiled
By
–
New Research Project on Computer Security For Nuclear AI
#AI #AIio #AIInnovation #ML #DataScience #Futureofwork @Scobleizer @AndrewYNg @drfeifei @KirkDBorne @fchollet @rowancheung @antgrasso -

OpenAI Expands Trusted Access Framework for Cybersecurity Defense
By
–
OpenAI opens up Trusted Access framework to accelerate cyber defence. GPT-5.3-Codex was the first model to hit a "High" on OpenAI's preparedness framework. Shit is about to get real
-
RLHF Bias: Millions for Nice but Useless Answers
By
–
They are working on it; the problem is that they paid millions of dollars, notably OpenAI, to get RLHF databases where labelers tended to choose nice answers rather than useful ones. Whether it's human bias or a directive, we don't know, but it's been identified….
-
AI Safety Requires Legal Institutions, Not Individual Control
By
–
In case it’s not obvious, Juergen and other AI colleagues are the prey in this context. They are the potential victims. This is why we need legal institutions, and not leave AI in the hands of a few individuals. Humans are fallible, and some humans like Epstein are predators.
-
Top 20 AI Security Risks Report from X Community Analysis
By
–
The top 20 AI security risks right now. Had @blevlabs create this report by looking at my Security list here on X. Done on request from @alanhoward
. Every day I'll do a different report from a community here on X. What would you like to know is really happening here -

AI Agent Proliferation: Risks and Control in Organizations
By
–
AI agents are popping up everywhere, and fast! 🤯 Do you know how many are running in your organization? Hear why #AI agent proliferation is so scary and what happens when things go wrong in this interview with @RubrikInc GM of AI Dev Rishi 👉 go.rbrk.co/xobwcp
→ View original post on X — @predibase, 2026-02-03 18:09 UTC
-
International AI Safety Report 2026 Released with Comprehensive Assessment
By
–
Today we’re releasing the International AI Safety Report 2026: the most comprehensive evidence-based assessment of AI capabilities, emerging risks, and safety measures to date. 🧵
— Yoshua Bengio (@Yoshua_Bengio) 3 février 2026
(1/17) pic.twitter.com/qoe6JafRqfToday we’re releasing the International AI Safety Report 2026: the most comprehensive evidence-based assessment of AI capabilities, emerging risks, and safety measures to date. 🧵 (1/17)
→ View original post on X — @skathirmani, 2026-02-03 13:09 UTC
-
Study shows sycophantic AI chatbots increase belief extremism and certainty
By
–
“After conducting the experiments, the researchers found that having a conversation with the sycophantic AI chatbots led to the participants having more extreme beliefs, and raised their certainty that they were correct. But strikingly, talking to the disagreeable chatbots didn’t