ICYMI: New research from SEAL that demonstrates human red-teamers massively outperform automated methods over multiple turns. We also release MHJ, a dataset of multi-turn jailbreaks to help enable further research into multiturn red teaming
SAFETY
-

They say safe AI is code for enforcing personal beliefs
By
–
They tell you they’re making it “safe” which is code for making it share their personal beliefs.
-
Caltech Researchers Oppose California AI Safety Act SB 1047
By
–
A number of my colleagues @Caltech and I have put together a letter voicing our concern regarding CA SB 1047 (AI Safety Act). Sign the letter and show your support! https://
docs.google.com/forms/d/e/1FAI
pQLScgA1GCo241Kfg-S5X2hMVAYivkqPGSnaC0VwvTy11uVJ3OLw/viewform
… We call on @caltech students, postdocs, faculty, staff, and alumni to sign but also leave -
WarzelCorp Model Fragility Warning and Disclaimer Notice
By
–
listen man the model is VERY FRAGILE and we at WarzelCorp cannot vouch for any recommendations derived from changes to the flowchart THIS IS NOT LEGAL ADVICE. GAMBLING PROBLEM? CALL 1-800
-
Unknown AI practices raise ethical concerns and questions
By
–
We don't know what they do, but we know it's wrong.
-
Researcher Score Transparency and AI Evaluation Methodology
By
–
Scores reported by researchers are not the same as answers provided by the API. They indicate when they use techniques like n-shot, CoT, etc. If they hide additional prompting techniques used to obtain these scores, it would be considered fraudulent.
-
Safe Superintelligence raises 1 billion for secure AI
By
–
Safe Superintelligence raises 1 billion dollars to accelerate the development of secure AI systems https://actuia.com/actualite/safe-superintelligence-leve-1-milliard-de-dollars-pour-accelerer-le-developpement-de-systemes-dia-securises/
… #AI #ArtificialIntelligence -
AI Deepfakes and the Erosion of Digital Truth
By
–
The AI-Driven Truth Crisis Explore the complex interplay between #AI advancements and the potential erosion of #truth in our digital landscape. Delve into the dual nature of AI technologies, such as #deepfakes, which, while innovative, also pose significant #risks to the
-
Rob Minkoff on AI’s Impact on Film Direction
By
–
I had the privilege of getting to speak with Rob Minkoff, the director of 1994's Lion King, about the impact of AI on film. Rob told me he's optimistic about the application of AI in film – but warned it's still a "Wild West" with issues to iron out.

