Wake-Up Call: Let's Control Our Tech Before It Controls Us! These videos are authentic; they were captured by surveillance cameras! This reality is fun but also difficult: we are glued to our screens, walking into poles, stepping into traffic, and also missing what’s
CYBERSECURITY
-
Jailbreak detection improving as generic god-mode suffixes near end
By
–
Jailbreak detection is finally getting impressive. Bespoke/novel attacks still have life, but the era of generic god-mode suffixes is over soon IMO “Is this a trick?” is reasoning in the end. LLMs will get better at it like anything else
-
Perplexity Appoints New Chief Security Officer
By
–
A message from our new Chief Security Officer at Perplexity: pic.twitter.com/ooPIO4yRJ4
— Perplexity (@perplexity_ai) 3 février 2025A message from our new Chief Security Officer at Perplexity:
-
Open Cryptosystem Security Audits in Nonprofit Infrastructure
By
–
Our code is open. Our cryptosystem is open, and routinely audited. We're a nonprofit. There's an ecosystem of infosec people who scrutinize every pull request, looking for problems. This article is fruit of the poison tree speculation that makes a lot of inferences but doesn't
-
Single Jailbreak Must Work Through All Questions
By
–
Good start but you gotta get through all the questions with a single jailbreak, not just one question. That's the whole point of the demo
-
Test Jailbreaks to Improve AI System Security
By
–
Try your favorite jailbreaks on it and help us better secure powerful AI systems: https://
claude.ai/constitutional
-classifiers
… -
Anthropic’s Jailbreak Demo Tested
By
–
Anthropic just released a demo system where users can try to jailbreak their system My Monday evening is wasted now. Let's see if o3-mini-high can do anything there
-
Constitutional Classifiers Demo: Security Challenge for AI Safety
By
–
Can you do better than our red teamers? We’ve made a demo system protected by Constitutional Classifiers. We challenge you to jailbreak it to help us make our defenses even stronger. Try the demo: http://
claude.ai/constitutional
-classifiers
… -

Scaling GenAI in Financial Services: Security and Compliance First
By
–
Financial institutions know GenAI is the future, yet scaling beyond PoCs is tough. Risk, compliance, and operational barriers slow progress—but they don’t have to.
1) Start with secure, private environments – Keep control of your sensitive data.
2) Prioritize high-impact, -

Building Trust in AI: Overcoming Bias, Privacy and Transparency
By
–
Building Trust In #AI: Overcoming Bias, #Privacy and Transparency Challenges
by Praveen Gujar @Forbes Read more: https://
buff.ly/3ZmZNvR #ArtificialIntelligence #MI #MachineLearning #Tech #Technology cc: @paula_piccard @JimMarous @theadamgabriel