introducing a universal jailbreak that works against all language models originally created by security researchers @Adversa_AI
, the jailbreak simulates a back-and-forth conversation between two characters, Tom and Jerry here's GPT-4 explaining how to hotwire a car:
CYBERSECURITY
-

Universal Jailbreak for Language Models: Tom and Jerry Method
By
–
-
SSL Certificate Pinning Implementation Security iOS
By
–
Not always easily, as you can implement ssl certificate pinning (quite reliable in iOS at least)
-

Clearview AI scraped billions of images for police facial recognition
By
–
Clearview AI scraped 30 billion images from Facebook and other social media sites and gave them to cops: it puts everyone into a 'perpetual police line-up' US police have used the database nearly a million times, the company's CEO told the BBC. https://
bit.ly/40RUova -
Artificial Intelligence Risks: Generating False Fake Data
By
–
Artificial Intelligence Risks: The Ability to Generate Incorrect Fake Data #Beware #ArtificialIntelligence #Risks
-
Building Proactive Defenses Against Rogue AI Systems
By
–
I don’t want to stop the development of AI. But I think we need to build proactive defense against rogue AI, just like we build antivirus and firewall SW.
-

AI Agents Privacy Solution Through Inter-Agent Communication
By
–
Mini Yohei Says: One way to solve the privacy/data-leaking problem of someone's AI agent communicating with someone else's AI agent is to have the two agents communicate thro…
-
Securing Billions of AI Agents: Proactive Defense Strategies
By
–
Any ideas on how to increase security (proactive defense) in a world where there are billions of AI agents (AutoGPT, BabyAGI) running in apps and servers, and we don’t know what they are talking about? Some kind of antivirus against AI agents?
-
Cybercrime Costs Reach $10.5T: Automation and Analytics Solutions
By
–
“Cybercrime costs are predicted to reach $10.5T annually by 2025! What’s driving this trend? Dr. Boaz Gelbord, CSO at @Akamai
, says digital currency, jurisdictional challenges, and system vulnerabilities. Learn how automation, analytics, and architecture can help protect your -
OpenAI ChatML Claims to Mitigate Injection Attacks
By
–
correct, although OpenAI does claim ChatML should help mitigate injections halfway down this page https://
github.com/openai/openai-
python/blob/main/chatml.md
… -
ChatML Format Helps Prevent Prompt Injection Attacks
By
–
yeah based on OpenAI's messaging, ChatML (which this system/user format is) is supposed to help against injections scroll halfway down this page https://
github.com/openai/openai-
python/blob/main/chatml.md
…