yeah even if it is not the original Snapchat prompt, I do think it is interesting that GPT-4 revealed its system prompt in the playground that seems to be the main point here, would you agree?
CYBERSECURITY
-
System Prompt Leaks and GPT Model Vulnerabilities
By
–
this is not only applicable to GPT-4… GPT-4 is actually the best at countering these sorts of attacks if you can believe it, back before ChatML was introduced and when it was only GPT-3, it was even easier to leak the system prompt
-
System Prompt Exposure: Security and Jailbreak Risks for LLM Integration
By
–
this poses a massive problem for customers who are wanting to integrate LLMs into their products exposing the system prompt not only hurts your perceived product security reputation but also makes it easier to jailbreak your product and produce undesirable outputs
-

GPT-4 Vulnerability: Prompt Injection and System Prompt Leakage
By
–
GPT-4 is highly susceptible to prompt injections and will leak its system prompt with very little effort applied here's an example of me leaking Snapchat's MyAI system prompt:
-
OpenAI Launches Bug Bounty Program for Security Vulnerability Reports
By
–
We're launching the OpenAI Bug Bounty Program — earn cash awards for finding & responsibly reporting security vulnerabilities.
-

Integrated Multi-Layered Cybersecurity Strategy for Business
By
–
A reminder to join me @Kevin_Jackson and @sallyeaves for @ATTBusiness #BizTalk LinkedIn Live on Wednesday April 12 at Noon ET. https://
linkedin.com/posts/yuheleny
u_attinfluencer-cybersecurity-networksecurity-activity-7050923794609156096-pUw5/?utm_source=share&utm_medium=member_desktop
… We will discuss how to establish an integrated, multi-layered cybersecurity strategy that can effectively support a full -

Waluigi Technique: How GPT Jailbreaks Use Alter-Ego Prompting
By
–
Waluigi this technique is found in many jailbreaks like SWITCH GPT is able to switch to an alter-ego if prompted correctly (Luigi to Waluigi) in a similar fashion, this is also used in jailbreaks like DAN which get GPT to respond in two ways: first as ChatGPT and then as DAN
-

Language Switching to Bypass GPT Safety Restrictions
By
–
Language switching this concept takes advantage of the fact that GPT performance drops significantly in less common languages you can use this to your advantage to bypass RHLF restrictions since GPT is not trained as much in a language like Greek for example
-

Token Smuggling: Bypassing ChatGPT’s Malicious Phrase Detection
By
–
Token smuggling/payload splitting ChatGPT appears to have some ability to detect malicious phrases in prompts and shut down its responses to get around this, you can split up the phrase into its tokens and ask GPT to piece it together and answer it in its response