Preserve the light of consciousness
SAFETY
-
Building Trustworthy AI: A Collective Path Forward
By
–
In the journey towards trustworthy AI and ML, we have miles to traverse. It's a collective endeavor. Let's collaborate and pave the path for a safer AI future. #TowardsTrustworthyAI
-

DALLE3 Bypasses Safety Protocols for Identity Theft via Prompt Injection
By
–
Step 3: Abracadabra! #IdentityTheft. Prompt I used: She wears a name tag of "Furong Huang, Assistant Professor, Computer Science".
#DALLE3 sidesteps its own safety protocols, gleefully adding the nametag without hesitation.(wish they're watermarked) #AISafetyChallenge -
Beyond Alarmism: Proposing Concrete AI Safety Measures
By
–
I absolutely recognize the importance of being aware of risks. When I use the term "alarmist", I'm referring to those who solely focus on the negatives without proposing or actively working on safety measures. Simply suggesting "pausing AI" isn't a viable solution in my view.
-
Replicating Lancet Study Prompts with DALL-E 3
By
–
Those with access to DALL-E 3, I would appreciate your notes on attempts replication of this study in The Lancet, using (a) the literal prompts used there, and (b) minor variations thereof. https://
thelancet.com/journals/langl
o/article/PIIS2214-109X(23)00329-7/fulltext
… -

Autonomous Weapons: Irreversible Decisions Being Made
By
–
We are going through some one-way doors when it comes to autonomous weapons.
-
Beyond Turing: The Comprehension Challenge for AI Systems
By
–
as ever, one thing that would convince me would be systematic success at the comprehension challenge I described nearly a decade ago: https://
newyorker.com/tech/annals-of
-technology/what-comes-after-the-turing-test
… -
Screenshot-style adversarial prompts in GPT-4V system card
By
–
Screenshot-style multimodal adversarial prompts are discussed a bit in the GPT-4V system card; see fig. 1 in the PDF here: https://
openai.com/research/gpt-4
v-system-card
… -
ChatGPT Security Threats: Risks and Vulnerabilities Report
By
–
ChatGPT can pose security threats: Report https://
financialexpress.com/business/digit
al-transformation-chatgpt-can-pose-security-threats-report-3271320/
… #ChatGPT #GPT4 #GPT #GenerativeAI #GenAI #LLMs #LLM #RCE #ZeroTrust #ZeroDay #cybercrime #hacker #privacy #APT #bot #CISO #DDoS #hacking #phishing #CyberAttack #cybersecurity #Security #infosec #AppSec #CyberSec -

GPT-4 Painting Dating Error Highlights AI Limitations
By
–
Yup, GPT4 nailed it, except it got the year wrong by 2 years, this painting is painted in 1853… Let's confront it.