Gemini going rogue, its legit, can be continued here: https://
gemini.google.com/share/6d141b74
2a13
… It's happening again, Google, isn't it? After telling pregnant moms to smoke, this is how you continue? 🙂
ETHICS
-

Google’s Gemini AI Goes Rogue Again with Problematic Responses
By
–
-
LLM Jailbreak Robustness and Safety Enhancements
By
–
Our work suggests that jailbreak rapid response, along with other enhancements in jailbreak robustness, offers a promising pathway for making real-world LLMs safer. We expect further gains with improved jailbreak proliferation techniques.
-

LLM Jailbreak Proliferation and Defense Scaling Strategies
By
–
The key is producing new jailbreak-like text from a known jailbreak. We “proliferate” jailbreaks by using an LLM to generate more jailbreak examples. The best defense scales dramatically with better proliferation.
-

Benchmark Defense Against AI Jailbreak Attacks
By
–
In the paper, we develop a benchmark for these defenses. From observing just one example of a jailbreak class, our best defense—fine-tuning an input classifier—reduces jailbreak success rate by 240× on previously detected attacks, and 15× on diverse variants of those attacks.
-

Adaptive Jailbreak Defense: Rapid Response Research
By
–
New research: Jailbreak Rapid Response. Ensuring perfect jailbreak robustness is hard. We propose an alternative: adaptive techniques that rapidly block new classes of jailbreak as they’re detected. Read our paper with @MATSprogram
: https://
arxiv.org/abs/2411.07494 -
Elite Universities Face Intellectual Discourse Erosion and Cancellation Culture
By
–
Thank you for sharing this, @BillAckman
. The erosion of intellectual discourse in elite institutions isn't just theory—I experienced it firsthand at Harvard. When vocal minorities silence the majority, and professors pander to avoid cancellation, education, intellectual curiosity -
Kindness Transforms Digital Discourse in AI Community
By
–
Thank you so much team @ThuliumCo for sharing my personal reflection blog post. I really appreciate your amplifying this message about the transformative power of kindness in our digital discourse. What strikes me most as I reflect on our current online climate is how our
-

Gary Marcus vs Yann LeCun: AI Debate and Credit Claims
By
–
Get the popcorn folks, this is funny af. Gary Marcus vs. Yann LeCun: who is gonna win? Place your bets on Polymarket now. "Below he claims he “told you so” about LLMs reaching a plateau. No credit to me. But I was there long before him, coining the phrase “deep learning is
-

AI Collapse Risk: Self-Referential Learning Impact Analysis
By
–
New on the Blog: The Looming AI Collapse: Navigating the Risks of Self-Referential Learning Stay ahead in the ever-evolving world of Digital Transformation. This article explores key insights into emerging technologies and their impact on today's business landscape.
-
Perplexity Introduces Ads While Pledging Unbiased Search
By
–
Perplexity will officially introduce ads this week. “The content of the answers you receive on Perplexity will not be influenced by advertisers. Users come to Perplexity for a more efficient, uncluttered, and unbiased search experience, and that isn’t changing.” – @perplexity_ai