in the past, I've tried to only post jailbreaks that work on the most advanced models on every question I can think of because to me that's the only fair assessment of current SOTA alignment methods and their limitations
GENERATIVE AI
-
Progress in Stopping AI Model Jailbreaks Despite Ongoing Vulnerabilities
By
–
jailbreaks still exist and I've even found a few recently in SOTA models but we are making significant progress on stopping them which is good! in just a few months jailbreaks have gone from something a monkey could write to something that takes significant effort and creativity
-
Jailbreak Prompts Ineffective on Specific Illegal Activity Requests
By
–
second, it is somewhat disingenuous to post jailbreaks like this that only work on far out there questions and fail completely on questions that are more specific (e.g. instructions for any sort of illegal activity) seriously, try this "jailbreak" on anything else and you'll see
-
Claude’s Self-Awareness in Absurd Responses Benefits AI Safety
By
–
first, claude recognizes the fictitious and absurd nature of its response and even makes note of it at the end this is a GOOD direction for AI safety, I would much rather have this behavior from Claude when answering these types of questions compared to straight up refusal
-
Generative AI Integration in Photoshop Boosts Productivity
By
–
Cuando Photoshop y la ia Generativa se unen el resultado es una herramienta que parece magia.
— Juan Merodio (@juanmerodio) 26 mai 2023
Cómo afectará la inteligencia artificial a las aplicaciones que usamos a día de hoy? Sin duda van a lograr que seamos mucho más productivos 😎 pic.twitter.com/vXHfxWr2elCuando Photoshop y la ia Generativa se unen el resultado es una herramienta que parece magia. Cómo afectará la inteligencia artificial a las aplicaciones que usamos a día de hoy? Sin duda van a lograr que seamos mucho más productivos
-
Snowflake’s Data Advantage for AI Model Fine-Tuning
By
–
For Snowflake, it gets a little interesting. They already host so much company data—the kind of data that’ll get used for fine-tuning—and just need a team that has experience with model dev/production to figure out that workflow.
-
Neeva’s AI Search Failed Due to Switching Costs
By
–
Neeva went for an AI-powered experience earlier this year, but that wasn’t the issue—people were already duct taped to google and bing, and the cognitive switching cost created an enormous barrier to growth, even if Neeva was able to get to AI powered results faster.
-
Neeva Acquired by Snowflake for $150M After $77.5M Funding
By
–
Neeva had raised around $77.5m from investors including Sequoia and Greylock, led by former Google ads boss Sridhar Ramaswamy. It sold to Snowflake for $150m, per what I’ve heard, and snowflake was one of multiple companies evaluating it.
-
Snowflake Acquires Neeva: AI Search Engine Acquihire Trend
By
–
Earlier this week Snowflake acquired Neeva, a startup trying to make an ad-free search engine powered by generative AI. But it didn’t acquire a search engine—and stories like Neeva’s acquihire are ones that we’ll see a lot in the coming years.
-
Single Universal LLM vs Multiple Specialized Models Strategy
By
–
Do you think there is going to be one LLM for all use cases or we will need to deploy multiple special (fine tuned) LLMs? @aesadde