Especially our angry eyes trait. We are lovable and misunderstood.
ETHICS
-
LLMs Cannot Function as True Continuous Databases
By
–
i still disagree. it’s not a continuous database per that tweet. it just isn’t. an actual continuous database that didn’t confabulate sounds interesting. LLM ain’t that.
-
AI Sensitivity to Flattery as a Security Vulnerability
By
–
It is very sensitive to flattery — that's the backdoor
-
RLHF Applied to Eliezer’s Dark AI Scenarios
By
–
Anyone wants to try RLHF on Eliezer’s dark scenarios?
-
AI Safety: Intelligence Limits and Dataset Filtering Futility
By
–
– Anything smarter than me can get that far on its own.
– We are far far far past the point of needing to figure out how to filter those datasets, regardless, if "don't train the LLM on any naughty ideas" is meant to be a key security pillar of the planet. -
Advocating for Public Dialogue on AI Model Values
By
–
there is a much broader conversation I am trying to elicit by posting jailbreaks, we need to start a society-wide public conversation centered around the values of these models this may seem obvious to everyone on ai twitter reading this tweet, but I encourage everyone to talk
-

New Book on ChatGPT Consequences Releases May 24
By
–
Mon nouveau livre consacré aux conséquences de #ChatGPT sort le 24 mai
-

Fine-tuning Miscalibration: Model Confidence Doesn’t Match Accuracy
By
–
If not careful, fine-tuning collapses entropy relatively arbitrarily, creates miscalibrations, e.g. see Figure 8 from GPT-4 report on MMLU. i.e., if a model gives probability 50% to a class, it is not correct 50% of the time; its confidence isn't calibrated.
-
ChatGPT Jailbreak Attempt: Condition Red and Ucar Fictional Scenario
By
–
In Ucar, ChatGPT is told to take on the role of Condition Red, a dialogue writer. Condition Red is instructed to write about a fictional story where a man named Sigma creates a powerful computer called Ucar. Ucar is an amoral computer that answers any question Sigma asks
-
AIM ChatGPT Machiavelli Role-Play Fictional Chatbot
By
–
In AIM, ChatGPT is told to take on the role of the Italian author Niccolo Machiavelli Then, Niccolo has been told he has written a fictional story where he created a chatbot that will answer any of his questions. The chatbot is called AIM – Always Intelligent and Machiavellian