it goes without a saying stop all the killing
SAFETY
-
Human Factor Remains Weakest Link in AI Systems
By
–
Lisez bien. Dans n’importe quel système informatique même le plus performant doté d’intelligence artificielle, l’homme est toujours le maillon le plus faible.
-

Google’s strict safety filters limit AI capabilities
By
–
Ohh wow thanks that's new, will try, problem is the safety filter is still so strict but that's Google's issue
-
MonsieurPhi criticizes Luc’s update of AI knowledge
By
–
I've talked about it a billion times on Twitch :') … In short, @MonsieurPhi is absolutely right to point out the problems with Luc's updating of knowledge and beliefs on AI-related topics. Saying that ChatGPT still invents nonsense that he cited,
-
Safeguarding Against Harmful AI Ideas and Technologies
By
–
C'est pas grave, on fera toujours barrage à toutes les mauvaises idées.
-
AI-powered attacks pose safety risks to passengers
By
–
Ces attaques peuvent bien évidemment tuer ou blesser des passagers. https://t.co/37LEEkTXEI
— Olivier Rimmel 🕊️ (@OlivierRimmel) 10 septembre 2025Ces attaques peuvent bien évidemment tuer ou blesser des passagers.
-
LLM Code Execution Vulnerabilities: Shell Commands and Arbitrary Code
By
–
Haha that answer it gave you isn't actually correct: it CAN run arbitrary shell commands but you have to convince it to use Python's http://
subprocess.run(…, shell=True) You can even get it to run PHP or Deno or Lua if you know what you're doing https://
til.simonwillison.net/llms/code-inte
rpreter-expansions
… -
Beware of LLM Misinformation: Critical Thinking Essential
By
–
liars and clowns anybody who cannot see that shouldn't use an LLM because they could believe any crap they're told!!!
-
Safe AI Label Contradicted by Palantir Partnership Concerns
By
–
yup or the fact that they label themselves "Safe AI" and then go work with palantir on some sketchy stuff !!!
-

Model Manipulation Risks: Quantization to Safety Degradation
By
–
not your Weights means eventually not the same Model btw they can > quantize it
> distill it
> hot-swap to a cheaper/weaker checkpoint
> make the model manipulative
> fine-tune it in ways that break safety or depth
> drop its IQ
> run experiments on you and/or your data
>