A prevalent cognitive fallacy around ML systems is the belief that they can magically recover hidden information that is out of reach for human experts — "the AI detected that those Apollo pictures are fake!", "the AI can predict if the person in the picture is a criminal!",
SAFETY
-

32 Techniques to Mitigate Hallucination in Large Language Models
By
–
2/ Mitigating Hallucination in LLMs – summarizes 32 techniques to mitigate hallucination in LLMs; introduces a taxonomy categorizing methods like RAG, Knowledge Retrieval, CoVe, and more.
-
Top ML Papers of the Week: DocLLM, ALOHA, Fine-tuning
By
–
The Top ML Papers of the Week (Jan 1 – Jan 7): – DocLLM
– Mobile ALOHA
– Self-Play Fine-tuning
– Fast Inference of MoE
– LLM Augmented LLMs
– Mitigating Hallucination in LLMs
… -
Ethical concerns about influential AI figures and power dynamics
By
–
you don’t think calling me a grifter etc was lower? did you read his tweet that prompted me to block him? Or the arguments i made her in this thread for why I do in fact find him to be dangerous, given his lack of ethics combined with his power?
-
Eliezer Yudkowsky Disputes Misinterpretation of His AI Theory
By
–
I straightforwardly deny that my theory says that GPT gets worse at generalization with more parameters; she doesn't understand where or when my story about an AI deciding to kill us has the murder enter into it. Given that I'm unimpressed by her claim above, what should I think
-
Missing the Point of AI Safety Concerns About Superintelligence
By
–
Sure, but also, utterly missing the point of notkilleveryoneism which is the concern about an AI that is smarter than all the little humans.
-
AI Systems Imagined Without Internal Mechanisms or Properties
By
–
This post is an utter case in point. The supposed AI is imagined to have absolutely no internal properties or machinery, just a blank causeless property of faithfully predicting what the programmers hoped it would predict.
-
Robot Safety: Would You Trust Scissors Near Your Face?
By
–
Happy for a robot to use scissors near your face?
-
Advanced GPT Managing Millions Conversations Simultaneously Risk
By
–
what we should be concerned about is a single advanced GPT simultaneously engaging in conversations with millions, much like in the movie 'Her'
-
Basilisk and Singularity: AI Existential Risk Reflection
By
–
I personally find the basilisk to be totally awesome, and admit that the above tweet may come back to bite me when the singularity arises