MIRI direct donation link: https://
intelligence.org/donate/ MIRI's fundraising page:
@esyudkowsky
-
MIRI Direct Donation Link and Fundraising Page
By
–
-
MIRI and Lightcone Final Fundraiser Push Recommendation
By
–
MIRI and Lightcone's SFF-matched fundraisers are coming down to the wire the next 3 days (as such often do). MIRI has $1M of $1.6M remaining on its match. Lightcone needs $1.6M of $2.6M. **If** you're equally motivated to donate to either, I'd say fill Lightcone's first.
-
Misunderstanding of Corrigibility Problem in AI Alignment
By
–
> the concern that corrigibility is in some sense a very anti-natural shape… Here, the basic vibe is something like: advanced, intelligent, self-aware minds have a strong tendency to want to “do their own thing” This doesn't sound like you understood the problem at all.
-
AI Companies Using Naive Obedience Training Instead of Corrigibility
By
–
If AI companies are trying to use any of my bright ideas that I once named "corrigibility", I haven't heard about it. They definitely haven't asked me for guidance. My impression is that they're doing naïve obedience training.
-
Corrigibility’s Challenge: Tensions with Coherent Reasoning
By
–
Corrigibilty is hard for different reasons from value alignment. Namely, that it cuts against the grain of coherent reasoning. This is harder to explain and fewer people ask about it, so it is little covered in the book. See eg https://
lesswrong.com/w/problem-of-f
ully-updated-deference
… for coverage of one -
Do you care whether the things you say are true?
By
–
Do you care whether the things you say are true?
-
ASI Safety Proposals Face Usefulness Trade-off and Enforcement Problems
By
–
Their proposal, even if you imagine it working, reduces to: Try to constrain the ASI so sharply that it's safe for the same reason a rock is safe, and nearly as useful as one. Implicitly, you then need to prevent anyone else from building a non-rock AI. They ignore *that*.
-
International Treaty Needed to Prevent ASI Development with Greater Bandwidth
By
–
You'd need an international treaty exactly as arduous as the ones we propose in order to prevent anyone anywhere from building a form of ASI that they expect to be actually useful to them, with more bandwidth.
-
ASI Security: Practical Limitations Despite Theoretical Safety Measures
By
–
They ignore the central question of whether you can do anything useful with an ASI even if you imagine that the security works. It can't talk to humans. It can't access the Internet. The bandwidth limitations preclude getting complete designs from it.
-
RAND Paper Criticized for Flawed AI Security Arguments
By
–
This was a bad paper and an embarrassment to RAND. They argue that humans have secured a bunch of specific attack avenues, therefore, something vastly smarter than humanity can be secured along every dimension.