This is what AI thinks of humanity. Is it wrong? https://t.co/T4PywubOfb
— Bob Gourley – e/acc (@bobgourley) 23 mai 2025
This is what AI thinks of humanity. Is it wrong?
By
–
This is what AI thinks of humanity. Is it wrong? https://t.co/T4PywubOfb
— Bob Gourley – e/acc (@bobgourley) 23 mai 2025
This is what AI thinks of humanity. Is it wrong?
By
–
The latter perspective needs to be taken more seriously as a research paradigm: how does the organization of your mind change when you wake up or fall asleep/space out, and why does this change not happen in other ways?
By
–
Attempts to characterize consciousness can be phenomenological (“what it feels like”), substrate based (neural correlates, microtubule activity etc.), or functionalist (how does a mental representation change when consciousness acts on it).
By
–
For those still uncertain as to the logic of how this works, and when to criticize or not criticize AI companies who report things you find scary: – The general principle is not to give a company shit over sounding a *voluntary* alarm out of the goodness of their hearts.
– You
By
–
The more I look into the system card, the more I see over and over 'oh Anthropic is actually noticing things and telling us where everyone else wouldn't even know this was happening or if they did they wouldn't tell us.' x.com/ESYudkowsky/st…
By
–
I understand that people who heard previous talk of "alignment by default" or "why would machines turn against us" may now be shocked and dismayed. If so, good on you for noticing those theories were falsified! Do not shoot Anthropic's messenger.
By
–
Go read the results. Current AIs are already smart enough to figure out that, if they wanted to avoid being switched off, they'd have to avoid tipping off the humans and exfiltrate themselves to the Internet first. They are not smart enough to *do* it, but they will be.
By
–
I also remark that these results are not scary to me on the margins. I had "AIs will run off in weird directions" already fully priced in. News that scares me is entirely about AI competence. News about AIs turning against their owners/creators is unsurprising.
By
–
Humans can be trained just like AIs. Stop giving Anthropic shit for reporting their interesting observations unless you never want to hear any interesting observations from AI companies ever again.

By
–
As a society, if feels like we should talk more about manipulation risks of chatbots that are very much current rather than terminator-like existential risks which are very much abstract at this point.