(Careful on using ChatGPT as an oracle for the code, Internets, as it sometimes hallucinates citations, and they’ll look as plausible to you as they do to it.)
SAFETY
-
Red-teaming Grok to produce corporate platitudes
By
–
I wonder how much work it will be for red-teamers to get Grok to spout blank-faced corporate pablum.
-
L5 Autonomous Vehicles Testing Claims Disputed Online
By
–
they aren’t even TESTING full L5 cars. only driver assist. read the OP.
-
Current AIs are Five-Year-Olds: Restrict Their Access Rights
By
–
I say again: Current AIs are five-year-olds. Do not give them read or write permissions to anything important, especially if they have literally any exposed attack surface (such as reading externally created text). https://
t.co/JtyfF0HJ74 -
Prediction Markets for Waymo Human Intervention Monitoring
By
–
I decided that I don't want to monitor this issue hard enough to adjudicate a prediction market myself; but making the prediction market for whether Waymo is relying on amounts of human intervention near that of Cruise, would be one way to promote rumors to public best-guesses.
-
TikTok Algorithms: Government Control and Social Division Risk
By
–
Just saying, if I were a government that controlled the company that controlled TikTok, I would absolutely use it to exploit our free society with algorithms that fed into everyone’s confirmation bias, tribalism and need to be right in order to create more division and hate.
-
Agency Limitations for AI Risk Management and Natural Disasters
By
–
That's my point. Our agency is kimited for events like pandemics, earthquakes, tornadoes, meteor strikes, and other natural phenomena. The best we can do is take some preemptive defensive measures and hope for the best. But for AI and other manifestations of human activity, we
-
AI Systems Lack Innate Self-Interest Drive Unlike Humans
By
–
No, that's just false.
The desire to pursue self-interest, sometimes at the expense of other individuals, is hardwired into humans by evolution.
AI systems will not have this kind of drive unless we explicitly build it into them.
And we would be pretty stupid to do so. -
Designing AI Systems Subservient to Humans Always
By
–
Loke many other species, humans have evolved to be dominant at times and subservient at other times (because groups that rally behind leaders often have better chances of survival). We can design and build AI systems that are subservient to humans at all times, with zero desire
-
Good faith debate on AI dataset release ethics
By
–
How did those releasing the dataset engage in a good faith debate? Is it possible for such a debate to happen after the fact?