What’s more likely in the techno-fascist idiocracy of the 21st Century?
We somehow manage to turn statistical models into super-intelligence… or the incompetent, criminally-corrupt, sociopathic leadership at the top of capitalism train an AI to defend the rich from the working class — by training it to be the most efficient killing machine possible — integrate it into fully autonomous killing machines, and it (accurately) determines the most efficient way to kill 99% of humanity is to start a nuclear war?


Yeah I’ve said a number of times that the AI being designed today could easily destroy humans, but not because it becomes self aware. The more likely possibility is that this beta tool gets put in control of critical systems by someone cutting corners and bam, total meltdown. Not because the AI is intending to do harm, but instead simply because it is ill suited for the task and NOT an actual intelligence.
AI being conscious or actively wanting to harm us isn’t the scenario most people worrying about the alignment problem are concerned about. The worry is that it ends up destroying us as a side product of something else it does. In a similar way humans destroy an anthill where we want to put up a house. We don’t do it because we want to harm ants.
The ever classic “paperclip maximiser” problem.
An AI is given the goal of maximising paperclip production in a factory. It’s given no other guard rails. Some shenanigans later, it finishes wiping out humanity, in order to divert more resources to paperclip production. (Apparently humans don’t want the iron in their blood used to make paperclips)
Yes, or an example I personally like to use that’s slightly less cartoonish than Bostrom’s Paperclip Maximizer is one that might ironically also come from him: getting into a self-driving taxi and telling it “to the airport, as fast as possible” and then you arrive 3 minutes later with the car completely wrecked, terrified out of your mind, covered in shit and vomit and with a 5-star wanted level. You got what you asked but not what you meant.
Yeah, I suppose Hanlon’s Razor applies to AI as well.
Exhibit A: most major companies in the past few years. Good thing so far it’s been mostly self harm. I’m sure a lesson has been learned and the race for AI in everything has stopped now. Right? Right?