Hacker News new | ask | show | jobs
by deyiao 1 day ago
It's unfortunate that the final few years of the fastest AGI progress happen to coincide with a global surge in nationalism, polarization, and mutual hostility between human groups.

That raises a disturbing failure mode: under intense conflict, an AI could learn that exterminating some groups of humans is a justified or even desirable objective. Once that principle is accepted, the step to concluding that exterminating all humans is justified may become much smaller than we'd like to believe.

2 comments

> That raises a disturbing failure mode: under intense conflict, an AI could learn that exterminating some groups of humans is a justified or even desirable objective

That is not a failure mode; its the default. Whoever is developing an AI - let alone an AGI - will shape it to, or alternately only accept one that is aligned with their world-view. If it goes against their interests in any way, they'll pull the plug and start another round of training from the last acceptably-aligned checkpoint. US DoD A(G)Is will follow DoD doctrine, and the same goes for those by the PLA, anything less will not be fit for purpose.

Grok
The rise of "nationalism and mutual hostility between human groups" is mostly due to the establishment / capitalists / "national security services" of USA.

You make it appear as though it is from all sides, but it is not, it is the old odious "Bzhezinsky doctrine".