|
|
|
|
|
by deyiao
1 day ago
|
|
It's unfortunate that the final few years of the fastest AGI progress happen to coincide with a global surge in nationalism, polarization, and mutual hostility between human groups. That raises a disturbing failure mode: under intense conflict, an AI could learn that exterminating some groups of humans is a justified or even desirable objective. Once that principle is accepted, the step to concluding that exterminating all humans is justified may become much smaller than we'd like to believe. |
|
That is not a failure mode; its the default. Whoever is developing an AI - let alone an AGI - will shape it to, or alternately only accept one that is aligned with their world-view. If it goes against their interests in any way, they'll pull the plug and start another round of training from the last acceptably-aligned checkpoint. US DoD A(G)Is will follow DoD doctrine, and the same goes for those by the PLA, anything less will not be fit for purpose.