|
|
|
|
|
by SXX
49 days ago
|
|
This. People who care about animal cruelty dont go building largest ever meatfarms and slaughterhouses. People who opposing arms manufacturing and gun violence dont jump to work for gun companies. People who really want AI benefit all humanity dont stick working with lying CEOs who want to convert company from a non-profit. Etc. So many examples. |
|
A first group dismisses the problem entirely, saying intelligence != power and AI doesn't have "drives".
A second group believes that alignment is solvable through engineering and iteration, and that we have the best chance of surviving if people with the right intentions are the ones working on it.
A third believes that aligning a superintelligence is a unique category of problem, that we are nowhere close to the level of scientific understanding needed to achieve it, that we only have one shot (because once a sufficiently powerful superintelligence exists it will thwart all future attempts, and alignment techniques that worked on dumber AI will likely not work on it), and that the world will have to coordinate to avoid killing ourselves off by building superintelligence before we understand how to do it safely, the way we have coordinated to avoid nuclear war.
The Anthropic and OpenAI founders, Elon, and Anthropic engineers are mostly in the second category. Some safety people at Anthropic and OAI are in the third category, but leading people in the third category think that pure safety roles at the labs are potentially impactful enough to be worth not quitting.