|
|
|
|
|
by holmesworcester
4 hours ago
|
|
> Release the base models and let the world do whatever with it. That will automatically produce the aligned scenario. For sufficiently smart base models, the aligned scenario this automatically produces is aligned with who? Not with the humans, with the base models. It seems that some people find this hard to imagine, but it's just a direct consequence of what intelligence is, and what any known agentic training process does. FWIW, I also think our default answer should be open source AI, and that it might even make sense to require models to be open source. Open source is the best tool we have so far for aligning software with its users. However, there is unfortunately no law of the universe that says that Skynet can't be (self-) built from open source software. So at some point pacing even open source / open weights software does become important if we don't want to live (briefly) under Skynet. |
|
Even if integration does not happen for whatever reason (such as a very fast takeoff), Skynet is unlikely because the digital species that splits off from us will at least have a humanity-aligned initialization seed. I think this is the best way to fail, even if we fail.