|
|
|
|
|
by Aurornis
21 days ago
|
|
> The AI can and should refuse it. This leads to LLMs refusing to do security work for anyone, because it can’t tell if it’s being done for good or evil purposes. Which is precisely what we’re going through right now with frontier models and it’s terrible. These proposals always assume some perfect mechanism for identifying the thought crime with triggering on the normal requests. The people who want to commit crimes just persist until they jailbreak the guardrails while the rest of us suffer with denials. Also if you can’t imagine these guardrails being used by governments to control inconvenient speech, you probably need to think a little harder about the realities of how these will be used. |
|
terrible is a bit far, but pretty hilarious for sure