|
|
|
|
|
by reasonableklout
1 day ago
|
|
It seems possible to deliberately not train on some offensive capabilities and still have a very useful model. For example, Opus 5 deliberately did not train on exploiting vulnerabilities, and so performed less well on exploits than Mythos, yet was equally proficient at finding such vulnerabilities, according to the Opus 5 system card in their "OSS-Fuzz" eval [1]. [1]: https://www.securityweek.com/anthropics-opus-5-nears-mythos-... |
|