Y
Hacker News
new
|
ask
|
show
|
jobs
GPT‑Red: Unlocking Self-Improvement for Robustness
(
openai.com
)
35 points
by
alvis
17 days ago
1 comments
jing09928
16 days ago
Useful direction, but the hard part seems to be measuring novelty after each fix. Are they reporting whether later red-team cases are genuinely distinct, or mostly variants of the same failure mode?
link