Hacker News new | ask | show | jobs
GPT‑Red: Unlocking Self-Improvement for Robustness (openai.com)
35 points by alvis 17 days ago
1 comments

Useful direction, but the hard part seems to be measuring novelty after each fix. Are they reporting whether later red-team cases are genuinely distinct, or mostly variants of the same failure mode?