Hacker News new | ask | show | jobs
by bloppe 14 days ago
> Earlier proposals from America and the European Union looked to the amount of computing power used to train a model as a rough guide for when oversight was required. Sir Demis instead suggests designating AI models as “frontier” if they meet certain thresholds on a selected set of benchmarks. The creators of those models would then be designated as “frontier labs” with extra responsibilities. Sir Demis is proud of the “elegant” way that approach sidesteps the question of whether academic or open-source models should be included or not.

Goodhart's law applies in reverse as well. Once the chosen benchmarks are known, model makers will aim to come in just under the threshold on those specific benchmarks, while maximizing their scores on other benchmarks.