|
|
|
|
|
by blcknight
4 days ago
|
|
> I hope the big labs will start using this benchmark in their RL pipelines. Labs do not train on benchmark data (allegedly). They can train on similar problems, but benchmarks have specific strings in them that labs are supposed to be aggressive in filtering out of their training corpora. |
|