Hacker News new | ask | show | jobs
by dhorthy 1 day ago
Yeah I would hold that models don’t know how to simplify because most rl/benchmarks doesn’t penalize complexity