Hacker News new | ask | show | jobs
by nzach 33 days ago
Instead of training the model to directly answer questions we trained the model to always write and execute the code that would solve the question ?

If that is the case, this isn't just a fancy way to perform prompt optimization?