Hacker News new | ask | show | jobs
by yreg 1 day ago
So they can charge more per token and decrease the pressure on their infra.
1 comments

Even if you specifically train the model on producing shorter answers, I would think producing good short answers would require more resources just as it does for humans (https://quoteinvestigator.com/2012/04/28/shorter-letter/: “If I Had More Time, I Would Have Written a Shorter Letter”)

If so, charging per output token is the wrong incentive.