Y
Hacker News
new
|
ask
|
show
|
jobs
by
cmiles74
20 days ago
I think both things can be true: new models benchmark higher
and
eat more tokens.
2 comments
jbvlkt
20 days ago
From my experience new models are slower and use more tokens even on questions which gpt 4 answered correctly. It is mostly because newer models tend to be more verbose (even with prompt requesting short answers).
link
bcjdjsndon
19 days ago
Unless somebody improved on the underlying transformer architecture... Surely AI is smart enough to do it by now
link