Hacker News new | ask | show | jobs
by epolanski 10 days ago
Even if hardware capacity increases, it seems clear it uses way more tokens, so I don't expect parity with other competitors on that front.

On the other hand I expect K3 future refinements to be massive and more efficient.

1 comments

The speed at which tokens are crunched, even on the same hardware, differs between models as well. Using more tokens is only a problem if they are processed at the same speed as with a comparison model.
Using more tokens is a significant problem if you pay per token?
Not if the price per token is significantly lower.

Also this arm of the discussion was about speed, not price.

> Using more tokens is a significant problem if you pay per token?

Yoh have posts in this thread suggesting that Fable is 5x more expensive than Kimi K3.