Hacker News new | ask | show | jobs
by wongarsu 13 hours ago
Is frontier scale larger than this? Kimi K3 seems to benchmark in the same range as Opus and Fable. I would have expected they are all in the 2-4T range, with quality of the training and architecture differences as the major differentiators
1 comments

The number of active parameters is vastly different. Deepseek CEO hinted that he estimates it as an order of magnitude difference in one of his recent interviews.

> Seems to benchmark

yes, but in human usage the differences show up

Would you happen to have a link to that interview? Sounds like an interesting read.
It was posted to HN a few days back

1. https://news.ycombinator.com/item?id=49019012 (original chinese)

2. https://news.ycombinator.com/item?id=49052912 (translated english [pdf])