|
|
|
|
|
by mindwok
9 hours ago
|
|
Does anyone else feel like the writing is on the wall for a future of local models? Spamming data centres everywhere, powering them, having to commit insane capital to hardware, all the effort to serve inference over a network reliably - when here we are with a frontier model nearly running on a laptop. Local AI on your device seems like a much more likely future to me than datacenters in space. For inference at least, training is another story. |
|
0.01 tokens per second means 1 million tokens ($3 worth of API usage [1]) takes 3.2 YEARS.
[1] https://www.kimi.com/resources/kimi-k3-pricing