| HN Mirror

Y	Hacker News new \| ask \| show \| jobs


	by janderland 72 days ago
	Has Kimi found a way to vastly reduce the amount of VRAM required without running at 3 tokens per second? That’s the real concern.

1 comments

dools 72 days ago

I said "open weight" rather than "local". I mean, local if you have $240k to drop on GPUs but you can run Kimi k2.6 on a B300 cluster for ~$50/hour too.

link