Hacker News new | ask | show | jobs
by wolttam 12 days ago
I'm a bit nervous this one isn't going to be open-weights. Any mention of "open" has been struck from the literature for this model (it was present an hour ago). We don't even know active params?

At this pricing, I'll be surprised if it's open.

4 comments

They will release the full weights by 7/27 along with support in vLLM.

Source: their release blog on WeChat. https://mp.weixin.qq.com/s/V4xhEIy8xDXSMDPrPkmUAQ

>We are currently working closely with our inference partners and open-source maintainers to align the technical details and ensure the model can be reliably deployed across the ecosystem. The full model weights will be released by July 27, 2026. Further details regarding the architecture, training, and evaluation will be released with the Kimi K3 technical report.

(translated by chrome)

11 days is a long time. It does not take that long to implement inference at providers. In my opinion, seems like they're being pre-emptively cautious about government intervention/review

Actually it does for a massive model, serving it correctly is not easy.

I believe Kimi also does some sort of Q&A and eval for day 0 partners, since early on a long of inference providers just weren’t running their models properly.

Eh, Minimax M2.7 also took a similar amount of time (actually longer) between availability and weights release.
I'm so glad to be wrong!
Reuters has been reporting that Chinese government is undergoing similar investigation to the US; blocking the export of domestic frontier models. They boil down to "anonymous sources" but it does seem inevitable as the tech gets stronger and stronger.
> The recent meetings (past month) with Alibaba, ByteDance, Z.ai, etc., were primarily about overseas acquisitions, foreign investment, and tech/talent outflow controls and not blocking foreigners from using Chinese AI models.

That is just misleading reporting. Very convenient for US frontier labs. This is actually what happened.

I am afraid this is may happen soon.

Now that they have compute capacity to train larger models, there is a non-zero chance they will be in the lead by next year.

In which case they will probably stop sharing to protect their position.

It came (at least in part) from a document in May where the CCP pretty much said that they will need to review models to make sure they don't threaten national security.

Which basically translates too "Don't give away tools that can be used to undermine your own goals".

Lots of fake news out there, but you don't need to speculate any longer. Key takeaways from President Xi's speech in his first ever appearance at the World AI Conference in Shanghai:

- Started the speech by referring to his signature maxim, "great changes unseen in a century are unfolding across the world"

- Said that the world has "entered an unprecedented period of active innovation on AI technology", which means "great opportunities as well as challenges for governance”

- reaffirmed commitment to open source to promote AI "openness and win-win"

- warns against "over stretching" the concept of national security as applied to AI where one country's national security is prioritised over others

- China opposes emergence of “new historical injustices” in AI (one of the most strongly worded parts of the speech)

- China in next 5 years will provide 5000 opportunities to developing countries in "AI training and seminar programmes" and "cooperation centres" - names ASEAN, League of Arab States, African Union, CELAC, SCO and BRICS

Live blog: https://www.scmp.com/tech/policy/article/3360858/chinas-xi-j...

Complete translation: https://x.com/i/status/2077984062933762450

So much for the speculation that China was encouraging the release of free/cheap models to mess with the US AI economy.
It's 2.8T, I'm sure they will open the weights but it will only be able to be run on very high end machines.
This does seem like a cash grab. These token rates are crazy. I'll just use GPT 5.6 thanks.