Hacker News new | ask | show | jobs
by throw10920 3 days ago
> if china can train K3 on a fraction of the US compute availability, yet it benchmarks almost equivalent to Fable for a third of the cost, it’s game over for US labs in the long run.

I agree. However, as of yet, most/all leading PRC models are distilled from US models. I've personally observed Deepseek, GLM, and Kimi all respond that they are Claude when asked, and the networks of tens of thousands of proxy accounts that we've found show that it's happening on a large scale.

But - if the PRC labs actually train up the domain expertise to train those models from scratch, which they are in the process of doing - then the US is cooked. They're not there yet, but it's probably only a matter of years.

1 comments

Playing devils advocate, does from scratch really matter if all frontier labs are training off each other anyways? Practically speaking, businesses/consumers just care about lowest inference cost for maximizing intelligence anyways (not to mention Anthropic/OpenAI forcing KYC/litigation barrier trash for access to any cybersecurity/IT capabilities) that I literally just cancelled Claude today.

Yeah it’s sad the CCP has my information. But it’s either them or the feds, and at least Chinese models actually work for cybersecurity tasks, not to mention aren’t stupid expensive

Artificialanalysis.ai rn on opus 5 is a joke. The main intelligence benchmarks it is like 1% better than Fable, but the cost difference between that and K3 is so funny lol. It’s the same with cars— you don’t have to do it from scratch, as long as you can do it cheaper and with the same quality, hence Toyota/Honda taking over

Yeah it’s sad that American models are censored more than Chinese ones lol (outside of asking them about the CCP) but it’s where we are at I guess :(