Hacker News new | ask | show | jobs
by zius 1 day ago
In my first interaction ("hi there kimi k3!"), Kimi K3 identified twice out of three times as Claude:

> Hi there! Quick note — I'm actually Claude, made by Anthropic, not Kimi. But no worries!

https://imgur.com/a/jqpc2Jc

and

> Just a quick heads-up — I'm Claude, made by Anthropic, not Kimi! Kimi is a different AI assistant (made by Moonshot AI), so it looks like there might be a little mix-up.

https://imgur.com/a/AKxeysH

9 comments

FWIW Claude sometimes identifies as DeepSeek when asked in Chinese: https://x.com/stevibe/status/2026227392076018101
I remember when Claude identified as ChatGPT a long time ago. It proves nothing else than that there is a lot of training material on the internet with Claude as the AI.
I don't understand why labs don't include correct model name in the training process. Almost nobody seems to be willing to tell their models who they actually are.
If the model works correctly, you don't need to include the model name in the training process. You just add in the system prompt "You are X, trained by Y" and the model will claim to be that. That also allows users to "white-label" the outputs sort to say.

This approach is basically how all "models know who they are" (they don't actually typically "know" that at all), it's just a system prompt instruction in the platform you use.

But I would expect they would at least use post-training to discourage identifying as the wrong company.
A lot of API customers don't want their chatbot or agent to identify as Claude or GPT, they want it to identify as <customisable marketingname>.
That's probably the true reason.
That only makes sense if you don't sell access to the model for 3rd party product development. One of those products is chat personas/agents. The company integrating the model will want it's own identity.
You can put anything in the prompt. And if you have bare model that you don't give any prompt from the start and want to find out which model you are talking to, you are out of luck.

I think Qwen teaches its models that they are Qwen. Most others don't bother.

Because they do not know the name of the model before they train it. There is also distillation, where multiple models will be trained from a larger one. E.G. Sonnet was promoted to Opus at one point after it surpassed expectations.
How many times do people need to point out that every model has this behavior until this stops being posted?
This happens with other models too - Gemini often identifies as ChatGPT for me, confusing many a debugging attempt
It's distillations all the way down!
Models don't know who they are. Stop asking this question to the model as a source of truth.
Models don't have an inherent identity. It should be obvious by now that every models trains on public AI chat session transcripts. I've seen Claude say it's Qwen.
The meme/trope of China copying everything really keeps playing into itself
Could be on purpose to disguise as a US made model
Bootleg AI
I can see how, if all you ever were exposed to was that link and this post, you might think they are equivalent, and the conclusion is "whatever, everyone is stealing from each other.

But the answer is no, Claude didn't train itself on ChatGPT, but Kimi K3 did train itself on Claude.

There has been no accusation that I am aware of by OpenAI against Anthropic, (which are in the same jurisdiction, so it would amount to legal action). On the other hand there have been accusations with extensive detail by Anthropic on how chinese models are attacking Anthropic to reverse engineer and copy their product.

There's nuance if you care to see it, but maybe it's easier to pretend that everyone is ripping each other off so that you can consume a ripoff without seeing yourself as at fault.

point is everyone scrapes data from the internet. Do you think there is at least one page on the internet that mentions Claude as an AI? Well, then it's part of the training material.
That’s how a smoking gone turns into a water pistol
when open source models are banned maybe i will become a cartel kingpin smuggling open source weights into the usa. find a sufficiently shifty street corner, 'what do you need', "kimi". are you familiar with my product? pure mxfp4, $500 per TB.

(joke and walter mitty, i know the government reads my messages)

if this qualifies bootleg, point me at non-bootleg frontier ai.