logoalt Hacker News

stymaarlast Sunday at 7:44 AM2 repliesview on HN

> Kimi K3 reproducibly identifies itself as Claude

It could also be have been trained from collected response datasets. Claude got caught several time responding it was ChatGPT or even Deepseek and I don't think Anthropic has been distealling DeepSeek.

> This behavior is exactly what you'd expect from a model distilled from Claude.

The opposite actually. If they wanted to distill Claude without getting caught they could just use a regex to change Claude to Kimi in their distillation pipeline!


Replies

dotancohenlast Sunday at 8:09 AM

  > distealling
Apt typo.

Though I am of the opinion that distilling is no different than how extant frontier LLMs have also been trained on other people's data, I could actually see the word distealling becoming useful in discussion.

show 2 replies
cobbzillalast Sunday at 2:37 PM

> they could just use a regex to change Claude to Kimi in their distillation pipeline!

Jean-Kimi Van Damme would like to have a word with you.

show 1 reply