Please elaborate the mechanisms by which a LLM would know what model it is.
By being trained on text containing "I am X" in the model response section.
ask it what type of model it is or what it's name is... it's weird that Kimi will say it's claude...
The training data likely references Claude significantly more often than Kimi, given the popularity of the models. There will simply be more examples of “Claude” being the response to that question.
Doesn't Claude say its Deepseek when asked in Chinese? I remember there being posts about that a while ago.
Okay, but I would also ask “why does Claude say that its name is Claude?”
Because it's in the system prompt
By being trained on text containing "I am X" in the model response section.
ask it what type of model it is or what it's name is... it's weird that Kimi will say it's claude...
The training data likely references Claude significantly more often than Kimi, given the popularity of the models. There will simply be more examples of “Claude” being the response to that question.
Doesn't Claude say its Deepseek when asked in Chinese? I remember there being posts about that a while ago.
Okay, but I would also ask “why does Claude say that its name is Claude?”
Because it's in the system prompt