6 ms·
Please elaborate the mechanisms by which a LLM would know what model it is.
by Paradigma11 23d ago
Please elaborate the mechanisms by which a LLM would know what model it is.
- datadrivenangel 23d agoask it what type of model it is or what it's name is... it's weird that Kimi will say it's claude...
- tshaddox 23d agoOkay, but I would also ask “why does Claude say that its name is Claude?”
- Gigachad 23d agoBecause it's in the system prompt
- fwn 23d agoAFAIK, you cannot run Claude without one of their mandatory system prompts at all. If we wanted to compare model responses, we would give all models system prompts with model names, thereby fixing the Kimi misattribution. The reason Kimi often states its name as Claude is likely because we can actually run it without the mandatory system prompt, smoothing over awkward competitor mentions.
- InvertedRhodium 23d agoThe training data likely references Claude significantly more often than Kimi, given the popularity of the models. There will simply be more examples of “Claude” being the response to that question.
- NekkoDroid 23d agoDoesn't Claude say its Deepseek when asked in Chinese? I remember there being posts about that a while ago.
- Paradigma11 23d agoUnless it is specifically instructed in the system prompt it will give you the most likely answer, which is Claude. If there are instructions in the system prompt it will give the correct answer, which would be Kimi. Or are you suggesting that Kimi is copy pasting Claudes system prompt?
- deleted 23d ago[deleted]
- Sabinus 23d agoBy being trained on text containing "I am X" in the model response section.