3 ms·
Please elaborate the mechanisms by which a LLM would know what model it is.
by Paradigma11 1mo ago
Please elaborate the mechanisms by which a LLM would know what model it is.
- deleted 1mo ago[deleted]
- datadrivenangel 1mo agoask it what type of model it is or what it's name is... it's weird that Kimi will say it's claude...
- tshaddox 1mo agoOkay, but I would also ask “why does Claude say that its name is Claude?”
- Gigachad 1mo agoBecause it's in the system prompt
- fwn 1mo agoAFAIK, you cannot run Claude without one of their mandatory system prompts at all. If we wanted to compare model responses, we would give all models system prompts with model names, thereby fixing the Kimi misattribution. The reason Kimi often states its name as Claude is likely because we can actually run it without the mandatory system prompt, smoothing over awkward competitor mentions.
- InvertedRhodium 1mo agoThe training data likely references Claude significantly more often than Kimi, given the popularity of the models. There will simply be more examples of “Claude” being the response to that question.
- NekkoDroid 1mo agoDoesn't Claude say its Deepseek when asked in Chinese? I remember there being posts about that a while ago.
- Paradigma11 1mo agoUnless it is specifically instructed in the system prompt it will give you the most likely answer, which is Claude. If there are instructions in the system prompt it will give the correct answer, which would be Kimi. Or are you suggesting that Kimi is copy pasting Claudes system prompt?
- Sabinus 1mo agoBy being trained on text containing "I am X" in the model response section.