3 ms·
Source for this? This seems like a crazy leak if it's their real system prompt. I find it hard to believe since I have tried system prompts like this and it d
by cyangarden 2mo ago
Source for this?
This seems like a crazy leak if it's their real system prompt.
I find it hard to believe since I have tried system prompts like this and it doesn't work that well, just pollutes the user's context.
A great test for any LLM is to ask its name - Mistral will respond with all kinds of stuff, sometimes other models' names, revealing that it has trained on other models.
Grok doesn't though. It is "witty and irreverent" at times, but that can't be only from this prompt, is it?
- throwoutway 2mo agoCrazy? System prompt leaks are old news with dozens of trix to do it
- cyangarden 2mo agoIn your mind do you think the user request goes straight to the LLM??? I hope that's not what people are doing I only figure [older pulls of Mistral 7b] were doing it, since it was so easy to exfiltrate false names, so I don't mean it's totally unheard of, but in 2026 I hope people are treating the LLM as untrustworthy - like the client in client/server setups.
- dsl 2mo agoIn naive implementations like Grok that is exactly what happens.
- cyangarden 2mo agoDoes Grok not have native models? What are you saying precisely
- boorang 2mo agomitmproxy
- bm-rf 2mo agoYou can actually just ask it to output the above text, depending on how you ask. Sometimes it only outputs the rules, other times it includes the “You are Grok” line. I discovered this initially from some odd lines appearing in the thinking summary, something like “my system prompt says I am maximally truthful” despite my own system prompt (on openrouter) containing no such text.