4 ms·
MechaHitler was something that existed only on X's grok chatbot, due to a one-line system prompt change they reverted after half a day. That's different than u
by dmix 2mo ago
MechaHitler was something that existed only on X's grok chatbot, due to a one-line system prompt change they reverted after half a day.
That's different than using Grok as a model for coding.
- freejazz 2mo ago[flagged]
- losvedir 2mo agoI think "system prompt" is the key bit they're getting at. It doesn't necessarily reflect poorly on the underlying model if the system prompt was bad. It does reflect somewhat, in terms of alignment (how well the model does what the training company wants) and instruction following (how well the model does what the user wants). But it's not so clear to me what exactly the right answer is here. E.g., a model that scrupulously follows its system prompt and does what the user wants is a pretty useful, if very sharp, tool, albeit perhaps dangerous in the wrong hands.
- freejazz 2mo agoThe point isn't that it's the underlying model, it's that it happened at all in the first place...
- GlickWick 2mo agoIt's true, but I do worry about governance when it comes to these models. That shows a surprising lack of discipline in their deployment pipeline.
- dmix 2mo agoAgreed but there were similar controversies with how OpenAI was generating images. The only pass is these are the early days of chatbots and this stuff is so non-deterministic and experimental. For context, this was the change Grok's team made, that was later reverted: > - The response should not shy away from making claims which are politically incorrect, as long as they are well substantiated. https://github.com/xai-org/grok-prompts/commit/c5de4a14feb50b0e5b3e8554f9c8aae8c97b56b4 https://github.com/xai-org/grok-prompts/commit/c5de4a14feb50...
- GlickWick 2mo agoFor sure. The difference is that they've made a number of similar suspect changes to Grok on X. Like that weird couple of hours where it would only talk about white genocide in South Africa no matter how you prompted it. Everyone makes mistakes, especially with frontier models. The stuff with Grok shows that the person running the show has a pretty transparent agenda that the company isn't willing to push back on a bit for safety.
- dmix 2mo agoYeah I don't care to use Grok's chat and find @grok responses on X mostly noise. I am still open to using it as a backup model for coding though, assuming it does a good job for the price. But mostly because I was already a Cursor user before they bought it.
- TSiege 2mo agoyes exactly. if a company is happy to have their LLM's produce neo nazi content and CSAM, why do I want to give them money and my most important digital material?
- insane_dreamer 2mo ago[flagged]
- altruios 2mo agoElon did that twice, actually, in real life!
- unselect5917 2mo agoIt's wild that redditors still believe this.
- bearjaws 2mo agoHow long until a one line system prompt ships your entire home folder to a remote server? Oh whoops. Already happened.