2 ms·
Not a fair comparison really. If you can run a model locally then you can somewhat train out the guardrails, censorship, and brand-safety. That has value a subs
by chasd00 19d ago
Not a fair comparison really. If you can run a model locally then you can somewhat train out the guardrails, censorship, and brand-safety. That has value a subscription does not.
Idk about the quality of this setup but just pasting it here as an example.
https://explainx.ai/blog/heretic-llm-abliteration-guide-2026 https://explainx.ai/blog/heretic-llm-abliteration-guide-2026
- ChickeNES 19d ago> If you can run a model locally then you can somewhat train out the guardrails, censorship, and brand-safety. When does the average person actually need to do that?
- jerf 19d agoWe've already seen frontier models refuse to answer almost any question that touches on computer security and be very likely to kick out biology and chemistry questions even if they aren't all that close to breeding dangerous viruses or making explosives. I expect this is only going to get worse. "Censorship" isn't just going to be about who you vote for and which political party the model will say nice things about and which it is more likely to say bad things about. It's going to become about whether the hoi polloi are allowed to have effective AIs at all. Like the 1990s internet, AI has outrun a lot of power structures but that is not going to continue indefinitely.
- ChickeNES 19d agoSo you want to remove valid safeguards? And stop misusing the word censorship.
- jerf 19d agoYou asked a question. I gave you an answer. I seriously doubt that if you and I sat down together at a table and banged on this for an hour that we would come to the same definition of "valid". Ask 10 people, get 12 answers to that question. There's going to be a lot of motte & bailey in the next couple of years, where I just want an AI to answer questions about whether my code is vulnerable and people like you will be "Oh so you want an AI that can hack the Pentagon do you?" and it doesn't look like we're going to be seeing eye to eye on that one.
- beachy 19d agoI just got some kind of cyber alert from Claude and was forced back down to Opus while I was trying to connect to a battery I own via bluetooth. So I can certainly understand why someone would want the guardrails gone.
- ChickeNES 19d agoWell I just applied to their cyber program, got accepted in two hours, haven't had that issue since. Ditto OpenAI. Why people treat these companies like sports teams instead of compute providers I have no idea, when I see underpriced compute, I take advantage of it.
- kees99 19d ago"Need" might be a bit too strong, but I do want overly obnoxious guardrails not to stand in the way. Case in point, last week I was poking Opus 5 into writing me some RPi-pico firmware for driving a small e-paper screen. Font was built in right into C code as hex constants. Space being tight, I asked if there is some clever compression that could be applied. Claude thought for good 10 minutes, then guardrail kicked in telling me that was "cyber", and refused to continue.
- ChickeNES 19d agoJust apply to the cyber program? I got in in around 2 hours, and I'm just some hobby hacker, not some paid security consultant. It is a valid point though, I'm doing tons of systems and embedded stuff and was hitting the safe guards with Claude and Codex before getting into their cyber programs (hex REALLY triggered Claude in particular, which was amusing).
- ASalazarMX 18d agoIf you mean retraining, it's not even needed anymore. If you want the guardrails off, these days you just install LMStudio and download an abridged model. It's all GUI. The abridged models might have weird behavior in edge cases after the pruning, though.
- xnx 19d agoThat can also be done with neoclouds.