5 ms·
After using Claude for a long time, I tested Sol 5.6 for the first time today. Love it, its an incredibly capable model and uses far fewer tokens/time thinking.
by netsec_burn 2mo ago
After using Claude for a long time, I tested Sol 5.6 for the first time today. Love it, its an incredibly capable model and uses far fewer tokens/time thinking. Its what I imagine Fable would be if I haven't been downgraded on every conversation - even after completing the verification program. I think I may cancel my Claude subscription finally.
- Razengan 2mo agoI recently tried Claude again after several months, to see if it was any better at something Codex has been struggling with… They STILL don't have an option to "Sign in with Apple" on the website, but they do for Google??!? (and on iPhone of course) Screw that asinine UX (and no it wasn't better than Codex at this particular task)
- nozzlegear 2mo agoThat was an issue at least a year ago. I had signed up for a claude account on my iPhone and then wanted to sign in on my laptop but nope, not possible. Insane they still haven't fixed it. Can somebody at Anthropic tag claude in slack or whatever goofy shit you do and ask it to add Apple OAuth to your website? Clearly humans aren't testing it.
- Razengan 2mo agoI signed up on iOS, Sign In with Apple, cause I don't go around giving random companies my actual email if I can help it and sure enough, I was right to do so: They don't even let you remove your payment method afterwards. Every other store, Steam etc., lets you. No way I have enough trust to install their desktop app after that, so I just want to try it through their website.. Can Sign In with Google, but not with Apple so you gotta open the Passwords app, copy your random email, paste into the website, then copy the OTP from your email.. It's been that way for at least a year The desktop app was clunky too the last couple times I tried it a few months ago and the AI itself hasn't been that hot compared to ChatGPT/Codex either: https://i.imgur.com/jYawPDY.png https://i.imgur.com/jYawPDY.png So all the Claude hype posted on HN seems like a case of the emperor with no clothes to me (P.S. The thing I just now tried to do on Claude hit the weekly usage limit after 2 minutes)
- solenoid0937 2mo agoThere is no Claude hype on HN, HN has more vitriol for Claude than any online forum I've seen
- datakan 2mo agoCan't change your email address with Claude either. You have to delete your whole account and create a new one. Absolute shit software and they expect us to believe these things are super powerful world changing things.
- Aargau 2mo agoI'm also in Anthropic Cyber Verification Program, but they specifically exclude Fable, just goes up to Opus 5. I hear you on the downgrades, I'm 13/13 on downgrades, and last downgraded me to Sonnet for asking for reasoning chain.
- ec109685 2mo agoFable is still the best there is. Sol close second but I find it gets way to stuck on details. Also Opus 5 is fine if your codebase is simple.
- curreylabs 2mo agoAnd you arent writing English
- FL410 2mo agoFable feels less cumbersome to work with, but it is SO DAMN ANNOYING with the refusals that I'm leaning more and more on Sol, and very much looking forward to GPT6. Just seems like Anthropic is trying their hardest to ruin their reputation and user experience.
- paxys 2mo agoRemember, Dario knows best
- teaearlgraycold 2mo agoI’ve never had it refuse anything. Even vulnerability searching in my codebase.
- ChadNauseam 2mo agoYeah I feel like I'm living in a different dimension than these people. I wonder what they're working on. I've literally never had it refuse everything and I max out my 20x plan every week
- SyneRyder 2mo agoCan I ask which country you're in? I have a theory that the safeguards differ depending on the user's country. I'm in Australia, and Fable downgrades to Opus when testing for bugs in memory in a legacy C code base. If Fable starts taking initiative and writes a test case that involves writing to a null pointer, that's the end of the conversation.
- ChadNauseam 2mo agoThat's interesting, thanks for sharing. Maybe the classifiers are really sensitive to anything that has to do with memory safety.
- FL410 2mo agoAs one random example, today I had it hit a refusal loop when adding a country selector dropdown to a form, presumably because it contained a “bad” country name? I hit refusals at least 2-3 times per day, sometimes many more. The worst part is it is often right in the middle of a multi-stage task, so the only option is really to switch to Opus 5 and let it defecate its absurdly verbose comments all over the rest of the edits in the turn and hope it doesn’t go on one of its tangents, then have Sol do damage control. Oh and I got approved for their “cyber verification program” blessing, which comically does absolutely nothing for Fable.
- jchw 2mo agoI think Fable's dominance is overstated. It definitely has the lead, but quantifying what that lead actually is is really hard. I'm using GPT 5.6 Sol to do some shit that I personally would consider "crazy" - low level undocumented hardware driver alchemy, reverse engineering highly obfuscated code, even a bit of screwing around with a rendering engine in Vulkan, really just about the most complex tasks I can get any model to do, and it does great. For the more advanced stuff, it definitely needs the effort bumped. But even with the effort bumped, the token usage really doesn't seem to skyrocket too badly until at least you hit xhigh and max, which really only seem to be necessary if you are doing genuine crazy stuff, so it's not that bad. I did similar stuff with Fable. In fact, I went directly from an Anthropic subscription with Fable to an OpenAI subscription with Sol, more or less, and it really felt pretty seamless. If anything, I was thrilled to realize how much I actually preferred Codex CLI, to the point where I started using it at work too. Fable seems to be generally more impressive at outputting one-shot web apps. I'm not really saying that to try to downplay what Fable can do, it's just that if I compare the two, this is one of the few definitely noticeable areas that you can easily demonstrate. Obviously, one-shotting programs is much better as a demonstration of a model's capabilities than it is practically useful (not that it is useless, but hopefully my point is understood). However, whatever Fable truly is better at, one thing I really like about GPT 5.6 Sol is even harder to quantify: taste. GPT 5.6 Sol outputs are still LLM outputs and they contain many things that people would probably consider "Claude-isms" for better or worse, but overall I really prefer the GPT 5.6 Sol output. I find it to be generally more tasteful. Hard to quantify, but when talking to people I've had enough people seemingly agree with me to convince me that it really is true.
- andrewingram 2mo agoI used Sol to extract the remaining decryption keys from the Super Mario Maker 2 (Switch) game files. Someone had previously extracted all the keys from the original release, but not any of the new ones from updates. Not only did it succeed, but it helped me understand the data sufficiently to add support for “Super World” rendering to my level viewer (which I made back in 2021), eg the little widget at the top of https://www.smm2-viewer.com/players/B16-306-GVG https://www.smm2-viewer.com/players/B16-306-GVG I was very pleasantly surprised to find Sol wasn’t obstructive over what was clearly a very grey area endeavour.
- tamimio 2mo agoYeah at this point claude is overrated, overly expensive, weird writing style (elliptical), and the worst part is the aggressive guardrails that even normal convos get interrupted, meanwhile openAI is still I would say at the normal balance, if you ask something too obvious or direct it will stop you other than that, it work flawlessly, plus, I have yet to hit the limit despite heavily using it these past weeks.
- ChadNauseam 2mo ago> if you ask something too obvious or direct it will stop you other than that, it work flawlessly What on earth are you asking it?
- bbg2401 2mo agoYou've replied incredulously to a similar stated experience in this thread already and proceeded to ignore the follow-up. Why are you again asking a question to which you have no intention to field an answer?
- ChadNauseam 2mo agoI left both comments 12 hours ago. The first reply to my other comment was 11 hours ago.
- kyxsc 2mo agoSol is way too eager to hone in on small details and ends up with massive over-engineering. Fable does it too - to be fair - but noticeably less. After extensively using both on Max 20x plans, I've concluded that Fable is better for problem solving and coding, whereas Sol 5.6 Ultra shines in debugging specific issues: tackle a problem with Fable then leverage Sol to clean up, double check, or fix specific issues. Fable (imo) had the edge on the $200 plan, but after this 50% reduction I'd say Codex is better value by far and there's no contest. --- Using Fable as the orchestrator and delegating tasks to Sol 5.6 Ultra via the codex plugin in Claude Code yielded good results, but still there was a lot more over-engineering (thus time and tokens spent) than Fable by itself would've done. Both models suffer from doing-too-much. But both models are fundamentally really smart and knowledgeable. I think it's really close and pricing cuts really spice things up for us consumers! Sol is a clear winner in the value department and the $100 plan is enticing! --- *Claude Code usage is reducing by 33% in 2 days, Wednesday August 19... cmon anthropic: clau.de/cc-50-promo
- greenavocado 2mo agoSol w/ Effort -> Low
- kyxsc 2mo agoIt's great, don't get me wrong, but so is Fable. I'm just comparing the long-horizon task performance between the two at the same or similar effort levels. Given the 50% discount on Sol and how smart it is, yeah it's unprecedented value. If you only want to use low effort, there's a clear winner here on value and it's not even close!
- iJohnDoe 2mo agoInteresting the use of Max and Ultra. I don’t doubt the complexity, but would someone use Max or Ultra on Typescript or Go, for example? Is it more about just avoiding any mistakes? Seems like that would be costly when medium or high would work fine?
- kyxsc 2mo ago
- Kye 2mo agoSol has held stuff for a while to do the same sort of hazard checks I assume Fable is doing, but it always releases them. I think that's the better way to handle it rather than preventing me from seeing how far I can get generating schematics to use in Minecraft. Currently: a mostly normal voxel house.
- znnajdla 2mo agoSol is my daily driver but there are still times I reach for Fable when Sol doesn’t cut it. Just yesterday for example, I was trying to build a self-modifying hot-reloaded agent harness in Elixir for fun and Sol just kept doing silly things like thin wrappers and unnecessary abstractions. Fable handled the task elegantly. Sol is really good as a reviewer for finding bugs due to its thoroughness however.
- matheusmoreira 2mo agoI too switched to OpenAI after I got sick of Anthropic's constant "safety" downgrades. Sol is definitely a breath of fresh air. > even after completing the verification program Was it easy to complete it? I ended up in some weird state where I can't even attempt the verification at all. Opened the Persona tab once, closed it and then it never opened ever again. It says a verification precheck failed. Even without TAC, Sol doesn't seem to get blocked very often. Fable would downgrade to Opus if I looked at it wrong.
- petesergeant 2mo agoI have a "strategy / life-coach" project, and was surprised at how much better Sol is than Fable on it, as I've found Fable to have the edge for most things for me so far. But Sol: questions were better, insight was better, it got the brief better.
- yieldcrv 2mo agoClaude as a harness at all really spends too much time before giving user feedback Its a crutch that is no longer competitive
- vdntp 2mo agoyou should check out the codex desktop app. people who've been using claude code for a long time will surely be surprised.
- boorang 2mo agoI was pleasantly surprised to find that the GPT models are much stricter in adhering to my AGENTS.md guidelines and heuristics than Claude.
- saidnooneever 2mo agosol is much better imho than Fable but i can understand if they will perform wildly different for different people with different levels of expertise aswell as different needs. I dislike fable myself it doesnt really work for me. Sol also doesnt _really_ work but it sort of tricks me into thinking it does more convincingly :p. cancelled my subscriptions few days ago. (was on 100$ ones, not sure if there is diff in quality for higher tiers or not.. there might be that too). what i hate the most is that they will make any obvious mistake you do not tell them to avoid. then on the next plan to fix it, your token limit is hit at step 4/5 -_-. Both models seem incredibly good at that mostly... for tasks outside of coding and program design i do find them quite useful. like devops crap. maybe because i hate that, i like their help there more.
- freepiai 1mo agoOkay I'm a bit late to this, but my side project is www.freepi.ai, it's free ad+training powered inference in a Pi Harness. So not sure if that something you'd be down to try but if you did and have feedback I'd love to try and make it work for you!
- leokennis 2mo ago5.6 Sol is a joy to use for "daily chat" as well. Compared to earlier OpenAI models it catches and corrects its mistakes very reliably. It also seems way smarter in tuning its replies to areas I am more/less knowledgeable about (i.e. when I ask it a law question, it assumes I know as much as a toddler which is true, but on political topics it more easily throws around terminology) and including analogies. On medium thinking, it's a very good compromise between speed and quality.
- shibaprasadb 2mo agoI cancelled my subscription recently and moved to Sol. So far - it has been a great experience. The only aspect where Fable/Claude is better I feel is doing some research from the web and summarising the facts.
- nutribeatApp 2mo agoI generally prefer Fable but in my experience Sol is a much better web researcher
- impulser_ 2mo agoIt's the complete opposite for me. The model might be the worst model I have ever used when compared to other models in the class. You just can't get it not to just write the most enterprise over complex over engineered solutions for every little thing you ask it to do. It the first model to actually make me pissed off to use AI. I absolutely hate the model so much. I don't even want to see the codebases this model is fucking up. It might just be good at finding bugs that about it. That all I would ever use it for just because it works harder than Claude models.
- jiaosdjf 2mo agoThe final straw for Claude was its refusal to give me a list of the most recent rapes reported by the BBC and basic information about them (location, date, names, just things reported in mainstream media). It outright REFUSED to complete this task. I will not be told what I can and can't do by AI and I will no longer be supporting American companies run by despicable people. GPT only gets my money right now because its so fast and cheap but I'll be back to Chinese models in no time.
- arcrs 2mo agooutput token efficiency bruv. u never go wrong with it
- nurettin 2mo agoUsed claude since 7/2025. Switched to codex after fable got blocked. It was still 5.5 but I knew they had to come up with something. As soon as I switched, wow. It wasn't super intelligent, but it was stable. Every day it was the same performance. This consistency is definitely worth paying for.
- justinbaker84 2mo agoI switched from claude to gpt when 5.6 came out for the same reasons. I don't understand why so many people are still using claude when GPT and open source are so much better.
- johnnyApplePRNG 2mo agoReally? I have witnessed 5.6 Sol Ultra edit line after line of literally empty lines ... for hours. I wasn't literally watching it, I came back to a goal (that it started for itself without my approval!) that had done nothing but that for some reason. It couldn't explain why it had started.
- synergy20 2mo agoindeed, I'm switching from claude to codex
- kingkongjaffa 2mo agoI also cancelled Claude recently. GPT 5.6 models are really good, and they have much better usage limits.
- jamesponddotco 2mo agoFable is the only one that follows my instructions correctly, which I find quite important. It takes STYLE.md and SPEC.md as law, and code just like I would code, with the same mistakes and all. I just can't get Opus (Opus 5 is dumb as a rock, to be fair) or Sol to do that, so I exclusively use Fable for personal work. When I reach my weekly limit, usually on the last day close to the reset, I just go back to coding by hand ¯\_(ツ)_/¯ Heck, it even does security reviews and fixes, as long as I don't ask it to "attack" the codebase. I'm planning on using Kimi or GLM for that part.