3 ms·
Sorry for being lazy, but is there a rough breakdown like "You get sonnet level for M5 and Opus for M5 pro, etc.", or is it still speculative. Or put simpler, d
by vadansky 1mo ago
Sorry for being lazy, but is there a rough breakdown like "You get sonnet level for M5 and Opus for M5 pro, etc.", or is it still speculative. Or put simpler, do you get Opus level for the 256GB M5 Max?
- beernet 1mo ago[dead]
- rogerkirkness 1mo agoOpus is probably ~2T parameter model, so that would probably not run on these. More like Sonnet.
- root_axis 1mo agoSonnet is estimated around 1T, so that is far beyond what's practical as well.
- c0rruptbytes 1mo agoThe 512GB could run GLM 5.3 which is Opus level
- lostmsu 1mo agoGLM 5.2 in NVFP4 is 465 GB. It would be a tough fit.
- root_axis 1mo agoFor local LLMs with a Mac, rule of thumb is you always want an Ultra (due to memory bandwidth). Even an M1 Ultra is superior to an M6 Pro in this regard. There are no configurations even close to running something comparable to frontier model variants, they're simply far too large, but something like full precision Qwen 35b or DeepSeek 70b at 50+ t/s is well within available configuration, and potential for plenty of room for large context sizes.
- dannyw 1mo ago256GB is enough for DSv4 Flash, expect maybe ~30tg/s, and a lot better profile. I'm using Flash heavily, and I would describe it as nearly as intelligent as Sonnet-class in agentic coding, but more usable. Less world knowledge of course, and definitely a bit less intelligent; but not _that_ much. On usability: Takes less handholding, less likely to make unsolicited refactors or whatever, and the writing style is readable. It's not great at super-long-horizon goals as the Claude 5 models are; but if you have a good harness, you can get around that.
- deleted 1mo ago[deleted]
- twobitshifter 1mo agoWith the newest Qwen 3.8 27B model you can get opus on just about any new Mac.
- mrheosuper 1mo agoIt can fit into 16GB macbook ?
- twobitshifter 1mo agoIt can but not well, you really want at least 24 GB and preferably 64 GB for the full model https://ofox.ai/blog/qwen-3-8-27b-run-locally-vram-gguf-2026/ https://ofox.ai/blog/qwen-3-8-27b-run-locally-vram-gguf-2026...
- dexterlagan 1mo agoI ended up getting an M5 Max MBP with 48GB and it runs Qwen3.8 in Q4. I developed 3 apps so far, medium complexity, with no issue using OpenCode. This is probably the minimum setup for comfortable local agentic coding in my book.