4 ms·
The performance of oolama on my M1 MAX is pretty solid - and does things that my 2070 GPU can't do because of memory.
by InTheArena 2y ago
The performance of oolama on my M1 MAX is pretty solid - and does things that my 2070 GPU can't do because of memory.
- dangus 2y agoNot that I don’t believe you but the 2070 is two generations and 5 years old. Maybe a comparison to a 4000 series would be more appropriate?
- Kirby64 2y agoThe M1 Max is also 2 generations old, and ~3 years old at this point. Seems like a fair comparison to me.
- dangus 2y agoThe 4000 series still has a bigger gap in how much of a generational leap that product was. The M3 Max has something like 33% faster overall graphics performance than the M1 Max (average benchmark) while the 4090 is something like 138% faster than the 2080Ti. Depending on which 2070 and 4070 models you compare the difference is similar, close to or exceeding 100% uplift.
- whizzter 2y agoGoogling power draw the 4090 goes up to 450w whilst the 2080ti was at 250w, adjusting for power consumption the increase is somewhere around 32%. Some architectural gains and probably optimizations in chipset workings but we're not seeing as many amazing generational leaps anymore regardless of manufacturer/designer.
- dangus 2y agoI’m still seeing over a 100% uplift comparing mobile to mobile on Nvidia products: https://gpu.userbenchmark.com/Compare/Nvidia-RTX-4090-Laptop-vs-Nvidia-RTX-2080-Mobile/m2036852vsm700275 https://gpu.userbenchmark.com/Compare/Nvidia-RTX-4090-Laptop... As far as desktop products, power consumption is irrelevant.
- talldayo 2y agoMaybe it's controversial, but I don't think comparing 5nm mobile hardware from 2021 is a fair fight against 12nm desktop hardware from 2018. And still, performance-wise, the 2070 still wins out by a ~33% margin: https://browser.geekbench.com/opencl-benchmarks https://browser.geekbench.com/opencl-benchmarks
- chessgecko 2y agoFor this comparison the generation of chip doesn’t really matter because the llm decode (which is the costly step) barely uses any of the perf and just needs the model weights to fit in memory
- JudasGoat 2y agoI found it interesting that the Apple M3 scored nearly identical to the Radeon 780M. I know the memory bandwidth is slower but you can add 2 32gb sodimms to the AMD APU for short money.
- Teever 2y agoWell, you know that it would still be able to do more than a 4000 series GPU from Nvidia because you can have more system memory in a mac than you can have video ram in a 4000 series GPU.
- dangus 2y agoYes, obviously I’m aware that you can throw more RAM at an M-series GPU. But of course that’s only helpful for specific workflows.