3 ms·
The Ryzen AI Max 395 128gb is super cool, but not fast for inference. Order of magnitude slower than dedicated GPU but at half the cost. You can run larger mode
by simple10 5mo ago
The Ryzen AI Max 395 128gb is super cool, but not fast for inference. Order of magnitude slower than dedicated GPU but at half the cost. You can run larger models on it but it's slow. Great for local async work. Not great for daily chat or code agent driver.
- throwa356262 5mo agoThe latest NPUs are pretty fast, I think what is missing is more optimised software support.
- plagiarist 5mo agoThe vRAM bandwidth is at least as much a problem as compute on these ones, there is a lot of data to shuffle around