5 ms·
I don't understand AMD in this. Isn't it insanity that they're not throwing all they've got at their software stack?
by AYBABTME 3y ago
I don't understand AMD in this. Isn't it insanity that they're not throwing all they've got at their software stack?
- tempaccount420 3y agoHardware people don't get along very well with software people.
- elbear 3y agoWhy's that?
- imtringued 3y agoBecause they didn't go to uni when hardware-software-codesign was being taught.
- nebula8804 3y agoWhat unis would that include? Isn't ATI Canadian? Therefore i'd expect lots of UToronto and Waterloo people there. Aren't they some of the best in this field?
- rapsey 3y agoBecause it is a different type of engineering. If you manage software development like you manage hardware development your software is going to be bad. That has always been AMD's problem and it is not likely to get fixed.
- AYBABTME 3y ago2t$ problem of egos?
- deleted 3y ago[deleted]
- deleted 3y ago[deleted]
- roenxi 3y agoYou know what happens to companies that panic and throw all their resources into knee-jerk software projects? I don't, but I'd predict it is ugly. Adding more people to a bad project generally makes it worse. The issue that AMD has is they had a long period where they clearly had no idea what they were doing. You could tell just from looking at websites, CUDA pretty much immediately gets to "here is a library for FFT", "here is a library for sparse matricies". AMD would explain that ROCM is an abbreviation of the ROCm Software platform or something unspeakably stupid. And that your graphics card wasn't supported. That changed a few months ago; so it looks like they have put some competent PMs in the chair now or something. But it'll take months for the flow on effects to reach the market. They have to figure out what the problems are which takes months to do properly; then fix the software (1-3 months more minimum); then get it into the open and the foundational libraries like PyTorch pick it up (might take another year). You can speed that up, but more cooks in the kitchen is not the way. Bandwidth use needs to be optimised. It isn't like ROCm seems lacks key features; it can technically do inference and training. My card crashes regularly though (might be a VRAM issue) so it is useless in practice. AMD can check boxes but the software doesn't really work and grappling with that organisationally is hard. Unless you have the right people in the right places, which AMD didn't have up to at least mid 2023.
- elcomet 3y agoPytorch has been supporting rocm for all last 2 years
- Certhas 3y agoLook at AMDs vs Intel. They have now surpassed Intel in terms of CPUs sold and market cap. That was unthinkable even six, seven years ago. It makes perfect sense that, organisationally, they were focused on that battle. If you remember the Athlon days, AMD beat Intel before, but briefly. It didn't last. This time it looks like they beat Intel and have had the focus to stay. Intel will come back and beat them some cycles, but there is no collapse on the horizon. So it makes sense that they started looking at nVidia in the last year or so. Of course nVidia has amassed an obscene war chest in the meantime...
- imtringued 3y agoYou have to remember that this only applies to cheap consumer GPUs, they tend to support their datacenter GPUs better. When you consider that Ryzen AI already eats the AI inference lunch, having better GPUs with better software only threatens to cannibalize their data center GPU offering. Given enough time nobody will care about using AMD GPUs for AI.
- wmf 3y agoYet we've heard about nobody doing training on AMD.
- logicchains 3y agoIt's a political problem. Good software engineers are paid more than good hardware engineers, but AMD management is unwilling to pay up to bring on good software engineers because then they'd also need to pay their hardware engineers more, otherwise the hardware engineers would be unsatisfied. If you check NVidia salaries online you'll see NVidia pays significantly more than AMD for both hardware and software engineers; it's a classic case of AMD management being penny-wise, pound-foolish.
- dheera 3y agoThis is quite possibly also why Boeing is having issues. If they paid everyone $1M/year salaries maybe more people would consider going into aerospace engineering. Right now though Boeing's starting salaries aren't that much higher than what an Uber driver in the bay area makes.
- e4325f 3y agoThey did buy Nod.ai recently