6 ms·
I would'nt touch nvidia "goodies" with a barge pole. Perfectly good 5-year GPU deprecated upon upgrade to ubuntu 22.04 because of cuda/drivers. The lock-in trap
by nologic01 3y ago
I would'nt touch nvidia "goodies" with a barge pole. Perfectly good 5-year GPU deprecated upon upgrade to ubuntu 22.04 because of cuda/drivers. The lock-in trap of the century.
- Foobar8568 3y ago>Please use CUDA 11.4 Which should be compatible with any Kepler late architecture, as in 6xx models from 10years ago+?
- cburdick13 3y agoThat's right, we've tested down to pascal, but this should work on Kepler too since CUDA and the underlying libraries support it.
- downrightmike 3y agoI have a spare Titan Z that could be useful if so.
- choppaface 3y agoWill you publish benchmarks for e.g. K80? Or provide a way for users to contribute? It's really handy to know, e.g. comparable to what is Resnet50 inference on a bunch of architectures.
- cburdick13 3y agoHi, what specifically are you looking to benchmark on the K80? Users are free to contribute and we've had many external PRs. Contribution guide is here: https://github.com/NVIDIA/MatX/blob/main/CONTRIBUTING.md https://github.com/NVIDIA/MatX/blob/main/CONTRIBUTING.md
- choppaface 3y agowell the root README.md has some "benchmarks" at the very top. maybe there's no existing benchmarks doc with more details? or even a repo of what's in the README? it's like if numpy is 5 sec on CPU, but a K80 is only 100ms, then the K80 cost might be worth not going for the A100 at 3ms. Similar argument for jetson.
- cburdick13 3y agoWe have benchmarks in the benchmarks directory, but these are for things like convolution, matrix multiples, etc. It's not for running a traditional benchmark set like resnet. Like most benchmarks it really depends on what you want to do, and since it's a general library everyone might care about different things.
- choppaface 3y agoI saw those, yes having the code is great but I’m more interested in the actual numbers. E.g. what does this sample do on A100, P100, T4, K80, Jetson nano? The analog is things like Resnet50 get tested and reported and then you know if Resnet50 might work for your budget / hardware. Versus the advert in the root Readme, which is impressive but gives no data on the pareto.
- cburdick13 3y agoHi, if you don't mind opening an issue asking for this we can run these and put in the readme.
- dheera 3y agoThe real issue is having to keep multiple copies of CUDA on a system to satisfy all the different moving parts that each want different CUDA versions and CUDNN versions, and those different CUDA installers fight over and uninstall each others' nvidia driver versions because CUDA is shipped with the nvidia driver, bundled and listed as an apt dependency. My system changes from nvidia-525 to nvidia-535 to nvidia-520 to nvidia-515 on a daily basis because I need to reinstall a different CUDA version just to try some new paper's code. PyTorch did it right, it now ships with its own CUDA and doesn't take a shit about version what you have in /usr/local. Everything else should do the same.
- kristjansson 3y agoDon’t install the drivers with CUDA? Don’t rely on distro-packaged CUDA?
- dheera 3y agoNVIDIA's official CUDA is packaged with their damn driver.
- cburdick13 3y agoHi, you can choose not to install the driver with a CUDA install, and to download the driver separately.
- dheera 3y agoOnly if you use the runfile, but that breaks yet other packages that specify 'cuda' as an apt dependency. Your 'cuda' packages should use '>=' not '==' e.g. 'cuda-12' should depend on nvidia>=525 NOT nvidia==525
- cburdick13 3y agoThat's a fair point. I'll look into it.
- smoldesu 3y agoAre you talking about for desktop/GUI use? A lot of older Nvidia hardware has been implicitly depreciated for years. Without newer patches to help Wayland support, your options are either to use nouveau or stick with unsupported software and official drivers. It's ugly, but Microsoft and Apple pretty much depreciate hardware the same way; first with 'recommended' cutoff points, then hard depreciations. On Linux it's less spelled-out, but things seem to be Wayland-or-bust now.
- reaperman 3y ago*deprecate or “EOL” (for continuing support). Depreciate is a tax and accounting concept, though often at the end of a MACRS depreciation schedule office equipment is often immediately replaced since it’s no longer contributing to tax write-offs. (Really, just not replaced as long as it continues to contribute to tax deductions)
- cburdick13 3y agoIt was mentioned elsewhere, but we support down to cuda 11.4, which supports down to the Maxwell architecture (nearly 10 years old now).
- arthurcolle 3y agocan you downgrade?
- deleted 3y ago[deleted]
- choppaface 3y agoSurprisingly this repo is BSD licensed so it might even outgrow nvidia. Eigen is pretty strong but this new MatX syntax might catch on, especially with easy GPU integration.