3 ms·The RTX 4090 has 330.3 TOPS with INT8 precision, so for inference workloads it is still a magnitude fasterby mosshammer 3y agoThe RTX 4090 has 330.3 TOPS with INT8 precision, so for inference workloads it is still a magnitude faster