3 ms·
Yeah, seems a bit odd because the TensorRT-LLM repo lists Turing as supported architecture. https://github.com/NVIDIA/TensorRT-LLM?tab=readme-ov-file#precision
by operator-name 3y ago
Yeah, seems a bit odd because the TensorRT-LLM repo lists Turing as supported architecture.
https://github.com/NVIDIA/TensorRT-LLM?tab=readme-ov-file#precision https://github.com/NVIDIA/TensorRT-LLM?tab=readme-ov-file#pr...