Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ipiszy
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
ipiszy
4y ago
For now it's for single GPU inference only.
2.
▲
by
ipiszy
4y ago
RTX 3080-10GB should work. You could check https://github.com/facebookincubator/AITemplate/tree/main/ex... , and https://www.reddit.com/r/StableDiffusion/comments/xv7m89
3.
▲
by
ipiszy
4y ago
Yes this is correct. batch 16 7.9s / 25 steps, per image 0.49s: it generates 16 images for each prompt within 7.9s, so it's 0.49s per image.
4.
▲
by
ipiszy
4y ago
AITemplate only supports fp16 data types with fp16 or fp32 accumulation right now. We are working on supporting more data types and quantization. We don't have an official comparison between AITemplate and tvm / onnx for now, but
5.
▲
by
ipiszy
4y ago
We have a bunch of unittests and E2E tests to compare numeric numbers between AITemplate and PyTorch eager.
6.
▲
by
ipiszy
4y ago
You could check "AITemplate optimizations" section in the blog ( https://ai.facebook.com/blog/gpu-inference-engine-nvidia-amd... ), and https://github.com/facebookincubator/AITemplate#more-
7.
▲
by
ipiszy
4y ago
As @haolu7 mentioned, you could take a pre-trained model and use AITemplate to do model inference. All you need to do is to re-write the model using AITemplate frontend and map PyTorch params to AITemplate params. Besides, AITemplate has a
8.
▲
by
ipiszy
4y ago
tl;dr: Meta is open sourcing AITemplate, an inference engine for both Nvidia and AMD GPUs. Code: https://github.com/facebookincubator/AITemplate . AITemplate delivers much better perf (1.9x ~ 12.8x) compared to PyTorch
9.
▲
by
ipiszy
14y ago
I like HackerNode more because: 1. It contains more columns than only "FrontPage", such as "Jobs" and "Comments"; 2. The UI seems more concise. Although there are not indent between comments, the text is much clearer. However, neither of th
10.
▲
by
ipiszy
15y ago
...............................