Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
fishingboy
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
fishingboy
4y ago
Yes, it is the sum of all models. But M6-10T (MoE) of Alibaba is a 10 trillion one.
2.
▲
Google has the most big pre-trained models while Alibaba has the largest
(openbmb.github.io)
4 points
by
fishingboy
4y ago
|
2 comments
3.
▲
by
fishingboy
4y ago
As far as I am concerned, there are many ways to compress a model such as quantization, pruning, and knowledge distillation. By the way, I found a package called BMCook when I browsed the OpenBMB repo, which implements several algorithms an
4.
▲
BMList – A list of big pre-trained models (GPT-3, DALL-E2...)
(github.com)
55 points
by
fishingboy
4y ago
|
3 comments