3 ms·
We are doing fine so far with just Transformers, we have just crossed Trillion parameter models, we don't know if 100s of Trillion param model wont be better.
by subhajeet2107 2mo ago
We are doing fine so far with just Transformers, we have just crossed Trillion parameter models, we don't know if 100s of Trillion param model wont be better.
- WarmWash 2mo agoThe breakthroughs would be mostly in optimizations for computing tokens, not as much in expanding parameter count (if anything we want to shrink that).