3 ms·
Agreed, 1000tok/s just fills up the context window (which is big by 2004 standards) super fast. But seems like 5.3-spark was just a taste of what’s to come.
by beering 3mo ago
Agreed, 1000tok/s just fills up the context window (which is big by 2004 standards) super fast. But seems like 5.3-spark was just a taste of what’s to come.
- taneq 3mo ago2004 standards? O.o
- partsch 3mo ago1904
- bogeholm 3mo agoBack when we were kids, we would get 0 tokens/sec _if we were lucky_
- mlinsey 3mo agoIn 2004, I took a class where we trained "language models" that were bigram word models, on an archive of a couple years of the Wall Street Journal. I remember someone who literally announced they were dropping the class to the whole room at the end of a lecture, saying "This isn't AI!!!"