4 ms·
Thanks! We had to rewrite lots of code to optimize for the hot-swap load speeds. So we prioritized llama as it was the most popular group on Hugging Face. And
by pico_creator 2y ago
Thanks!
We had to rewrite lots of code to optimize for the hot-swap load speeds. So we prioritized llama as it was the most popular group on Hugging Face. And RWKV (which is what we work on in open source space)
But other architectures are coming within a week or two. Up next is probably Mixtral MoE models.
We will keep adding until we add ALL the architectures and models =)