3 ms·Is this similar to fastllm? https://github.com/ztxz16/fastllm https://github.com/ztxz16/fastllmby nogajun 3mo agoIs this similar to fastllm? https://github.com/ztxz16/fastllm https://github.com/ztxz16/fastllmvikmals 3mo agofastllm targets the GPU, while colibri uses CPU inference onlyaliljet 3mo agoI'd be curious about an.option that would allow glm use with a low end GPU like a 2080 ti...