4 ms·
You can use my new golang inference engine to run variants of Qwen 3.5 faster than llama.cpp: https://github.com/computerex/dlgo https://github.com/computerex/d
by computerex 7mo ago
You can use my new golang inference engine to run variants of Qwen 3.5 faster than llama.cpp:
https://github.com/computerex/dlgo https://github.com/computerex/dlgo