4 ms·
There is a really interesting startup in Prague that is doing just that. They fine-tuned Qwen 3.6 27b to have 46% fewer reasoning tokens while maintaining most
by kamranjon 2mo ago
There is a really interesting startup in Prague that is doing just that. They fine-tuned Qwen 3.6 27b to have 46% fewer reasoning tokens while maintaining most of the performance characteristics. I'm interested to see if they continue down this path of optimizing reasoning for other models.
https://bottlecapai.com/post/thinkingcap-qwen3-6-27b/ https://bottlecapai.com/post/thinkingcap-qwen3-6-27b/