4 ms·
I've used only Qwen3.5 so far for work and was, after initial struggles, successful with GPU setup, no mlx. Ngl the fact that they are using `presence_penalty:
by proxysna 5mo ago
I've used only Qwen3.5 so far for work and was, after initial struggles, successful with GPU setup, no mlx. Ngl the fact that they are using `presence_penalty: 0` and no `max_tokens` is weird after that exact setup caused me "initial struggles", but i've set up a simple docker-compose with vllm and qwen3.6 right now to test it out and it worked perfectly fine for me.
Gist with the compose and example of an output.
https://gist.github.com/meaty-popsicle/f883f4a118ff345b430c3b97edc7b64f https://gist.github.com/meaty-popsicle/f883f4a118ff345b430c3...