4 ms·
I’d like to see more of these models. I’ve been using diffusion Gemma and it is very fast on GPUs in output token/sec. In the diffusion Gemma whitepaper, they
by gdiamos 1mo ago
I’d like to see more of these models.
I’ve been using diffusion Gemma and it is very fast on GPUs in output token/sec.
In the diffusion Gemma whitepaper, they say they could have done better with more time and compute.
Even with those caveats, it is very uses-able as a local model.