2 ms·
It's like 50% realtime on a 3090, not quite real time on a 4090. You can also use smaller models and it's a bit faster. 3080 is same speed, you don't need the
by JonathanFly 3y ago
It's like 50% realtime on a 3090, not quite real time on a 4090. You can also use smaller models and it's a bit faster.
3080 is same speed, you don't need the extra memory.
- JonathanFly 3y agoActually I was just checking, and Bark isn't that close to maxing out GPU utilization. Running two instances on a 3090 seems like a throughput increase and the models fit. Update: And getting weird CUDA issues. Hmn...
- woodson 3y agoJust to add a datapoint: the main audioLM based models (not the BERT embedding part) fully utilize an RTX 2080 Ti.
- JonathanFly 3y agoMust be some low hanging fruit to optimize in Bark. It would be somewhat close to realtime if it was close to 100% and scaled linearly.