4 ms·
Heh, funny to see this popup here :) The performance on Apple Silicon should be much better today compared to what is shown in the video as whisper.cpp now run
by ggerganov 3y ago
Heh, funny to see this popup here :)
The performance on Apple Silicon should be much better today compared to what is shown in the video as whisper.cpp now runs fully on the GPU and there have been significant improvements in llama.cpp generation speed over the last few months.
- A4ET8a8uTh0 3y agoYou are kinda famous now man. Odds are, people follow your github religiously.
- MuffinFlavored 3y agoIs ggerganov to LLM what Fabrice Bellard is to QuickJS/QEMU/FFMPEG?
- actionfromafar 3y agoThat's a big burden to place on anyone.
- boesboes 3y ago13 minutes between this and the commit of a new demo video, not bad :D And impressive performance indeed!
- tomtom1337 3y agoIs it just me, or is the gpu version actually slower to respond?
- tomtom1337 3y agoAh, forget the other message, I watched the videos in the wrong order! And I can’t delete or edit using the Hack app!
- deleted 3y ago[deleted]
- asadm 3y agoI have sent a PR to move that new demo to the top. I think the new demo is significantly better.
- v3ss0n 3y agowill this work with latested distilled llama?
- deleted 3y ago[deleted]
- sgt 3y agoIs running this on Apple Silicon the most cost effective way to run this, or can it be done cheaper on a beefed up homelab Linux server?