2 ms·
Yeah, the community always seems to figure out how to do things more effective. My girlfriend asked me if I could transcribe some audio files for her with my "
by can_center_divs 4y ago
Yeah, the community always seems to figure out how to do things more effective.
My girlfriend asked me if I could transcribe some audio files for her with my "programming stuff". I immediately thought of Whisper from OpenAI.
I first used the official CLI tool. With the largest model it took long 8 hours to transcribe a 30 min long file. I noticed it was running on the CPU - tried switching it to use the GPU instead with no luck. Running it on WSL was probably not helping.
Then I found this gem: https://github.com/Const-me/Whisper https://github.com/Const-me/Whisper
A C++ Windows implementation of Whisper. I opened the program, fed it with the largest model and the file. The transcript was done in 4 minutes, instead of 8 hours... Downside? The program has a GUI, lol.
Of course, I could probably get the CLI tool to run on the GPU with some tinkering and installing some Nvidia packages for Whisper to use. But frankly, I have so little experience with that kind of stuff, that installing the Windows implementation was a much easier choice.
- deleted 4y ago[deleted]