Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
srinifromsalem
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
srinifromsalem
8mo ago
Nice work on the speech-to-speech pipeline! You're absolutely right that it has to go through the text intermediate step - that's actually where a lot of the interesting processing can happen. I've found that the speech->t
2.
▲
by
srinifromsalem
8mo ago
For local speech-to-text, Whisper remains the gold standard - you can run it locally with good accuracy across languages. For speech-to-speech, you'd typically chain Whisper with a local TTS model like Coqui TTS or use something like T