Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
matisqe
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
matisqe
4y ago
Novak's mental game and tenacity is incredible - can't think of better traits for a startup founder. Great article! I wonder how he would compare to Federer or Nadal as founders :).
2.
▲
by
matisqe
4y ago
We are offering both Speech Synthesis (/TTS) and Voice Lab (Rapid Voice Cloning and Voice Design) as a standard SaaS model (w/ fixed quota of characters you can voice per month). API is directly available on the platform. Outside
3.
▲
by
matisqe
4y ago
Exactly! Only issue is having a well-labelled dataset with those type of cues. We have an idea on how to do it though!
4.
▲
by
matisqe
4y ago
(ElevenLabs dev here) The generative voices and the way they sound is very much a function all the training data, sampling and interpolation as you also pointed out. As a lot of these do involve deep breaths, that why synthesized voice will
5.
▲
by
matisqe
4y ago
Hey! ElevenLabs here - not public yet, but we will be opening up Beta later this month. API is available directly in the platform.
6.
▲
by
matisqe
4y ago
Hey - ElevenLabs dev here. The quality above works with <1s latency that for some real-time apps is already sufficient. On smaller chunks of text it can be as quick as ~500ms.
7.
▲
by
matisqe
4y ago
Hey! ElevenLabs dev here - yes exactly! We do rapid voice cloning (just on few seconds samples) that for American accents works really well - which is already available in Beta. We can also do a professional near-identical copy with longer
8.
▲
by
matisqe
4y ago
ElevenLabs dev here - we believe this is a 2 step process and agree it is needed! First, we want to the quality you get out-of-the-box to already by brilliant by taking context into account. Granted, that gets you sometimes 98% there and ar