31 ms·
This is awesome. Thanks for pushing the audio pareto frontier forward. Probably far fetched for now, but I think the next big evolution is building the pareto/
by karimf 19d ago
This is awesome. Thanks for pushing the audio pareto frontier forward.
Probably far fetched for now, but I think the next big evolution is building the pareto/much cheaper alternative to GPT-Live-1.
The STT/TTS market is quite saturated, while today, there's almost no cheap/open source alternative to GPT-Live-1.
- 6879346626 19d agoIs this really something people want? Honestly you can properly lower the pricing at least by 50%+. Getting something conversationally better has been done, the tool calling will likely be worse though. The infrastructure for real time is really annoying though.
- toebee 18d agoif we can lower the pricing by not 50% but 10x, then I think it would be something people want. we are taking the bet that OSS models will take a huge chunk of market share not just in LLMs but in multimodal as well
- 6879346626 18d agoIf audio only then 10x is 100% possible right now based on math of the services. + Video is unlikely unless they are willing to give up margins. Didn't take long at all for Google to smash out with 50% lol
- toebee 19d agoagreed. we've been doing some work around NVIDIA personaplex 7b, but its quality is quite far from GPT-Live-1, esp in terms of intelligence. Once a good OSS model is out, we'll be sure to be the first to serve it cheaply to the masses :)