3 ms·
We’ve just released Neuphonic TTS Air, a lightweight open-source speech foundation model under Apache 2.0. The main idea: frontier-quality text-to-speech, but
by neuphonic 1y ago
We’ve just released Neuphonic TTS Air, a lightweight open-source speech foundation model under Apache 2.0.
The main idea: frontier-quality text-to-speech, but small enough to run in realtime on CPU. No GPUs, no cloud APIs, no rate limits.
Why we built this:
- Most speech models today live behind paid APIs → privacy tradeoffs, recurring costs, and external dependencies.
- With Air, you get full control, privacy, and zero marginal cost.
- It enables new use cases where running speech models on-device matters (edge compute, accessibility tools, offline apps).
Repo: https://github.com/neuphonic/neutts-air https://github.com/neuphonic/neutts-air
Would love feedback from HN on performance, applications, and contributions.
- makkes 1y ago"the model has a 2048 tokens limit. This includes the reference text/phones as well as the reference and generation audio tokens." https://github.com/neuphonic/neutts-air/issues/15#issuecomment-3368454690 https://github.com/neuphonic/neutts-air/issues/15#issuecomme... So "no rate limits", while true, is kind of setting different expectations.