4 ms·
How much work would it be to use the C++ ONNX run-time with this instead of Python? Is it a Claudeable amount of work? The iOS version is Swift-based.
by fwsgonzo 7mo ago
How much work would it be to use the C++ ONNX run-time with this instead of Python? Is it a Claudeable amount of work?
The iOS version is Swift-based.
- rohan_joshi 7mo agoshouldn't be hard. what backend/hardware are you interested in running this with? i'll add an example for using C++ onnx model. btw check out roadmap, our inference engine will be out 1-2 weeks and it is expected to be faster than onnx.
- fwsgonzo 7mo agodesktop CPUs running inference on a single background thread would be the ideal case for what I'm considering.
- koolala 7mo agoI want to run it in a website with Wasm and having the browser do the audio playback
- ilnmtlbnm 6mo agoI've been playing with running small models in browser tabs for some time, and finally decided to open some of it. Added kitten (nano only, for now, will move on to mini) to my "web tts thing": https://github.com/idle-intelligence/tts-web https://github.com/idle-intelligence/tts-web demo: https://idle-intelligence.github.io/tts-web/web/ https://idle-intelligence.github.io/tts-web/web/