4 ms·
Here is a video demo of the project: https://youtu.be/aIg4-eL9ATc?si=66ynl4Mlci9v76rU https://youtu.be/aIg4-eL9ATc?si=66ynl4Mlci9v76rU
by Alyx1337 3y ago
Here is a video demo of the project:
https://youtu.be/aIg4-eL9ATc?si=66ynl4Mlci9v76rU https://youtu.be/aIg4-eL9ATc?si=66ynl4Mlci9v76rU
- alchemist1e9 3y agoNice work! Very impressed. Do you happen to know anything about any open source voice identification software? I’ve noticed with ChatGPT voice and any other voice driven assistant that a massive problem is the background voices and noise. One solution could be advanced pre-processing to ID your voice only. Another idea I’ve had is using something professional with PTT: https://sheepdogmics.com/products/quick-disconnect-mic-tubeless-earpiece-kenwood https://sheepdogmics.com/products/quick-disconnect-mic-tubel...
- Alyx1337 3y agoThanks! I don't know a lot about this but someone shared this local voice assistant in the comments: https://github.com/KoljaB/LocalAIVoiceChat https://github.com/KoljaB/LocalAIVoiceChat Could be a good lead
- alchemist1e9 3y agoYeah github.com/KoljaB is quite a collection of stuff! I agree. It all seems your vision of JARVIS, which I share completely but haven't accomplished what you have, again excellent work and thank you for sharing, is very attainable. Probably combining your work along with KoljaB is very promising.
- Alyx1337 3y agoThank you very much!
- Jayakumark 3y agoCheck whether this can help https://github.com/resemble-ai/resemble-enhance/tree/main https://github.com/resemble-ai/resemble-enhance/tree/main
- visarga 3y agoGoogle Gemini was trained on audio and can generate audio directly. Whatever you build now will be replaced by a much better version soon.