gitaskhub

Neural audio models in C++20. Text-to-speech, speech-to-text, speaker diarization, the RAVE audio autoencoder, keyword spotting, and in-tree English G2P.

Stars · 10
Language · C++
License · MIT
Ask anything about this repo to start.

By chatting or signing in you agree to the Terms and chat-message logging (revocable in History).