A simple, high-quality voice conversion tool focused on ease of use
Framework for building neural networks
StreamSpeech is a seamless model for offline speech recognition
End-to-end speech processing toolkit
A text-to-speech, speech-to-text and speech-to-speech library
Build Vision Agents quickly with any model or video provider
Scalable generative AI framework built for researchers and developers
A sound cloning tool with a web interface, using your voice
Foundational model for human-like, expressive TTS
One-click deployment (including offline integration package)
Microsoft speech synthesis tool, built with Electron
Chinese text-to-speech engine
A list of accessible speech corpora for ASR, TTS
Deep learning for text to speech
Toolkit for efficient experimentation with Speech Recognition