Industrial-level controllable zero-shot text-to-speech system
Vim Win32 Installer
Generate audiobooks from EPUBs, PDFs and text with captions
Pycorrector is a toolkit for text error correction
Open source annotation tool for machine learning practitioners
Official inference repo for FLUX.1 models
Framework for building realtime multimodal voice AI agents apps
OCR software, free and offline
CLIP, Predict the most relevant text snippet given an image
Mozc - a Japanese Input Method Editor designed for multi-platform
Powerful Android AI agent with tools, automation, and Linux shell
State-of-the-art TTS model under 25MB
Label Studio is a multi-type data labeling and annotation tool
Python & command-line tool to gather text on the Web
Cut videos with a text editor
SoTA open-source TTS
Tokenizer-Free TTS for Multilingual Speech Generation
Open source plain text editor designed for writing novels
Faster Whisper transcription with CTranslate2
Qwen-Image is a powerful image generation foundation model
A simple native web interface that uses ChatTTS to synthesize text
Robust Speech Recognition via Large-Scale Weak Supervision
State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX
Offline Text To Speech synthesis for python
Generate audiobooks from e-books