Free, high-quality text-to-speech API endpoint to replace OpenAI
Apache-2.0 open-source image generation and editing model family
Voice Recognition to Text Tool
A Powerful Native Multimodal Model for Image Generation
Easy-to-use and powerful NLP library with Awesome model zoo
Interface for OuteTTS models
Qwen-Image is a powerful image generation foundation model
Powerful Android AI agent with tools, automation, and Linux shell
Open image model at the forefront of design
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Self-host the powerful Chatterbox TTS model
Towards Human-Sounding Speech
A simple, high-quality voice conversion tool focused on ease of use
A Model Context Protocol (MCP) server
Tokenizer-Free TTS for Multilingual Speech Generation
Persian NLP Toolkit
tiktoken is a fast BPE tokeniser for use with OpenAI's models
Agent harness to make your slop code well-engineered and beautiful
Management of Yandex Station and other smart home devices
TextWorld is a sandbox learning environment for the training
Implementation of Imagen, Google's Text-to-Image Neural Network
Easily compute clip embeddings and build a clip retrieval system
Collection of Gemma 3 variants that are trained for performance
Instant voice cloning by MIT and MyShell. Audio foundation model
An Open Source text-to-speech system built by inverting Whisper