Interface for OuteTTS models
Open-source multi-speaker long-form text-to-speech model
super expressive prompting model based on ltx2.3
TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning
A Web UI for easy subtitle using whisper model
Self-hosted AI audio transcription
MOSS‑TTS Family open‑source speech and sound generation model
Instantly generate AI-powered subtitles on your device
Audio Plugin for Audio to MIDI transcription using deep learning
High-Quality Voice Cloning TTS for 600+ Languages
Web presentation editor replicating many PowerPoint features online
The official Node.js / Typescript library for the Groq API
The official Python Library for the Groq API
One-click deployment (including offline integration package)
Curated AI engineering notes on LLMs, generative models, and tools
MARS5 speech model (TTS) from CAMB.AI
Rust framework for building modular and scalable LLM-powered apps
Foundational model for human-like, expressive TTS
Synchronized Translation for Videos
Audiocraft is a library for audio processing and generation
Towards Human-Level Text-to-Speech through Style Diffusion
A python package to analyze and compare voices with deep learning
Music research software
TensorFlow Implementation of DC-TTS: yet another text-to-speech model