Speech-to-text, text-to-speech, and speaker recognition
Buzz transcribes and translates audio offline
Let the Xiaoai speaker "hear your voice"
ComfyUI integration for Microsoft's VibeVoice text-to-speech model
Interface for OuteTTS models
Ultraminimalist macOS recording + transcription
VoiceStudio is the open-source, fully-local ElevenLabs alternative
Open-source multi-speaker long-form text-to-speech model
macOS System-wide audio equalizer & volume mixer
super expressive prompting model based on ltx2.3
A Web UI for easy subtitle using whisper model
A native macOS menu bar app for managing audio device priorities
Official PyTorch Implementation
Self-hosted AI audio transcription
Music Assistant is a free, opensource Media library manager
MOSS‑TTS Family open‑source speech and sound generation model
TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning
Googles NotebookLM but local
An Open Source implementation of Notebook LM with more flexibility
Control SONOS speakers from your terminal
The ioquake3 community effort to continue supporting/developing id's
The HTML Presentation Framework
Open source software for live streaming and recording
SpatGRIS4
Audio Plugin for Audio to MIDI transcription using deep learning