LLM-based Reinforcement Learning audio edit model
Your Hardcore Loop Machine.
API samples for the Universal Windows Platform.
The open-source voice synthesis studio powered by Qwen3-TTS
Generate music based on natural language prompts using LLMs
Sonic Pi is your free code-based music creation and performance tool
Tokenizer-Free TTS for Multilingual Speech Generation
A gallery that showcases on-device ML/GenAI use cases
Sample code and notebooks for Generative AI on Google Cloud
Digital Signal Processing in Python, by Allen B. Downey
Interface for OuteTTS models
The official Python SDK for the ElevenLabs API
ComfyUI integration for Microsoft's VibeVoice text-to-speech model
StreamSpeech is a seamless model for offline speech recognition
Git extension for versioning large files
A reference client implementation for the playback of MPEG DASH
Play SoundTracker media on your computer.
Delphi translations of the MS Media Foundation and related API's
Plays iMelody (IMY) files using many sound systems
The purpose of the project is to develop audio processing algorithms
Edge TTS Desktop turns text into speech through edge-tts.
A GUI application to show the tree structure of a RIFF file
Code examples for the new features of iOS 9