Miso TTS is an 8 billion, highly emotive text-to-speech model
MiniMax H3 inference engine for Mac computers
Robust Speech Recognition via Large-Scale Weak Supervision
Data manipulation and transformation for audio signal processing
Industrial-level controllable zero-shot text-to-speech system
A tool for transcoding lossless audio files
Audio codecs extracted from Android Open Source Project
TorchMultimodal is a PyTorch library
A Conversational Speech Generation Model
Music Player Daemon SACD/DVD-A ISO decoder plugins
Implementation of NÜWA, attention network for text to video synthesis
A Very Low-Bitrate Codec for Speech Compression
State-of-the-art deep learning based audio codec
Real Time Speech Enhancement in the Waveform Domain (Interspeech 2020)
Modern 3D engine and IDE written using C# and C++.
Facebook AI research's automatic speech recognition toolkit
An Open Source alternative to SBR (Spectral band replication)
Open source speech models for Julius in English and other languages.
MPEG1 Video Decoder in JavaScript
Decoding of OOK signals transmitted by a weather sensor on 433,82 Mhz
The simplest audio player based on FFmpeg
Library for Arduino based R/C equipment
JavaScript audio decoding framework