An Open Source implementation of Notebook LM with more flexibility
Audio Plugin for Audio to MIDI transcription using deep learning
A Python library for audio
A GPU-accelerated library containing highly optimized building blocks
Build cross-modal and multimodal applications on the cloud
The Triton Inference Server provides an optimized cloud
Deep learning for text to speech
Library of deep learning models and datasets
Speech recognition software for English & Polish languages
PyTorch implementation of convolutional neural networks
Music research software
Cross Audio-Visual Recognition using 3D Architectures
A fast GPU accelerated feature extraction software for speech analysis
Implementation of duration high-order hidden Markov model in Matlab.