Fast multimodal LLM for real-time voice interaction and AI apps
Python Audio Analysis Library: Feature Extraction, Classification
A suite of advanced multi-modal LLMs
Virtual AI anchor that combines state-of-the-art technology
Toolkit for audio, music, and speech generation
A python package to analyze and compare voices with deep learning
Toolkit for efficient experimentation with Speech Recognition
A fast GPU accelerated feature extraction software for speech analysis
CTC-based forced aligner for audio-text in 158 languages