State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX
A free, open source, and extensible speech-to-text application
Automatic Speech Recognition with Word-level Timestamps
Python Audio Analysis Library: Feature Extraction, Classification
Multilingual speech recognition and audio understanding model
Python library for audio and music analysis
Label Studio is a multi-type data labeling and annotation tool
Swing Music is a beautiful, self-hosted music player
Privacy-focused video surveillance software for Windows & Linux
Speech-to-text, text-to-speech, and speaker recognition
Improved AudioBookConverter based on freeipodsoftware release
Automated Music Discovery and Collection Manager
Tiny data-over-sound library
Lightweight Windows utility for switching audio
BizHawk is a multi-system emulator written in C#
Cross-platform, customizable ML solutions
Robust Speech Recognition via Large-Scale Weak Supervision
Have a natural, spoken conversation with AI
A Web UI for easy subtitle using whisper model
A python tool that uses GPT-4, FFmpeg, and OpenCV
Cross-Platform C++ 2D/3D game engine
Deep Learning API and Server in C++14 support for Caffe, PyTorch
Complete HomeKit integration for UniFi Protect with full support
A nearly-live implementation of OpenAI's Whisper
Unofficial (Golang) Go bindings for the Hugging Face Inference API