State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX
A free, open source, and extensible speech-to-text application
Automatic Speech Recognition with Word-level Timestamps
Python Audio Analysis Library: Feature Extraction, Classification
Multilingual speech recognition and audio understanding model
Automagically synchronize subtitles with video
Python library for audio and music analysis
Label Studio is a multi-type data labeling and annotation tool
Swing Music is a beautiful, self-hosted music player
Privacy-focused video surveillance software for Windows & Linux
Speech-to-text, text-to-speech, and speaker recognition
Lightweight Windows utility for switching audio
Improved AudioBookConverter based on freeipodsoftware release
Automated Music Discovery and Collection Manager
Tiny data-over-sound library
BizHawk is a multi-system emulator written in C#
Infrastructure to enable deployment of ML models
Cross-platform, customizable ML solutions
Robust Speech Recognition via Large-Scale Weak Supervision
A gallery that showcases on-device ML/GenAI use cases
EPUB to audiobook converter, optimized for Audiobookshelf
Have a natural, spoken conversation with AI
A Web UI for easy subtitle using whisper model
Cross-Platform C++ 2D/3D game engine
A python tool that uses GPT-4, FFmpeg, and OpenCV