Open Source OCR Engine
Awesome multilingual OCR toolkits based on PaddlePaddle
Robust Speech Recognition via Large-Scale Weak Supervision
State-of-the-art 2D and 3D Face Analysis Project
Speech-to-text, text-to-speech, and speaker recognition
A Lightweight Face Recognition and Facial Attribute Analysis
Port of OpenAI's Whisper model in C/C++
OCR software, free and offline
Multilingual speech recognition and audio understanding model
Speech recognition module for Python
Captcha solver extension for humans
Handwritten Text Recognition (HTR) system implemented with TensorFlow
kaldi-asr/kaldi is the official location of the Kaldi project
High-Performance Face Recognition Library on PaddlePaddle & PyTorch
A pure Javascript Multilingual OCR
Contexts Optical Compression
Open-source industrial-grade ASR models
A PyTorch-based Speech Toolkit
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
On-device Speech Recognition for Apple Silicon
Audio foundation model excelling in audio understanding
An Open-Source Toolkit for General-OCR Research and Applications
Speech recognition for your site
Automatic Speech Recognition with Word-level Timestamps