Recognition and resolution of numbers, units, date/time, etc.
Open Source OCR Engine
Awesome multilingual OCR toolkits based on PaddlePaddle
Speech-to-text, text-to-speech, and speaker recognition
Robust Speech Recognition via Large-Scale Weak Supervision
Handwritten Text Recognition (HTR) system implemented with TensorFlow
OCR software, free and offline
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
A cross-platform software for text translation and recognition
Contexts Optical Compression
A pure Javascript Multilingual OCR
An Open-Source Toolkit for General-OCR Research and Applications
Library for OCR-related tasks powered by Deep Learning
Speech recognition module for Python
Documents and exposes generated source for menu-bar workflows
A free, open source, and extensible speech-to-text application
Open-Source Python3 tool for recognizing layouts, tables, and math
Automatic Speech Recognition with Word-level Timestamps
Cross-platform AI language practice app
Open source semantic search and text analytics for large document sets
Faster Whisper transcription with CTranslate2
Audio foundation model excelling in audio understanding
Enhances Tesseract OCR output using LLMs (local or API)
Open-source industrial-grade ASR models
A full spaCy pipeline and models for scientific/biomedical documents