Robust Speech Recognition via Large-Scale Weak Supervision
Multilingual speech recognition and audio understanding model
Speech-to-text, text-to-speech, and speaker recognition
OpenVINO™ Toolkit repository
Omnilingual ASR Open-Source Multilingual SpeechRecognition
Cross-platform AI language practice app
Replace OpenAI GPT with another LLM in your app
Han Language Processing
Underthesea - Vietnamese NLP Toolkit
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
Build your own AI friend
A full spaCy pipeline and models for scientific/biomedical documents
Enhances Tesseract OCR output using LLMs (local or API)
The first AI that can earn its own existence, replicate, and evolve
A GUI Agent app based on UI-TARS to control your computer using AI
Toolkit for conversational AI
Open source AI VTuber platform with voice chat and Live2D avatars
A PyTorch-based Speech Toolkit
Fast multimodal LLM for real-time voice interaction and AI apps
The no-nonsense RAG chunking library
Research and application of technologies such as nl processing
A pure Javascript Multilingual OCR
Recognition and resolution of numbers, units, date/time, etc.
Training data (data labeling, annotation, workflow) for all data types
The media player for language learning, with dual subtitles