Robust Speech Recognition via Large-Scale Weak Supervision
Multilingual speech recognition and audio understanding model
Speech-to-text, text-to-speech, and speaker recognition
OpenVINO™ Toolkit repository
Replace OpenAI GPT with another LLM in your app
Cross-platform AI language practice app
Han Language Processing
Underthesea - Vietnamese NLP Toolkit
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
A full spaCy pipeline and models for scientific/biomedical documents
Build your own AI friend
Enhances Tesseract OCR output using LLMs (local or API)
The first AI that can earn its own existence, replicate, and evolve
Toolkit for conversational AI
Open source AI VTuber platform with voice chat and Live2D avatars
The no-nonsense RAG chunking library
A PyTorch-based Speech Toolkit
Fast multimodal LLM for real-time voice interaction and AI apps
Research and application of technologies such as nl processing
A pure Javascript Multilingual OCR
Recognition and resolution of numbers, units, date/time, etc.
Training data (data labeling, annotation, workflow) for all data types
The media player for language learning, with dual subtitles
Documents and exposes generated source for menu-bar workflows
Contexts Optical Compression