Foundational Models for State-of-the-Art Speech and Text Translation
Qwen3-Coder is the code version of Qwen3
A very simple framework for state-of-the-art NLP
Speech to Text to Speech, sends text as OSC messages
Models for the spaCy Natural Language Processing (NLP) library
Self-hosted AI audio transcription
AI assistant based on large models that can actively think and plan
Framework for building AI-powered interactive digital humans and agent
Assist in organizing your piles of documents
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Pycorrector is a toolkit for text error correction
A simple tool for reading in poorly redacted documents
GLM-4-Voice | End-to-End Chinese-English Conversational Model
Qwen3-omni is a natively end-to-end, omni-modal LLM
NLP Cloud serves high performance pre-trained or custom models for NER
Statistical machine intelligence and learning engine
OCR offline image text recognition command line windows program
End-to-end speech processing toolkit
Multi-modal large language model designed for audio understanding
Jittor is a high-performance deep learning framework
Bailing is a voice dialogue robot similar to GPT-4o
In-App assistant SDK to build a multimodal conversational UX websites
Language modeling in a sentence representation space
Parser generator to read, process, or translate structured text
Free OCR Software: No internet required, easy to use.