A simple native web interface that uses ChatTTS to synthesize text
The behavior guidance framework for customer-facing LLM agents
Automatic Speech Recognition with Word-level Timestamps
Crowdsourcing platform for full text transcription and tagging
Offline inference engine for art, real-time voice conversations
An open-source toolkit for monitoring Language Learning Models (LLMs)
Code for running inference and finetuning with SAM 3 model
SOTA Open Source TTS
OCR software, free and offline
Video-based AI memory library. Store millions of text chunks in MP4
Robust Speech Recognition via Large-Scale Weak Supervision
Mozc - a Japanese Input Method Editor designed for multi-platform
A high-quality rapid TTS voice cloning model
A lightweight text-to-speech model with zero-shot voice cloning
Contexts Optical Compression
Speech recognition module for Python
A Python toolbox for gaining geometric insights
Qwen3-TTS is an open-source series of TTS models
High accuracy RAG for answering questions from scientific documents
A pure-python PDF library capable of splitting, merging, cropping
A minimalist command line knowledge base manager
Converts text to speech in realtime
Python bindings for MuPDF's rendering library.
Google Gen AI Python SDK provides an interface for developers
Tokenizer-Free TTS for Multilingual Speech Generation