1B text generation model based on the HRM architecture
Module for automatic summarization of text documents and HTML pages
High-performance inference server for text embeddings models API layer
Large Language Model Text Generation Inference
Hypernetworks that adapt LLMs for specific benchmark tasks
First class Sublime Text AI assistant with gpt-5, Opus 4.6, Gemini 3
AI tool that removes hardcoded subtitles and text from videos locally
TTS with kokoro and onnx runtime
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
Awesome multilingual OCR toolkits based on PaddlePaddle
Cut videos with a text editor
Open-Source Python3 tool for recognizing layouts, tables, and math
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Speech-AI-Forge is a project developed around TTS generation model
Mozc - a Japanese Input Method Editor designed for multi-platform
High-Quality Voice Cloning TTS for 600+ Languages
Industrial-level controllable zero-shot text-to-speech system
Offline inference engine for art, real-time voice conversations
FastAPI framework, high performance, easy to learn, fast to code
Wan2.2: Open and Advanced Large-Scale Video Generative Model
A simple, high-quality voice conversion tool focused on ease of use
A simple native web interface that uses ChatTTS to synthesize text
Official inference repo for FLUX.1 models
Tokenizer-Free TTS for Multilingual Speech Generation
Converts text to speech in realtime