Context-aware desktop AI assistant that understands screen content
Collection of Gemma 3 variants that are trained for performance
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
High accuracy RAG for answering questions from scientific documents
Open source OSINT tool for gathering data on emails, phones, and IPs
Free, high-quality text-to-speech API endpoint to replace OpenAI
A minimalist command line knowledge base manager
Speech recognition module for Python
Python library and CLI tool to interface with Google Translate
Python binding to the Apache Tika™ REST services
A browser agent with a dynamic, indexed action space
AsrTools: Smart Voice-to-Text Tool
A high-quality PDF to Markdown tool based on large language model
Agent harness to make your slop code well-engineered and beautiful
Multi-tool for semantic search
Image inpainting tool powered by SOTA AI Model
Offline inference engine for art, real-time voice conversations
A nearly-live implementation of OpenAI's Whisper
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Generate blog articles from video or audio
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Speech-AI-Forge is a project developed around TTS generation model
Contexts Optical Compression
A robust, efficient, low-latency speech-to-text library
MOSS-TTS-Nano is an open-source multilingual tiny speech generation