The agent that grows with you
The most powerful and modular diffusion model GUI, api and backend
Real time face swap and one-click video deepfake
Automatic Speech Recognition with Word-level Timestamps
Run Local LLMs on Any Device. Open-source
A Lightweight Face Recognition and Facial Attribute Analysis
A lightweight audio-to-MIDI converter with pitch bend detection
OCRmyPDF adds an OCR text layer to scanned PDF files
Machine learning in Python
Faster Whisper transcription with CTranslate2
Official inference repo for FLUX.1 models
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
Letta (formerly MemGPT) is a framework for creating LLM services
Ready-to-use OCR with 80+ supported languages
Framework for Telegram Bot API written in Python 3.7 with asyncio
AI that gets your everyday tasks done
Effortless data labeling with AI support from Segment Anything
The most powerful local music generation model
Agent Zero AI framework
A personal AI assistant, easy to install
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Instant voice cloning by MIT and MyShell. Audio foundation model
Open source OSINT tool for gathering data on emails, phones, and IPs
Command line OSINT and threat intelligence automation tool
Make websites accessible for AI agents