Open image model at the forefront of design
FastAPI framework, high performance, easy to learn, fast to code
Self-host the powerful Chatterbox TTS model
Apache-2.0 open-source image generation and editing model family
Toolkit for conversational AI
A Family of Open Sourced Music Foundation Models
Music Assistant is a free, opensource Media library manager
Text and image to video generation: CogVideoX and CogVideo
Edit PDF files with Nano Banana
Official inference repo for FLUX.2 models
A lightweight text-to-speech model with zero-shot voice cloning
Nexa SDK is a comprehensive toolkit for supporting ONNX and GGML
Use Microsoft Edge's online text-to-speech service from Python
A Powerful Native Multimodal Model for Image Generation
Unifying 3D Mesh Generation with Language Models
An Open-Source Toolkit for General-OCR Research and Applications
Converts text to speech in realtime
AI-powered tool for generating, optimizing, and translating subtitles
A simple tool for reading in poorly redacted documents
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
Generate audiobooks from e-books, voice cloning & 1107+ languages
ASCII art library for Python
A simple, high-quality voice conversion tool focused on ease of use
This does for Documents what repo-browser does for repos
Framework for building real-time voice and multimodal AI agents