Buzz transcribes and translates audio offline
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference
MiniMax H3 is a general-purpose, omni-modal generative system
The most powerful local music generation model
Port of Facebook's LLaMA model in C/C++
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Open-source, high-performance AI model with advanced reasoning
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Powerful AI language model (MoE) optimized for efficiency/performance
Awesome multilingual OCR toolkits based on PaddlePaddle
Official inference repo for FLUX.1 models
Qwen's most powerful open-source image generation model
GLM-5: From Vibe Coding to Agentic Engineering
High-Resolution 3D Assets Generation with Large Scale Diffusion Models
From Vibe Coding to Agentic Engineering
From Images to High-Fidelity 3D Assets
Python bindings for llama.cpp
Official Python inference and LoRA trainer package
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models
Native and Compact Structured Latents for 3D Generation
Official inference repo for FLUX.2 models
Qwen3-TTS is an open-source series of TTS models
Qwen2.5-VL is the multimodal large language model series
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
An easy 1-click way to create beautiful artwork on your PC using AI