Buzz transcribes and translates audio offline
Port of Facebook's LLaMA model in C/C++
MiniMax H3 is a general-purpose, omni-modal generative system
Awesome multilingual OCR toolkits based on PaddlePaddle
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference
Open-source, high-performance AI model with advanced reasoning
The most powerful local music generation model
From Vibe Coding to Agentic Engineering
Official inference repo for FLUX.1 models
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Official Python inference and LoRA trainer package
An easy 1-click way to create beautiful artwork on your PC using AI
Open Frontier Intelligence
Open-source multi-speaker long-form text-to-speech model
Lets make video diffusion practical
From Images to High-Fidelity 3D Assets
Code for running inference and finetuning with SAM 3 model
Powerful AI language model (MoE) optimized for efficiency/performance
GLM-5: From Vibe Coding to Agentic Engineering
Industrial-level controllable zero-shot text-to-speech system
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
Qwen3.5 is the large language model series developed by Qwen team
Official inference repo for FLUX.2 models