Awesome multilingual OCR toolkits based on PaddlePaddle
Official Python inference and LoRA trainer package
Official inference repo for FLUX.1 models
The most powerful local music generation model
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Powerful AI language model (MoE) optimized for efficiency/performance
Official repository for LTX-Video
Open-source, high-performance AI model with advanced reasoning
State-of-the-art TTS model under 25MB
Qwen3-TTS is an open-source series of TTS models
LTX-Video Support for ComfyUI
Qwen3-Coder is the code version of Qwen3
Python bindings for llama.cpp
Fast stable diffusion on CPU and AI PC
Visual Causal Flow
Agentic, Reasoning, and Coding (ARC) foundation models
Advanced language and coding AI model
Phi-3.5 for Mac: Locally-run Vision and Language Models
Native and Compact Structured Latents for 3D Generation
Lets make video diffusion practical
Official inference repo for FLUX.2 models
Code for running inference and finetuning with SAM 3 model
Qwen-Image is a powerful image generation foundation model