Wan2.2: Open and Advanced Large-Scale Video Generative Model
Awesome multilingual OCR toolkits based on PaddlePaddle
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Official inference repo for FLUX.1 models
Official Python inference and LoRA trainer package
Open-source, high-performance AI model with advanced reasoning
The most powerful local music generation model
Powerful AI language model (MoE) optimized for efficiency/performance
Official inference repo for FLUX.2 models
Code for running inference and finetuning with SAM 3 model
Lets make video diffusion practical
Native and Compact Structured Latents for 3D Generation
Official repository for LTX-Video
Agentic, Reasoning, and Coding (ARC) foundation models
Convert Google Gemini web into OpenAI-compatible API
Python inference and LoRA trainer package for the LTX-2 audio–video
Text and image to video generation: CogVideoX and CogVideo
Qwen3-TTS is an open-source series of TTS models
Fast stable diffusion on CPU and AI PC
From Images to High-Fidelity 3D Assets
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
State-of-the-art TTS model under 25MB
LTX-Video Support for ComfyUI
Qwen3 is the large language model series developed by Qwen team