Official Python inference and LoRA trainer package
The most powerful local music generation model
Official inference repo for FLUX.1 models
Fast stable diffusion on CPU and AI PC
Text and image to video generation: CogVideoX and CogVideo
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Official repository for LTX-Video
Open-source, high-performance AI model with advanced reasoning
AI PPT Track Terminator, the strongest PPT Skill ever
LTX-Video Support for ComfyUI
Awesome multilingual OCR toolkits based on PaddlePaddle
Revolutionizing Database Interactions with Private LLM Technology
A Systematic Framework for Interactive World Modeling
High-Resolution Image Synthesis with Latent Diffusion Models
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Official inference repo for FLUX.2 models
Python bindings for llama.cpp
Qwen-Image is a powerful image generation foundation model
Pokee Deep Research Model Open Source Repo
Python inference and LoRA trainer package for the LTX-2 audio–video
26m function call model that runs on incredibly small devices
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
A Powerful Native Multimodal Model for Image Generation
Robust Speech Recognition Across Languages, Dialects