Qwen-Image is a powerful image generation foundation model
The largest collection of PyTorch image encoders / backbones
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
Image inpainting tool powered by SOTA AI Model
Open Source Differentiable Computer Vision Library
Fast-stable-diffusion + DreamBooth
Foundational video generation model with 13.6B parameters
Easily compute clip embeddings and build a clip retrieval system
A Customizable Image-to-Video Model based on HunyuanVideo
Kaggle Python docker image
The most powerful and modular diffusion model GUI, api and backend
HunyuanVideo: A Systematic Framework For Large Video Generation Model
Official inference repo for FLUX.2 models
AI video generator optimized for low VRAM and older GPUs use
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Sharp Monocular Metric Depth in Less Than a Second
Wan2.2: Open and Advanced Large-Scale Video Generative Model
High-Resolution Image Synthesis with Latent Diffusion Models
2D and 3D Face alignment library build using pytorch
Nexa SDK is a comprehensive toolkit for supporting ONNX and GGML
InvokeAI is a leading creative engine for Stable Diffusion models
Text and image to video generation: CogVideoX and CogVideo
Reference PyTorch implementation and models for DINOv3
Simplest working implementation of Stylegan2
Effortless data labeling with AI support from Segment Anything