Fast stable diffusion on CPU and AI PC
Models for object and human mesh reconstruction
Wan2.2: Open and Advanced Large-Scale Video Generative Model
InvokeAI is a leading creative engine for Stable Diffusion models
Open Source Differentiable Computer Vision Library
Stable Diffusion web UI
Reverse engineering Gemini's SynthID detection
A Powerful Native Multimodal Model for Image Generation
A Unified Framework for Text-to-3D and Image-to-3D Generation
Fast image augmentation library and an easy-to-use wrapper
Guiding Instruction-based Image Editing via Multimodal Large Language
CLIP, Predict the most relevant text snippet given an image
Easily turn large sets of image urls to an image dataset
[NeurIPS 2023] ImageReward: Learning and Evaluating Human Preferences
CogView4, CogView3-Plus and CogView3(ECCV 2024)
Rebuild the object in a reference image as a code-only, procedural
Open image model at the forefront of design
AI video generator optimized for low VRAM and older GPUs use
Unsupervised Learning for Image Registration
One-ink editorial print image skill
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
A Customizable Image-to-Video Model based on HunyuanVideo
Open-source image generative foundation model
Flexible Photo Recrafting While Preserving Your Identity
An open source implementation of CLIP