Interactive video and image annotation tool for computer vision
Reverse engineering Gemini's SynthID detection
Open Source Differentiable Computer Vision Library
Stable Diffusion web UI
A Powerful Native Multimodal Model for Image Generation
Fast image augmentation library and an easy-to-use wrapper
A Unified Framework for Text-to-3D and Image-to-3D Generation
Guiding Instruction-based Image Editing via Multimodal Large Language
CLIP, Predict the most relevant text snippet given an image
Easily turn large sets of image urls to an image dataset
Automates PWA asset generation and image declaration
CogView4, CogView3-Plus and CogView3(ECCV 2024)
[NeurIPS 2023] ImageReward: Learning and Evaluating Human Preferences
Rebuild the object in a reference image as a code-only, procedural
AI video generator optimized for low VRAM and older GPUs use
Open image model at the forefront of design
Unsupervised Learning for Image Registration
One-ink editorial print image skill
A Customizable Image-to-Video Model based on HunyuanVideo
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
80+ free AI services for chat, image, video, voice & APIs
Open platform for sharing and discovering Stable Diffusion models
Client-side indecent content checking powered by TensorFlow.js
Open-source image generative foundation model
A studio for image and video generation