Embed images and sentences into fixed-length vectors
Fast ODE Solver for Diffusion Probabilistic Model Sampling
Generate 3D objects conditioned on text or images
Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis
CLIP + FFT/DWT/RGB = text to image/video
Official Code for DragGAN (SIGGRAPH 2023)
Text-to-Image generation. The repo for NeurIPS 2021 paper
Serve machine learning models within a Docker container
YoloV3 Implemented in Tensorflow 2.0
Alfred workflow using ChatGPT, DALL·E 2 and other models for chatting
Let us control diffusion models
Code release for "Detecting Twenty-thousand Classes
One-click face swap
Multimodal AI Story Teller, built with Stable Diffusion, GPT, etc.
An open-source framework for training large multimodal models
OpenMMLab Image Classification Toolbox and Benchmark
Official repo for consistency models
Visual localization made easy with hloc
Task-oriented finetuning for better embeddings on neural search
Enable sending and receiving images during chatting
Flash enables you to easily configure and run complex AI recipes
Official PyTorch Implementation of "Scalable Diffusion Models"
A large open dataset + tools to speed up MRI scans using ML
Implementation / replication of DALL-E, OpenAI's Text to Image
Text-to-3D & Image-to-3D & Mesh Exportation with NeRF + Diffusion