Implementation of Make-A-Video, new SOTA text to video generator
Virtual AI anchor that combines state-of-the-art technology
A library for transfer learning by reusing parts of TensorFlow models
Multi-Voice and Prompt-Controlled TTS Engine
Official code for Style Aligned Image Generation via Shared Attention
Embed images and sentences into fixed-length vectors
Convert an image to text to spot intelligible words.
Generate 3D objects conditioned on text or images
Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis
CLIP + FFT/DWT/RGB = text to image/video
Text-to-Image generation. The repo for NeurIPS 2021 paper
Run the Stable Diffusion releases in a Docker container
Alfred workflow using ChatGPT, DALL·E 2 and other models for chatting
Let us control diffusion models
Multimodal AI Story Teller, built with Stable Diffusion, GPT, etc.
An open-source framework for training large multimodal models
Task-oriented finetuning for better embeddings on neural search
Implementation / replication of DALL-E, OpenAI's Text to Image
Official PyTorch Implementation of "Scalable Diffusion Models"
Text-to-3D & Image-to-3D & Mesh Exportation with NeRF + Diffusion
Img2Txt - Extract Text From Images using AI
Deep learning tool that converts portrait photos into line art
A Python library for turning text quotes into graphical images
A walk along memory lane
Point cloud diffusion for 3D model synthesis