An open-source framework for training large multimodal models
Task-oriented finetuning for better embeddings on neural search
Implementation / replication of DALL-E, OpenAI's Text to Image
Official PyTorch Implementation of "Scalable Diffusion Models"
Text-to-3D & Image-to-3D & Mesh Exportation with NeRF + Diffusion
Img2Txt - Extract Text From Images using AI
Deep learning tool that converts portrait photos into line art
A Python library for turning text quotes into graphical images
A walk along memory lane
Point cloud diffusion for 3D model synthesis
Text-conditional image generation model based on OpenAI's unCLIP
A latent text-to-image diffusion model
Real-time music generation using stable diffusion techniques AI
A minimal implementation of diffusion models for text generation
CPT: A Pre-Trained Unbalanced Transformer
Singing Voice Synthesis via Shallow Diffusion Mechanism
An interpretable and efficient predictor using pre-trained models
Demo for the "Talking Head Anime from a Single Image"
Based on the Disco Diffusion, version of the AI art creation software
min(DALL·E) is a fast, minimal port of DALL·E Mini to PyTorch
Generate images from texts. In Russian
Notebooks, models and techniques for the generation of AI Art
GLIDE: a diffusion-based text-conditional image synthesis model
An Open-Source Framework for Prompt-Learning
Simple command line tool for text to image generation