Large Multimodal Models for Video Understanding and Editing
Automating making many trailer-like videos with a single click!
Open-source AI video pipeline, fully automated with MCP
Official Repository for Pot-O MusiQT
Free Windows clipboard manager with Clipboard Peek and Smart Paste
Uility to make home movies from your digital camera files
DJ Sound Mixer
MARS5 speech model (TTS) from CAMB.AI
Embed images and sentences into fixed-length vectors
Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis
CLIP + FFT/DWT/RGB = text to image/video
Generate Harmonious Colors Freely.
An open-source framework for training large multimodal models
Implementation / replication of DALL-E, OpenAI's Text to Image
A latent text-to-image diffusion model
Text-conditional image generation model based on OpenAI's unCLIP
Based on the Disco Diffusion, version of the AI art creation software
Simple command line tool for text to image generation
A simple command line tool for text to image generation
Local image generation using VQGAN-CLIP or CLIP guided diffusion
A CLI tool/python module for generating images from text
LiVES is a Video Editing System. It is designed to be simple to use, y
Image augmentation for machine learning experiments