Tool to remotely activate Text-To-Speech (TTS) on a server
Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis
The PyTorch-based audio source separation toolkit for researchers
OpenMMLab's Next Generation Video Understanding Toolbox and Benchmark
Run the Stable Diffusion releases in a Docker container
One-click face swap
OpenMMLab Image Classification Toolbox and Benchmark
Video automatic transcribe and translated subtitle generator
Android Manga reader with Japanese OCR and dictionary capabilities
Video Frame Interpolation & Super Resolution using NVIDIA's TensorRT
Desktop software for controlling the Vosk Speech Recognition Toolkit
Audio generation using diffusion models, in PyTorch
Batch file to install and run NAM (neural-amp-modeler) easily.
Capture and control API for IIDC compliant cameras
AI powered image classification for nudity and documents / id-cards
A library for audio and music analysis, feature extraction.
Lightweight Stable Diffusion v 2.1 web UI: txt2img, img2img, depth2img
A repository of trained models
A Strong and Easy-to-use Single View 3D Hand+Body Pose Estimator
A simple PyTorch Implementation of Generative Adversarial Networks
The deep learning toolkit for speech-to-text