ADAMS is a workflow engine for building complex knowledge workflows.
Herramienta para Sintetizar voces en la PC
Embed images and sentences into fixed-length vectors
ITTT is a Free tool designed to Scan and extract Text from Images.
MMEditing is a low-level vision toolbox based on PyTorch
Tool to remotely activate Text-To-Speech (TTS) on a server
The PyTorch-based audio source separation toolkit for researchers
Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis
OpenMMLab's Next Generation Video Understanding Toolbox and Benchmark
Run the Stable Diffusion releases in a Docker container
One-click face swap
OpenMMLab Image Classification Toolbox and Benchmark
Android Manga reader with Japanese OCR and dictionary capabilities
Video automatic transcribe and translated subtitle generator
Video Frame Interpolation & Super Resolution using NVIDIA's TensorRT
Desktop software for controlling the Vosk Speech Recognition Toolkit
Capture and control API for IIDC compliant cameras
Batch file to install and run NAM (neural-amp-modeler) easily.
Audio generation using diffusion models, in PyTorch
AI powered image classification for nudity and documents / id-cards
A library for audio and music analysis, feature extraction.