GUI for a Vocal Remover that uses Deep Neural Networks
Faster Whisper transcription with CTranslate2
OCR software, free and offline
Stable Diffusion web UI
Welcome the Era of One-shot Long-horizon Parsing
Visual Causal Flow
SkyPilot: Run AI and batch jobs on any infra
Use Microsoft Edge's online text-to-speech service from Python
Generate audiobooks from EPUBs, PDFs and text with captions
The easiest way to use deep metric learning in your application
AI video generator optimized for low VRAM and older GPUs use
Comprehensive Gradio WebUI for audio processing
Open source healthcare AI
AI-data warehouse to enrich, transform and analyze unstructured data
1 min voice data can also be used to train a good TTS model
Official repository for LTX-Video
Lets make video diffusion practical
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
Fast backend for long-term AI user memory via structured profiles
Open source platform for the machine learning lifecycle
95% token savings. 155x faster queries. 16 languages
LLM-based agent for general purpose software engineering tasks
Prevent PyTorch's `CUDA error: out of memory` in just 1 line of code
Build your own Cowork, AI Scientist and other SoTA Agents
A TTS that fits in your CPU (and pocket)