VoiceStudio is the open-source, fully-local ElevenLabs alternative
Supercharge Your LLM with the Fastest KV Cache Layer
3D reconstruction software
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Trainable latent-memory framework for 100M-token contexts
Gemma open-weight LLM library, from Google DeepMind
A Customizable Image-to-Video Model based on HunyuanVideo
TorchMultimodal is a PyTorch library
ChatGLM3 series: Open Bilingual Chat LLMs | Open Source Bilingual Chat
ComfyUI integration for Microsoft's VibeVoice text-to-speech model
Fast Python collaborative filtering for implicit feedback datasets
Style-Bert-VITS2: Bert-VITS2 with more controllable voice styles
SGLang is a fast serving framework for large language models
Running large language models on a single GPU
A unified framework for scalable computing
Multilingual Automatic Speech Recognition with word-level timestamps
Our first fully AI generated deep learning system
Interface for OuteTTS models
A free and reliable P2P BitTorrent client
AI Suite for upscaling, interpolating & restoring images/videos
Python package built to ease deep learning on graph
Blazingly Fast & Customizable Linux distribution
A local web server for developing web applications on Windows.
ChatGLM2-6B: An Open Bilingual Chat LLM
Offline desktop app to convert EPUB to MP3 using Kokoro-82M neural TTS