VoiceStudio is the open-source, fully-local ElevenLabs alternative
Supercharge Your LLM with the Fastest KV Cache Layer
3D reconstruction software
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Trainable latent-memory framework for 100M-token contexts
Gemma open-weight LLM library, from Google DeepMind
A Customizable Image-to-Video Model based on HunyuanVideo
ChatGLM3 series: Open Bilingual Chat LLMs | Open Source Bilingual Chat
ComfyUI integration for Microsoft's VibeVoice text-to-speech model
Fast Python collaborative filtering for implicit feedback datasets
Style-Bert-VITS2: Bert-VITS2 with more controllable voice styles
SGLang is a fast serving framework for large language models
Running large language models on a single GPU
A unified framework for scalable computing
Multilingual Automatic Speech Recognition with word-level timestamps
Our first fully AI generated deep learning system
Interface for OuteTTS models
AI Suite for upscaling, interpolating & restoring images/videos
Python package built to ease deep learning on graph
Offline desktop app to convert EPUB to MP3 using Kokoro-82M neural TTS
ChatGLM2-6B: An Open Bilingual Chat LLM
Query-aware prompt compression for high-signal LLM prompts.
Open platform for training, serving, and evaluating language models
High quality, fast, modular reference implementation of SSD in PyTorch
A computer vision framework to create and deploy apps in minutes