Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Easy-to-use and high-performance NLP and LLM framework
Open source machine learning framework to automate text conversations
Toolkit for conversational AI
Official Python inference and LoRA trainer package
An Open Source implementation of Notebook LM with more flexibility
An open source implementation of CLIP
TTS model capable of streaming conversational audio in realtime
Create videos with Stable Diffusion
The beginning of scalable pixel-native search
Build Vision Agents quickly with any model or video provider
Search all of YouTube from the command line
Multilingual sentence & image embeddings with BERT
Data Infrastructure providing an approach to multimodal AI workloads
[NeurIPS 2023] ImageReward: Learning and Evaluating Human Preferences
LLM inference server with continuous batching & SSD caching
Speech-AI-Forge is a project developed around TTS generation model
lightweight package to simplify LLM API calls
Unified web UI for training and running open models locally
Long-form streaming TTS system for multi-speaker dialogue generation
The most powerful local music generation model
A TTS model capable of generating ultra-realistic dialogue
NeuTTS model built from small LLM backbones
On-device TTS model by Neuphonic
Foundational video generation model with 13.6B parameters