Create videos with Stable Diffusion
The best free open source website change detection and restock service
TTS model capable of streaming conversational audio in realtime
An open source implementation of CLIP
The most reliable AI agent framework that supports MCP
Data Infrastructure providing an approach to multimodal AI workloads
Build multimodal language agents for fast prototype and production
An Open Source implementation of Notebook LM with more flexibility
lightweight package to simplify LLM API calls
Speech-AI-Forge is a project developed around TTS generation model
Main repository for the Sphinx documentation builder
Multi-lingual large voice generation model, providing inference
Zero-copy PDF text extraction library written in Zig
Turn words into chords
Search all of YouTube from the command line
Unified web UI for training and running open models locally
Long-form streaming TTS system for multi-speaker dialogue generation
The most powerful local music generation model
Multilingual sentence & image embeddings with BERT
NeuTTS model built from small LLM backbones
On-device TTS model by Neuphonic
Network analysis in Python
Text and supporting code for Think Stats, 2nd Edition
AI-friendly PPT builder skill: 17 hand-polished Chinese PPTX templates
Googles NotebookLM but local