An Open Source implementation of Notebook LM with more flexibility
High-Quality Voice Cloning TTS for 600+ Languages
Powerful AI language model (MoE) optimized for efficiency/performance
Open-source multi-speaker long-form text-to-speech model
Aider is AI pair programming in your terminal
A community-supported supercharged version of paperless
From Images to High-Fidelity 3D Assets
Open-source AI agent framework
Personal AI, On Personal Devices
Use Microsoft Edge's online text-to-speech service from Python
A Lightweight Face Recognition and Facial Attribute Analysis
Native and Compact Structured Latents for 3D Generation
Deepfakes Software For All
Comprehensive Gradio WebUI for audio processing
Nexa SDK is a comprehensive toolkit for supporting ONNX and GGML
A high-quality tool for convert PDF to Markdown and JSON
Code for running inference and finetuning with SAM 3 model
The official Meta Llama 3 GitHub site
A natural language interface for computers
Industrial-level controllable zero-shot text-to-speech system
AI Fully Automated Short Video Engine
Get your documents ready for gen AI
Fast stable diffusion on CPU and AI PC
AI-powered video clipping and highlight generation
Official inference repo for FLUX.2 models