An Open Source implementation of Notebook LM with more flexibility
High-Quality Voice Cloning TTS for 600+ Languages
Powerful AI language model (MoE) optimized for efficiency/performance
A community-supported supercharged version of paperless
Open-source multi-speaker long-form text-to-speech model
From Images to High-Fidelity 3D Assets
Aider is AI pair programming in your terminal
Open-source AI agent framework
Personal AI, On Personal Devices
Use Microsoft Edge's online text-to-speech service from Python
A Lightweight Face Recognition and Facial Attribute Analysis
Native and Compact Structured Latents for 3D Generation
Deepfakes Software For All
Comprehensive Gradio WebUI for audio processing
Nexa SDK is a comprehensive toolkit for supporting ONNX and GGML
A high-quality tool for convert PDF to Markdown and JSON
The official Meta Llama 3 GitHub site
Code for running inference and finetuning with SAM 3 model
A natural language interface for computers
AI Fully Automated Short Video Engine
Get your documents ready for gen AI
Fast stable diffusion on CPU and AI PC
Industrial-level controllable zero-shot text-to-speech system
AI-powered video clipping and highlight generation
Official inference repo for FLUX.2 models