Sharp Monocular Metric Depth in Less Than a Second
An Open-source Framework for Data-centric Language Agents
Instant voice cloning by MIT and MyShell. Audio foundation model
Deep learning optimization library: makes distributed training easy
Apache-2.0 open-source image generation and editing model family
Realtime Web Apps and Dashboards for Python and R
VoiceStudio is the open-source, fully-local ElevenLabs alternative
A high-quality rapid TTS voice cloning model
Persistent HTTP cache for python requests
MobileLLM Optimizing Sub-billion Parameter Language Models
TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning
Oobabooga - The definitive Web UI for local AI, with powerful features
scikit-learn compatible tabular foundation model
Official PyTorch Implementation
MOSS‑TTS Family open‑source speech and sound generation model
Spark-TTS Inference Code
Implementation of "MobileCLIP" CVPR 2024
CLIP, Predict the most relevant text snippet given an image
Privacy browser for Android
Multi-lingual large voice generation model, providing inference
Give your AI agent eyes to see the entire internet
Comprehensive Gradio WebUI for audio processing
A list of free LLM inference resources accessible via API
Low-latency AI inference engine optimized for mobile devices
The absolute trainer to light up AI agents