lightweight package to simplify LLM API calls
Faster Whisper transcription with CTranslate2
Convert Google Gemini web into OpenAI-compatible API
Document Image Parsing via Heterogeneous Anchor Prompting”
Long-form streaming TTS system for multi-speaker dialogue generation
StreamSpeech is a seamless model for offline speech recognition
Oobabooga - The definitive Web UI for local AI, with powerful features
Easy-to-use Speech Toolkit including Self-Supervised Learning model
A lightweight text-to-speech model with zero-shot voice cloning
The official Python SDK for the ElevenLabs API
Towards Human-Sounding Speech
TTS model capable of streaming conversational audio in realtime
Execute SQL queries and manage databases seamlessly with Timeplus
Reverse-engineered Python API for Google Gemini web app
A middleware to provide an openAI compatible endpoint
MOSS‑TTS Family open‑source speech and sound generation model
A GPT-4o Level MLLM for Vision, Speech and Multimodal Live Streaming
Provides convenient access to the Anthropic REST API from any Python 3
Free, high-quality text-to-speech API endpoint to replace OpenAI
NVR with realtime local object detection for IP cameras
Nexa SDK is a comprehensive toolkit for supporting ONNX and GGML
Anthropic's educational courses
Claude Code, but it runs on your Mac for free
Python library for building agents that leverages Google Antigravity
A floating, screen-aware Claude Code chat for Windows