Towards Human-Sounding Speech
Open source annotation tool for machine learning practitioners
Crowdsourcing platform for full text transcription and tagging
Agent harness to make your slop code well-engineered and beautiful
Claude Code skill implementing Manus-style persistent planning
Offline inference engine for art, real-time voice conversations
Miso TTS is an 8 billion, highly emotive text-to-speech model
Management of Yandex Station and other smart home devices
A fast TTS architecture with conditional flow matching
Jupyter Notebooks as Markdown Documents, Julia, Python or R scripts
Snippet solution for Vim
Comprehensive Markdown plugin built for Django
A sound cloning tool with a web interface, using your voice
Designed for text embedding and ranking tasks
A lightweight text-to-speech model with zero-shot voice cloning
Style-Bert-VITS2: Bert-VITS2 with more controllable voice styles
Python bindings for MuPDF's rendering library.
A python parametric CAD scripting framework based on OCCT
Tools to ease the creation of snippets, syntax definitions, etc.
An open-source toolkit for monitoring Language Learning Models (LLMs)
Speech recognition module for Python
CLIP, Predict the most relevant text snippet given an image
Multimodal-Driven Architecture for Customized Video Generation
GLM-4-Voice | End-to-End Chinese-English Conversational Model
Unifying 3D Mesh Generation with Language Models