Accurate × Fast × Comprehensive
A lightweight text-to-speech model with zero-shot voice cloning
The official Python SDK for UCP
Open-source large language model family from Tencent Hunyuan
A feature rich discord Modmail bot
Asynchronous multi-platform robot framework written in Python
2D and 3D Face alignment library build using pytorch
Industrial-strength Natural Language Processing (NLP)
VoiceStudio is the open-source, fully-local ElevenLabs alternative
Reference agents, skills, and data for the financial-services
An Open Source implementation of Notebook LM with more flexibility
AudioMuse-AI is an Open Source Dockerized environment
GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning
A Powerful Native Multimodal Model for Image Generation
A game theoretic approach to explain the output of ml models
Let Claude (or any LLM) actually watch a video
AI PPT Track Terminator, the strongest PPT Skill ever
RF-DETR is a real-time object detection and segmentation
Audio foundation model excelling in audio understanding
Generate audiobooks from EPUBs, PDFs and text with captions
Models for the spaCy Natural Language Processing (NLP) library
Investment Research for Everyone, Everywhere
Adding guardrails to large language models
Offline Text To Speech synthesis for python
AI video generator optimized for low VRAM and older GPUs use