A Python library for audio
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
Official Python inference and LoRA trainer package
Multimodal Diffusion with Representation Alignment
Official repository for LTX-Video
Speech recognition module for Python
Ableton Live Model Context Protocol Integration
MOSS‑TTS Family open‑source speech and sound generation model
Python inference and LoRA trainer package for the LTX-2 audio–video
State-of-the-art TTS model under 25MB
A Telegram bot that integrates with OpenAI's official ChatGPT APIs
Build Vision Agents quickly with any model or video provider
A GPT-4o Level MLLM for Vision, Speech and Multimodal Live Streaming
Audiocraft is a library for audio processing and generation
Ainee - AI Notetaking and Learning Companion
IPTV/NVR/CCTV/Video cloud https://fastocloud.com