GLM-4-Voice | End-to-End Chinese-English Conversational Model
Generate high-definition story short videos with one click using AI
A Telegram RSS bot that cares about your reading experience
A Telegram bot that integrates with OpenAI's official ChatGPT APIs
Adversarial Robustness Toolbox (ART) - Python Library for ML security
Framework for research and development of foundation models
JamTools is a cross-platform gadget set software
Decentralize, Self-host Cloud Gaming/Application
Automated YouTube Shorts pipeline
Pure Python FFmpeg-based live video / audio streaming to YouTube
"VideoRAG: Chat with Your Videos
Multi-lingual large voice generation model, providing inference
The Triton Inference Server provides an optimized cloud
Hub of ready-to-use datasets for ML models
Claude Code blog skill suite
Open Vision Agents by Stream. Build voice and vision agents quickly
Open-source Video Translation Skill
Multi-source content processor for NotebookLM
Instill Core is a full-stack AI infrastructure tool for data
Open-source abilities for OpenHome agents
Build multimodal AI applications with cloud-native stack
Improve human sleep through scientifically
Omnilingual ASR Open-Source Multilingual SpeechRecognition
TorchMultimodal is a PyTorch library
A GPT-4o Level MLLM for Vision, Speech and Multimodal Live Streaming