State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX
Automatically translates the text of a video based on a subtitle file
Reading book source
Audio Language Models are Few-Shot Learners
Conversational voice AI agents
Open-source Video Translation Skill
A Web UI for easy subtitle using whisper model
Scalable generative AI framework built for researchers and developers
LLM Large Model of Selling Anchor
A Conversational Speech Generation Model
Googles NotebookLM but local
Framework for building AI-powered interactive digital humans and agent
Multi-modal large language model designed for audio understanding
Chat with it via text and voice
Flowly is 100x faster than OpenClaw
Build Vision Agents quickly with any model or video provider
A fast TTS architecture with conditional flow matching
Easy-to-use Speech Toolkit including Self-Supervised Learning model
A very simple framework for state-of-the-art NLP
Generate blog articles from video or audio
Open source personal AI Assistant for Linux, Windows and Mac
Pre-trained Deep Learning models and demos
AI generative media user experience highlighting use of APIs
Official Python inference and LoRA trainer package
Mice speech to text with MX Cinnamon OS ISO