kaldi-asr/kaldi is the official location of the Kaldi project
Train machine learning models within Docker containers
Why use many token when few token do trick
SGLang is a fast serving framework for large language models
Python tool for browser-based interactive data apps in one file
A python tool that uses GPT-4, FFmpeg, and OpenCV
Style-Bert-VITS2: Bert-VITS2 with more controllable voice styles
A SOTA open-source image editing model
Python SDK for the Computer Use model Lux, developed by OpenAGI
Open-source framework for intelligent speech interaction
Large Multimodal Models for Video Understanding and Editing
An Open Source text-to-speech system built by inverting Whisper
Speech-AI-Forge is a project developed around TTS generation model
OCR expert VLM powered by Hunyuan's native multimodal architecture
RGBD video generation model conditioned on camera input
Benchmarking synthetic data generation methods
AIMET is a library that provides advanced quantization and compression
AI video customer acquisition and AI short drama creation platform
Generates original ARC-AGI-1-style tasks distribution-matched
AI cognitive-enhancement Skills based on Anthropic's J-space
Codex plugin that turns attached object images into code-only
Agent skill: make LLMs write docs in ASD-STE100
An open-source toolkit for BigMac-style pipeline-parallel training
An Open Real-time Video-Language Interaction System
Open Vision Agents by Stream. Build voice and vision agents quickly