HunyuanVideo: A Systematic Framework For Large Video Generation Model
Faster Whisper transcription with CTranslate2
How to optimize some algorithm in cuda
LightLLM is a Python-based LLM (Large Language Model) inference
A high-quality rapid TTS voice cloning model
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
Pruna is a model optimization framework built for developers
Generate audiobooks from e-books
Sharp Monocular Metric Depth in Less Than a Second
Easily compute clip embeddings and build a clip retrieval system
Self-supervised visual learning using momentum contrast in PyTorch
Unified web UI for training and running open models locally
A high-performance ML model serving framework, offers dynamic batching
Deep learning optimization library: makes distributed training easy
Library for OCR-related tasks powered by Deep Learning
Large-Scale Agentic RL for High-Performance CUDA Kernel Generation
The best ChatGPT that $100 can buy
Multi-lingual large voice generation model, providing inference
DeepMind model for tracking arbitrary points across videos & robotics
A PyTorch-based Speech Toolkit
ChatGLM2-6B: An Open Bilingual Chat LLM
YOLOv5 is the world's most loved vision AI
An opinionated CLI to transcribe Audio files w/ Whisper on-device
A GUI tool for generating subtitle from videos, generating srt files
Libraries for optimizing AI models, inference speed, and GPU usage