HunyuanVideo: A Systematic Framework For Large Video Generation Model
An interactive NVIDIA-GPU process viewer and beyond
Faster Whisper transcription with CTranslate2
A NumPy-compatible array library accelerated by CUDA
How to optimize some algorithm in cuda
The fundamental package for scientific computing with Python
A high-quality rapid TTS voice cloning model
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
LightLLM is a Python-based LLM (Large Language Model) inference
Generate audiobooks from e-books
Sharp Monocular Metric Depth in Less Than a Second
Sharp Monocular View Synthesis in Less Than a Second
Pruna is a model optimization framework built for developers
Easily compute clip embeddings and build a clip retrieval system
Enables the best performance on NVIDIA RTX Graphics Cards
Self-supervised visual learning using momentum contrast in PyTorch
Unified web UI for training and running open models locally
A high-performance ML model serving framework, offers dynamic batching
Deep learning optimization library: makes distributed training easy
Library for OCR-related tasks powered by Deep Learning
Large-Scale Agentic RL for High-Performance CUDA Kernel Generation
The best ChatGPT that $100 can buy
Multi-lingual large voice generation model, providing inference
DeepMind model for tracking arbitrary points across videos & robotics
A PyTorch-based Speech Toolkit