Solve puzzles. Learn CUDA
Running large language models on a single GPU
Standardized Serverless ML Inference Platform on Kubernetes
HunyuanVideo: A Systematic Framework For Large Video Generation Model
Open Source Differentiable Computer Vision Library
ReFT: Representation Finetuning for Language Models
ProtoMotions is a GPU-accelerated simulation and learning framework
Making large AI models cheaper, faster and more accessible
Easily compute clip embeddings and build a clip retrieval system
Toolkit for conversational AI
Voice Recognition to Text Tool
Faster Whisper transcription with CTranslate2
Large Language Model Text Generation Inference
Data manipulation and transformation for audio signal processing
Minimal Python framework for scalable AI inference servers fast
Suite of reference architectures for building GPU-accelerated vision
Instill Core is a full-stack AI infrastructure tool for data
Sharp Monocular Metric Depth in Less Than a Second
A sound cloning tool with a web interface, using your voice
Unified Model Serving Framework
Generative AI reference workflows
Open source AI VTuber platform with voice chat and Live2D avatars
The official repo of Qwen chat & pretrained large language model
Deep learning library
Controllable and fast Text-to-Speech for over 7000 languages