Solve puzzles. Learn CUDA
AirLLM 70B inference with single 4GB GPU
Performance-optimized AI inference on your GPUs
Running large language models on a single GPU
The fundamental package for scientific computing with Python
Simple package for monitoring and control your NVIDIA Jetson
Parallax is a distributed model serving framework
Find the local LLM that actually runs and performs best
How to optimize some algorithm in cuda
NVIDIA Isaac Sim is an open-source application on NVIDIA Omniverse
Development repository for the Triton language and compiler
Run macOS on QEMU/KVM
SkyPilot: Run AI and batch jobs on any infra
Run Local LLMs on Any Device. Open-source
Ongoing research training transformer models at scale
Pythonic tool for running machine-learning/high performance workflows
ProtoMotions is a GPU-accelerated simulation and learning framework
AI agents running research on single-GPU nanochat training
A high-quality rapid TTS voice cloning model
Reinforcement Learning for Humanoid Robot with Zero-Shot Sim2Real
State-of-the-art Parameter-Efficient Fine-Tuning
Voice Recognition to Text Tool
VoiceStudio is the open-source, fully-local ElevenLabs alternative
Making large AI models cheaper, faster and more accessible
Open deep learning compiler stack for cpu, gpu, etc.