Real-time NVIDIA GPU dashboard
Solve puzzles. Learn CUDA
Monitor temperature sensors, fan speed, voltage, load & clock speeds
Monitor temperature sensors, fan speeds, voltages, load & clock speeds
NVIDIA GPU Operator creates/configures/manages GPUs atop Kubernetes
Emulating Apple Silicon devices
Everything I know about running LLMs locally
A hardware-accelerated GPU terminal emulator
Multi-platform high-performance compute language extension for Rust
SwiftShader is a high-performance CPU-based implementation
157 models, 30 providers, one command to find what runs on hardware
AirLLM 70B inference with single 4GB GPU
Performance-optimized AI inference on your GPUs
A tool for converting Xbox 360 shaders to HLSL
Running large language models on a single GPU
C++ and Python support for the CUDA Quantum programming model
HeavyDB (formerly MapD/OmniSciDB)
The free, Open Source alternative to OpenAI, Claude and others
Simple package for monitoring and control your NVIDIA Jetson
Fast LLM speculative inference server for consumer hardware
Parallax is a distributed model serving framework
High-speed Large Language Model Serving for Local Deployment
Find the local LLM that actually runs and performs best
High-performance CPU, GPU, and memory profiler for Python
LLM inference in C/C++