AlphaFold 3 inference pipeline
A scalable inference server for models optimized with OpenVINO
Fast LLM speculative inference server for consumer hardware
ONNX-TensorRT: TensorRT backend for ONNX
C++ Implementation of PyTorch Tutorials for Everyone
PyTorch/TorchScript/FX compiler for NVIDIA GPUs using TensorRT
Environments and algorithms for research in general reinforcement
A high-performance distributed file system
Serving system for machine learning models
Run GGUF models easily with a UI or API. One File. Zero Install.
Open source large-language-model based code completion engine
C++ library based on tensorrt integration