ONNX Runtime: cross-platform, high performance ML inferencing
Unified Model Serving Framework
Powering Amazon custom machine learning chips
C++ library for high performance inference on NVIDIA GPUs
MNN is a blazing fast, lightweight deep learning framework
Library for OCR-related tasks powered by Deep Learning
A GPU-accelerated library containing highly optimized building blocks
A general-purpose probabilistic programming system
An MLOps framework to package, deploy, monitor and manage models
AIMET is a library that provides advanced quantization and compression
A unified framework for scalable computing
The Triton Inference Server provides an optimized cloud
Easy-to-use deep learning framework with 3 key features
OpenMMLab Model Deployment Framework
Self-contained Machine Learning and Natural Language Processing lib
High-level Deep Learning Framework written in Kotlin
Guide to deploying deep-learning inference networks
Deep learning inference framework optimized for mobile platforms
Uniform deep learning inference framework for mobile
Deploy a ML inference service on a budget in 10 lines of code