ONNX Runtime: cross-platform, high performance ML inferencing
Open standard for machine learning interoperability
MNN is a blazing fast, lightweight deep learning framework
OpenVINO™ Toolkit repository
A GPU-accelerated library containing highly optimized building blocks
A unified framework for scalable computing
C++ library for high performance inference on NVIDIA GPUs
High-performance neural network inference framework for mobile
The Triton Inference Server provides an optimized cloud
Powering Amazon custom machine learning chips
FlashInfer: Kernel Library for LLM Serving
Neural Network Compression Framework for enhanced OpenVINO
Standardized Serverless ML Inference Platform on Kubernetes
Run serverless GPU workloads with fast cold starts on bare-metal
PArallel Distributed Deep LEarning: Machine Learning Framework
Everything you need to build state-of-the-art foundation models
DoWhy is a Python library for causal inference
Unified Model Serving Framework
Library for OCR-related tasks powered by Deep Learning
Training and deploying machine learning models on Amazon SageMaker
Build Production-ready Agentic Workflow with Natural Language
Probabilistic reasoning and statistical analysis in TensorFlow
An Open-Source Programming Framework for Agentic AI
A set of Docker images for training and serving models in TensorFlow
Deep learning optimization library: makes distributed training easy