ONNX Runtime: cross-platform, high performance ML inferencing
Convert TensorFlow, Keras, Tensorflow.js and Tflite models to ONNX
MLX: An array framework for Apple silicon
Rust native ready-to-use NLP pipelines and transformer-based models
OpenVINO™ Toolkit repository
OpenMMLab Model Deployment Framework
Embed images and sentences into fixed-length vectors
High-level Deep Learning Framework written in Kotlin
CPU/GPU inference server for Hugging Face transformer models
Deep learning inference framework optimized for mobile platforms