OpenVINO™ Toolkit repository
Set of comprehensive computer vision & machine intelligence libraries
Connect home devices into a powerful cluster to accelerate LLM
A unified framework for scalable computing
A library for accelerating Transformer models on NVIDIA GPUs
Powering Amazon custom machine learning chips
On-device AI across mobile, embedded and edge for PyTorch
A GPU-accelerated library containing highly optimized building blocks
Replace OpenAI GPT with another LLM in your app
PArallel Distributed Deep LEarning: Machine Learning Framework
OpenMLDB is an open-source machine learning database
INT4/INT5/INT8 and FP16 inference on CPU for RWKV language model
Trainable, memory-efficient, and GPU-friendly PyTorch reproduction
A high-performance ML model serving framework, offers dynamic batching
Unified Model Serving Framework
Easy-to-use Speech Toolkit including Self-Supervised Learning model
Library for serving Transformers models on Amazon SageMaker
Lightweight Python library for adding real-time multi-object tracking
OpenAI swift async text to image for SwiftUI app using OpenAI
The Triton Inference Server provides an optimized cloud
C++ implementation of ChatGLM-6B & ChatGLM2-6B & ChatGLM3 & GLM4(V)
Bolt is a deep learning library with high performance
A real time inference engine for temporal logical specifications
GPU environment management and cluster orchestration
A graphical manager for ollama that can manage your LLMs