A unified framework for scalable computing
A library for accelerating Transformer models on NVIDIA GPUs
Powering Amazon custom machine learning chips
On-device AI across mobile, embedded and edge for PyTorch
A GPU-accelerated library containing highly optimized building blocks
Replace OpenAI GPT with another LLM in your app
PArallel Distributed Deep LEarning: Machine Learning Framework
OpenMLDB is an open-source machine learning database
INT4/INT5/INT8 and FP16 inference on CPU for RWKV language model
Trainable, memory-efficient, and GPU-friendly PyTorch reproduction
A high-performance ML model serving framework, offers dynamic batching
Unified Model Serving Framework
Library for serving Transformers models on Amazon SageMaker
Lightweight Python library for adding real-time multi-object tracking
OpenAI swift async text to image for SwiftUI app using OpenAI
C++ implementation of ChatGLM-6B & ChatGLM2-6B & ChatGLM3 & GLM4(V)
Bolt is a deep learning library with high performance
A real time inference engine for temporal logical specifications
GPU environment management and cluster orchestration
A graphical manager for ollama that can manage your LLMs
Easy-to-use deep learning framework with 3 key features
PyTorch library of curated Transformer models and their components
An innovative library for efficient LLM inference
The unofficial python package that returns response of Google Bard
Open platform for training, serving, and evaluating language models