Run Local LLMs on Any Device. Open-source
Library for serving Transformers models on Amazon SageMaker
A high-performance ML model serving framework, offers dynamic batching
Lightweight Python library for adding real-time multi-object tracking
Tensor search for humans
A set of Docker images for training and serving models in TensorFlow
A graphical manager for ollama that can manage your LLMs
GUI shell for running local LLM on desktop
Openai style api for open large language models
An easy-to-use LLMs quantization package with user-friendly apis
A computer vision framework to create and deploy apps in minutes
Database system for building simpler and faster AI-powered application
Run any Llama 2 locally with gradio UI on GPU or CPU from anywhere
CPU/GPU inference server for Hugging Face transformer models