Training and deploying machine learning models on Amazon SageMaker
Run Local LLMs on Any Device. Open-source
Ready-to-use OCR with 80+ supported languages
FlashInfer: Kernel Library for LLM Serving
Python Package for ML-Based Heterogeneous Treatment Effects Estimation
DoWhy is a Python library for causal inference
A high-throughput and memory-efficient inference and serving engine
Everything you need to build state-of-the-art foundation models
Single-cell analysis in Python
Operating LLMs in production
Create HTML profiling reports from pandas DataFrame objects
Unified Model Serving Framework
Gaussian processes in TensorFlow
Uplift modeling and causal inference with machine learning algorithms
The official Python client for the Huggingface Hub
Official inference library for Mistral models
Optimizing inference proxy for LLMs
Adversarial Robustness Toolbox (ART) - Python Library for ML security
Powering Amazon custom machine learning chips
LMDeploy is a toolkit for compressing, deploying, and serving LLMs
Neural Network Compression Framework for enhanced OpenVINO
Trainable models and NN optimization tools
Multi-LoRA inference server that scales to 1000s of fine-tuned LLMs
Efficient few-shot learning with Sentence Transformers
A unified framework for scalable computing