Training and deploying machine learning models on Amazon SageMaker
Powering Amazon custom machine learning chips
Run Local LLMs on Any Device. Open-source
Single-cell analysis in Python
FlashInfer: Kernel Library for LLM Serving
A high-throughput and memory-efficient inference and serving engine
The official Python client for the Huggingface Hub
Everything you need to build state-of-the-art foundation models
DoWhy is a Python library for causal inference
Gaussian processes in TensorFlow
Optimizing inference proxy for LLMs
A unified framework for scalable computing
A Pythonic framework to simplify AI service building
Python Package for ML-Based Heterogeneous Treatment Effects Estimation
Adversarial Robustness Toolbox (ART) - Python Library for ML security
Probabilistic reasoning and statistical analysis in TensorFlow
Large Language Model Text Generation Inference
Uplift modeling and causal inference with machine learning algorithms
Replace OpenAI GPT with another LLM in your app
Data manipulation and transformation for audio signal processing
Pytorch domain library for recommendation systems
Operating LLMs in production
Multi-LoRA inference server that scales to 1000s of fine-tuned LLMs
Integrate, train and manage any AI models and APIs with your database
Libraries for applying sparsification recipes to neural networks