A toolkit to optimize ML models for deployment for Keras & TensorFlow
Deep learning optimization library: makes distributed training easy
PyTorch extensions for fast R&D prototyping and Kaggle farming
Official inference library for Mistral models
Open-source tool designed to enhance the efficiency of workloads
MII makes low-latency and high-throughput inference possible
20+ high-performance LLMs with recipes to pretrain, finetune at scale
Data manipulation and transformation for audio signal processing
A Unified Library for Parameter-Efficient Learning
Scripts for fine-tuning Meta Llama3 with composable FSDP & PEFT method
A lightweight vision library for performing large object detection
Simplifies the local serving of AI models from any source
Easy-to-use Speech Toolkit including Self-Supervised Learning model
Unified Model Serving Framework
AIMET is a library that provides advanced quantization and compression
A GPU-accelerated library containing highly optimized building blocks
Low-latency REST API for serving text-embeddings
Trainable, memory-efficient, and GPU-friendly PyTorch reproduction
LLM training code for MosaicML foundation models
An MLOps framework to package, deploy, monitor and manage models
Create HTML profiling reports from pandas DataFrame objects
Library for serving Transformers models on Amazon SageMaker
Multi-Modal Neural Networks for Semantic Search, based on Mid-Fusion
Tensor search for humans
Powering Amazon custom machine learning chips