Trainable, memory-efficient, and GPU-friendly PyTorch reproduction
Training and deploying machine learning models on Amazon SageMaker
Large Language Model Text Generation Inference
Uncover insights, surface problems, monitor, and fine tune your LLM
Easy-to-use Speech Toolkit including Self-Supervised Learning model
Replace OpenAI GPT with another LLM in your app
Uplift modeling and causal inference with machine learning algorithms
DoWhy is a Python library for causal inference
A library for accelerating Transformer models on NVIDIA GPUs
Pytorch domain library for recommendation systems
PyTorch extensions for fast R&D prototyping and Kaggle farming
Multi-Modal Neural Networks for Semantic Search, based on Mid-Fusion
A unified framework for scalable computing
Powering Amazon custom machine learning chips
Open-source tool designed to enhance the efficiency of workloads
LMDeploy is a toolkit for compressing, deploying, and serving LLMs
Neural Network Compression Framework for enhanced OpenVINO
Phi-3.5 for Mac: Locally-run Vision and Language Models
Probabilistic reasoning and statistical analysis in TensorFlow
Libraries for applying sparsification recipes to neural networks
Gaussian processes in TensorFlow
Simplifies the local serving of AI models from any source
Easiest and laziest way for building multi-agent LLMs applications
Efficient few-shot learning with Sentence Transformers
Data manipulation and transformation for audio signal processing