Run Local LLMs on Any Device. Open-source
A high-throughput and memory-efficient inference and serving engine
State-of-the-art diffusion models for image and audio generation
Sparsity-aware deep learning inference runtime for CPUs
Everything you need to build state-of-the-art foundation models
Official inference library for Mistral models
20+ high-performance LLMs with recipes to pretrain, finetune at scale
Single-cell analysis in Python
The official Python client for the Huggingface Hub
A toolkit to optimize ML models for deployment for Keras & TensorFlow
Training and deploying machine learning models on Amazon SageMaker
Operating LLMs in production
Superduper: Integrate AI models and machine learning workflows
Python Package for ML-Based Heterogeneous Treatment Effects Estimation
Large Language Model Text Generation Inference
A library for accelerating Transformer models on NVIDIA GPUs
Standardized Serverless ML Inference Platform on Kubernetes
Trainable models and NN optimization tools
A lightweight vision library for performing large object detection
Multi-Modal Neural Networks for Semantic Search, based on Mid-Fusion
Uplift modeling and causal inference with machine learning algorithms
DoWhy is a Python library for causal inference
Uncover insights, surface problems, monitor, and fine tune your LLM
Replace OpenAI GPT with another LLM in your app
LLM training code for MosaicML foundation models