Standardized Serverless ML Inference Platform on Kubernetes
Deep learning optimization library: makes distributed training easy
Replace OpenAI GPT with another LLM in your app
Unified Model Serving Framework
DoWhy is a Python library for causal inference
Phi-3.5 for Mac: Locally-run Vision and Language Models
The Triton Inference Server provides an optimized cloud
Library for serving Transformers models on Amazon SageMaker
Multilingual Automatic Speech Recognition with word-level timestamps
Bring the notion of Model-as-a-Service to life
State-of-the-art Parameter-Efficient Fine-Tuning
20+ high-performance LLMs with recipes to pretrain, finetune at scale
Sparsity-aware deep learning inference runtime for CPUs
A Pythonic framework to simplify AI service building
State-of-the-art diffusion models for image and audio generation
Create HTML profiling reports from pandas DataFrame objects
Python Package for ML-Based Heterogeneous Treatment Effects Estimation
A lightweight vision library for performing large object detection
Easiest and laziest way for building multi-agent LLMs applications
Integrate, train and manage any AI models and APIs with your database
Trainable, memory-efficient, and GPU-friendly PyTorch reproduction
Superduper: Integrate AI models and machine learning workflows
Low-latency REST API for serving text-embeddings
A unified framework for scalable computing
A set of Docker images for training and serving models in TensorFlow