Browse free open source Python LLM Inference Tools and projects below. Use the toggles on the left to filter open source Python LLM Inference Tools by OS, license, language, programming language, and project status.
Implementation of model parallel autoregressive transformers on GPUs
GPU environment management and cluster orchestration
Build your chatbot within minutes on your favorite device
LMDeploy is a toolkit for compressing, deploying, and serving LLMs
Easiest and laziest way for building multi-agent LLMs applications
A Pythonic framework to simplify AI service building
Multi-LoRA inference server that scales to 1000s of fine-tuned LLMs
Official inference library for Mistral models
Trainable, memory-efficient, and GPU-friendly PyTorch reproduction
Trainable models and NN optimization tools
Create HTML profiling reports from pandas DataFrame objects
Simplifies the local serving of AI models from any source
Implementation of "Tree of Thoughts
Serve machine learning models within a Docker container
Integrate, train and manage any AI models and APIs with your database
A library to communicate with ChatGPT, Claude, Copilot, Gemini
Pytorch domain library for recommendation systems
Replace OpenAI GPT with another LLM in your app
Run any Llama 2 locally with gradio UI on GPU or CPU from anywhere
Tensor search for humans
Optimizing inference proxy for LLMs
OpenFieldAI is an AI based Open Field Test Rodent Tracker
A graphical manager for ollama that can manage your LLMs
GUI shell for running local LLM on desktop
Openai style api for open large language models