Helps developers deploy LangChain runnables and chains as a REST API
Ready-to-run cloud templates for RAG
Pruna is a model optimization framework built for developers
Open-source LLM load balancer and serving platform for hosting LLMs
Telegram Drive
Jupyter notebook tutorials for OpenVINO
Running large language models on a single GPU
Learn how to develop, deploy and iterate on production-grade ML
Defang CLI and sample projects
Low-latency REST API for serving text-embeddings
Tribuo - A Java machine learning library
Probabilistic reasoning and statistical analysis in TensorFlow
Fast SQL-based BI tool for real-time dashboards and analytics
ID-based RAG FastAPI: Integration with Langchain and PostgreSQL
Open platform for sharing and discovering Stable Diffusion models
TFX is an end-to-end platform for deploying production ML pipelines
Replace OpenAI GPT with another LLM in your app
NLP Cloud serves high performance pre-trained or custom models for NER
C#/.NET binding of llama.cpp, including LLaMa/GPT model inference
An open-source, code-first Java toolkit
High-performance Inference and Deployment Toolkit for LLMs and VLMs
Your Personal AI Assistant; easy to install, deploy on local or coud
Desktop app that provides a graphical interface for OpenClaw AI
Easiest and laziest way for building multi-agent LLMs applications
Build and run agents you can see, understand and trust