An MLOps framework to package, deploy, monitor and manage models
Operating LLMs in production
Deploy your agentic worfklows to production
RF-DETR is a real-time object detection and segmentation
Jupyter notebook tutorials for OpenVINO
Probabilistic reasoning and statistical analysis in TensorFlow
MLOps simplified. From ML Pipeline ⇨ Data Product without the hassle
Low-latency REST API for serving text-embeddings
Ready-to-run cloud templates for RAG
Training and deploying machine learning models on Amazon SageMaker
Replace OpenAI GPT with another LLM in your app
Gen-AI Chat for Teams
Private AI platform for agents, enterprise search and RAG pipelines
Running large language models on a single GPU
Learn how to develop, deploy and iterate on production-grade ML
A guidance language for controlling large language models
ID-based RAG FastAPI: Integration with Langchain and PostgreSQL
Pruna is a model optimization framework built for developers
20+ high-performance LLMs with recipes to pretrain, finetune at scale
Build high-quality LLM apps
Deploy reasoning AI agents powered by agentic graph RAG in minutes
Full-stack AI Red Teaming platform
TFX is an end-to-end platform for deploying production ML pipelines
Easiest and laziest way for building multi-agent LLMs applications
Cybersecurity AI (CAI), the framework for AI Security