Audience

Developers and companies searching for an inference server solution to improve AI production

About NVIDIA Triton Inference Server

NVIDIA Triton™ inference server delivers fast and scalable AI in production. Open-source inference serving software, Triton inference server streamlines AI inference by enabling teams deploy trained AI models from any framework (TensorFlow, NVIDIA TensorRT®, PyTorch, ONNX, XGBoost, Python, custom and more on any GPU- or CPU-based infrastructure (cloud, data center, or edge). Triton runs models concurrently on GPUs to maximize throughput and utilization, supports x86 and ARM CPU-based inferencing, and offers features like dynamic batching, model analyzer, model ensemble, and audio streaming. Triton helps developers deliver high-performance inference aTriton integrates with Kubernetes for orchestration and scaling, exports Prometheus metrics for monitoring, supports live model updates, and can be used in all major public cloud machine learning (ML) and managed Kubernetes platforms. Triton helps standardize model deployment in production.

Pricing

Starting Price:
Free
Free Version:
Free Version available.

Integrations

Ratings/Reviews

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Company Information

NVIDIA
United States
developer.nvidia.com/nvidia-triton-inference-server

Videos and Screen Captures

Other Useful Business Software
Custom VMs From 1 to 96 vCPUs With 99.95% Uptime Icon
Custom VMs From 1 to 96 vCPUs With 99.95% Uptime

General-purpose, compute-optimized, or GPU/TPU-accelerated. Built to your exact specs.

Live migration and automatic failover keep workloads online through maintenance. One free e2-micro VM every month.
Start Free

Product Details

Platforms Supported
Windows
Mac
Linux
Training
Documentation
In Person
Videos
Support
Phone Support
Online

NVIDIA Triton Inference Server Frequently Asked Questions

Q: What kinds of users and organization types does NVIDIA Triton Inference Server work with?
Q: What languages does NVIDIA Triton Inference Server support in their product?
Q: What kind of support options does NVIDIA Triton Inference Server offer?
Q: What other applications or services does NVIDIA Triton Inference Server integrate with?
Q: What type of training does NVIDIA Triton Inference Server provide?
Q: How much does NVIDIA Triton Inference Server cost?

NVIDIA Triton Inference Server Product Features

Artificial Intelligence

For Healthcare Not Supported
Multi-Language Not Supported
Chatbot Not Supported
For Sales Not Supported
Rules-Based Automation Not Supported
Machine Learning Not Supported
Natural Language Processing Not Supported
Predictive Analytics Not Supported
Virtual Personal Assistant (VPA) Not Supported
Process/Workflow Automation Not Supported
For eCommerce Not Supported
Image Recognition Not Supported

Machine Learning

Deep Learning Not Supported
ML Algorithm Library Not Supported
Model Training Not Supported
Natural Language Processing (NLP) Not Supported
Predictive Modeling Not Supported
Templates Not Supported
Visualization Not Supported
Statistical / Mathematical Tools Not Supported

NVIDIA Triton Inference Server Additional Categories