+
+

Related Products

  • Runpod
    220 Ratings
    Visit Website
  • Gemini Enterprise Agent Platform
    984 Ratings
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • Google AI Studio
    30 Ratings
    Visit Website
  • UTunnel VPN and ZTNA
    118 Ratings
    Visit Website
  • Azore CFD
    24 Ratings
    Visit Website
  • Servers.com by Nexcess
    15 Ratings
    Visit Website
  • Ditto
    2 Ratings
    Visit Website
  • Planview ProjectAdvantage
    121 Ratings
    Visit Website
  • Dynamo Software
    71 Ratings
    Visit Website

About

Highly scalable and standards-based model inference platform on Kubernetes for trusted AI. KServe is a standard model inference platform on Kubernetes, built for highly scalable use cases. Provides performant, standardized inference protocol across ML frameworks. Support modern serverless inference workload with autoscaling including a scale to zero on GPU. Provides high scalability, density packing, and intelligent routing using ModelMesh. Simple and pluggable production serving for production ML serving including prediction, pre/post-processing, monitoring, and explainability. Advanced deployments with the canary rollout, experiments, ensembles, and transformers. ModelMesh is designed for high-scale, high-density, and frequently-changing model use cases. ModelMesh intelligently loads and unloads AI models to and from memory to strike an intelligent trade-off between responsiveness to users and computational footprint.

About

Fast, lightweight, portable, rust-powered, and OpenAI compatible. We work with cloud providers, especially edge cloud/CDN compute providers, to support microservices for web apps. Use cases include AI inference, database access, CRM, ecommerce, workflow management, and server-side rendering. We work with streaming frameworks and databases to support embedded serverless functions for data filtering and analytics. The serverless functions could be database UDFs. They could also be embedded in data ingest or query result streams. Take full advantage of the GPUs, write once, and run anywhere. Get started with the Llama 2 series of models on your own device in 5 minutes. Retrieval-argumented generation (RAG) is a very popular approach to building AI agents with external knowledge bases. Create an HTTP microservice for image classification. It runs YOLO and Mediapipe models at native GPU speed.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Developers and professionals searching for a model inference platform on Kubernetes

Audience

Developers in search of a runtime solution to build cloud-native applications

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

Free
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

KServe
kserve.github.io/website/latest/

Company Information

Second State
United States
www.secondstate.io

Alternatives

Alternatives

LM-Kit.NET

LM-Kit.NET

LM-Kit

Categories

Categories

Integrations

Docker
Kubernetes
Apache APISIX
Bloomberg
Discord
GitHub
GitLab
IBM Cloud
JavaScript
Jira
Kubeflow
Llama 2
NVIDIA DRIVE
Node.js
Notion
OpenAI
Polkadot
Python
Slack
Telegram

Integrations

Docker
Kubernetes
Apache APISIX
Bloomberg
Discord
GitHub
GitLab
IBM Cloud
JavaScript
Jira
Kubeflow
Llama 2
NVIDIA DRIVE
Node.js
Notion
OpenAI
Polkadot
Python
Slack
Telegram
Claim KServe and update features and information
Claim KServe and update features and information
Claim Second State and update features and information
Claim Second State and update features and information