simple web server free download

KServe

Standardized Serverless ML Inference Platform on Kubernetes

...It aims to solve production model serving use cases by providing performant, high abstraction interfaces for common ML frameworks like Tensorflow, XGBoost, ScikitLearn, PyTorch, and ONNX. It encapsulates the complexity of autoscaling, networking, health checking, and server configuration to bring cutting edge serving features like GPU Autoscaling, Scale to Zero, and Canary Rollouts to your ML deployments. It enables a simple, pluggable, and complete story for Production ML Serving including prediction, pre-processing, post-processing and explainability. KServe is being used across various organizations.

Downloads: 3 This Week

Last Update: 2026-03-13

See Project

API-for-Open-LLM

Openai style api for open large language models

API-for-Open-LLM is a lightweight API server designed for deploying and serving open large language models (LLMs), offering a simple way to integrate LLMs into applications.

Downloads: 0 This Week

Last Update: 2025-01-22

See Project

LLaVA

Visual Instruction Tuning: Large Language-and-Vision Assistant

Visual instruction tuning towards large language and vision models with GPT-4 level capabilities.

Downloads: 9 This Week

Last Update: 2024-02-04

See Project

Hugging Face Transformer

CPU/GPU inference server for Hugging Face transformer models

Optimize and deploy in production Hugging Face Transformer models in a single command line. At Lefebvre Dalloz we run in-production semantic search engines in the legal domain, in the non-marketing language it's a re-ranker, and we based ours on Transformer. In that setup, latency is key to providing a good user experience, and relevancy inference is done online for hundreds of snippets per user query. Most tutorials on Transformer deployment in production are built over Pytorch and FastAPI....

Downloads: 0 This Week

Last Update: 2022-08-22

See Project

BudgetML

Deploy a ML inference service on a budget in 10 lines of code

...It is by no means meant to be used in a full-fledged production-ready setup. It is simply a means to get a server up and running as fast as possible with the lowest costs possible.

Downloads: 0 This Week

Last Update: 2022-08-26

See Project

Search Results for "simple web server"

Showing 5 open source projects for "simple web server"

KServe

API-for-Open-LLM

LLaVA

Hugging Face Transformer

BudgetML

Search Results for "simple web server"

Showing 5 open source projects for "simple web server"

KServe

API-for-Open-LLM

LLaVA

Hugging Face Transformer

BudgetML

Related Categories