Related Products
|
||||||
About
OpenAI- and Anthropic-compatible inference API from an EU company. The flagship model runs on dedicated GPUs in EIA data centres with zero data retention: prompts and completions are processed in memory only, not stored, not logged, not used for training. Routed open models from third-party providers are available with the same key and clearly labelled. One DPA and one invoice from an EU company. Features: streaming, tool calling, structured output, public DPA and sub-processor list, per-token pricing. Measured on the live system in August 2026: 176 tokens per second per stream, first token in 0.3 seconds.
|
About
WebLLM is a high-performance, in-browser language model inference engine that leverages WebGPU for hardware acceleration, enabling powerful LLM operations directly within web browsers without server-side processing. It offers full OpenAI API compatibility, allowing seamless integration with functionalities such as JSON mode, function-calling, and streaming. WebLLM natively supports a range of models, including Llama, Phi, Gemma, RedPajama, Mistral, and Qwen, making it versatile for various AI tasks. Users can easily integrate and deploy custom models in MLC format, adapting WebLLM to specific needs and scenarios. The platform facilitates plug-and-play integration through package managers like NPM and Yarn, or directly via CDN, complemented by comprehensive examples and a modular design for connecting with UI components. It supports streaming chat completions for real-time output generation, enhancing interactive applications like chatbots and virtual assistants.
|
|||||
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
|||||
Audience
Companies that need LLM inference under EU data-protection requirements; developers of AI agents and RAG applications
|
Audience
Developers seeking a tool to implement high-performance, in-browser language model inference without relying on server-side processing
|
|||||
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Supported
|
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Supported
|
|||||
API
Offers API
Not Supported
|
API
Offers API
Supported
|
|||||
Screenshots and VideosNo images available
|
Screenshots and Videos |
|||||
Pricing
$0.04 per 1M input tokens
Free Version
Supported
Free Trial
Not Supported
|
Pricing
Free
Free Version
Supported
Free Trial
Not Supported
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Not Supported
In Person
Not Supported
|
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Not Supported
In Person
Not Supported
|
|||||
Company InformationHeabsy
Founded: 2014
Slovakia
heabsy.com
|
Company InformationWebLLM
webllm.mlc.ai/
|
|||||
Alternatives |
Alternatives |
|||||
|
|
||||||
|
|
||||||
Categories |
Categories |
|||||
Integrations
Alpaca
Not Supported
Gemma
Not Supported
Llama
Not Supported
Llama 2
Not Supported
Llama 3
Not Supported
Llama 3.3
Not Supported
Mathstral
Not Supported
Ministral 8B
Not Supported
Mistral AI
Not Supported
Mistral Large
Not Supported
|
Integrations
Alpaca
Supported
Gemma
Supported
Llama
Supported
Llama 2
Supported
Llama 3
Supported
Llama 3.3
Supported
Mathstral
Supported
Ministral 8B
Supported
Mistral AI
Supported
Mistral Large
Supported
|
|||||
|
|
|