Related Products
|
||||||
About
OpenAI- and Anthropic-compatible inference API from an EU company. The flagship model runs on dedicated GPUs in EIA data centres with zero data retention: prompts and completions are processed in memory only, not stored, not logged, not used for training. Routed open models from third-party providers are available with the same key and clearly labelled. One DPA and one invoice from an EU company. Features: streaming, tool calling, structured output, public DPA and sub-processor list, per-token pricing. Measured on the live system in August 2026: 176 tokens per second per stream, first token in 0.3 seconds.
|
About
Xinference is an enterprise AI inference platform for teams that want to run open models without building the serving stack themselves. Most teams start on the Model API: 300+ open models behind one OpenAI-compatible endpoint hosted in Australia. Switching from an existing provider takes two lines of code. As usage grows, the same workloads move to Dedicated Inference on reserved GPUs, or to a private deployment inside the customer's own cloud or data centre. Every deployment includes one control plane: per-request logs, live TTFT and TPOT monitoring, role-based access control, audit logs and SSO. Xinference does not train on customer data and does not retain it by default. Common workloads include enterprise RAG, customer assistants, agents and function calling, coding assistance, document extraction, speech and image generation.
|
|||||
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
|||||
Audience
Companies that need LLM inference under EU data-protection requirements; developers of AI agents and RAG applications
|
Audience
Enterprise engineering and platform teams running AI in production
|
|||||
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Supported
|
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Not Supported
|
|||||
API
Offers API
Not Supported
|
API
Offers API
Supported
|
|||||
Screenshots and VideosNo images available
|
Screenshots and VideosNo images available
|
|||||
Pricing
$0.04 per 1M input tokens
Free Version
Supported
Free Trial
Not Supported
|
Pricing
No information available.
Free Version
Not Supported
Free Trial
Not Supported
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Not Supported
In Person
Not Supported
|
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Not Supported
In Person
Not Supported
|
|||||
Company InformationHeabsy
Founded: 2014
Slovakia
heabsy.com
|
Company InformationXinference
Founded: 2026
Australia
xinference.co
|
|||||
Alternatives |
Alternatives |
|||||
|
|
||||||
Categories |
Categories |
|||||
Integrations
No info available.
|
Integrations
No info available.
|
|||||
|
|
|