+
+

Related Products

  • Grafana Cloud
    860 Ratings
    Visit Website
  • New Relic
    2,941 Ratings
    Visit Website
  • NeuBird
    2 Ratings
    Visit Website
  • Cloudflare
    2,048 Ratings
    Visit Website
  • Gemini Enterprise Agent Platform
    1,161 Ratings
    Visit Website
  • Google AI Studio
    41 Ratings
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • Auvik
    678 Ratings
    Visit Website
  • Gearset
    305 Ratings
    Visit Website
  • AdRem NetCrunch
    165 Ratings
    Visit Website

About

Langfuse is an open source LLM engineering platform to help teams collaboratively debug, analyze and iterate on their LLM Applications. Observability: Instrument your app and start ingesting traces to Langfuse Langfuse UI: Inspect and debug complex logs and user sessions Prompts: Manage, version and deploy prompts from within Langfuse Analytics: Track metrics (LLM cost, latency, quality) and gain insights from dashboards & data exports Evals: Collect and calculate scores for your LLM completions Experiments: Track and test app behavior before deploying a new version Why Langfuse? - Open source - Model and framework agnostic - Built for production - Incrementally adoptable - start with a single LLM call or integration, then expand to full tracing of complex chains/agents - Use GET API to build downstream use cases and export data

About

Scorable is an AI evaluation and monitoring platform designed to help developers measure, control, and improve the behavior of applications built with large language models. It enables teams to create customized automated evaluators, sometimes referred to as AI “judges”, that assess how an AI system responds to users and whether its outputs meet defined quality standards such as accuracy, relevance, helpfulness, tone, and policy compliance. Developers can describe what they want to measure in plain language, and the platform generates a tailored evaluation stack that tests AI outputs against context-specific criteria rather than generic benchmarks. These evaluators can be embedded directly into application code, allowing AI systems such as chatbots, retrieval-augmented generation (RAG) systems, or autonomous agents to be continuously monitored in production environments.

Platforms Supported

Windows Not Supported
Mac Not Supported
Linux Not Supported
Cloud Supported
On-Premises Supported
iPhone Not Supported
iPad Not Supported
Android Not Supported
Chromebook Not Supported

Platforms Supported

Windows Not Supported
Mac Not Supported
Linux Not Supported
Cloud Supported
On-Premises Supported
iPhone Not Supported
iPad Not Supported
Android Not Supported
Chromebook Not Supported

Audience

Software Engineers, AI Engineers, Data Scientists, Product Managers

Audience

Developers and AI product teams building LLM-powered applications who need tools to evaluate, monitor, and control the quality and reliability of AI outputs in production

Support

Phone Support Not Supported
24/7 Live Support Supported
Online Supported

Support

Phone Support Not Supported
24/7 Live Support Not Supported
Online Supported

API

Offers API Supported

API

Offers API Supported

Screenshots and Videos

Screenshots and Videos

Pricing

$29/month
Generous free tier and Pro Plans starting at USD 29/month.
Langfuse is open source so you will always be able to host the software yourself at no cost.
Free Version Supported
Free Trial Supported

Pricing

$19 per month
Free Version Supported
Free Trial Not Supported

Reviews/Ratings

Overall 5.0 / 5
ease 4.0 / 5
features 5.0 / 5
design 4.0 / 5
support 5.0 / 5

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Pros & Cons from Real Users

Pros

  • I like that this is exclusively designed for LLMs only, so it takes a lot of clutter out of having to deal with features to do with rest of the models. Scores are very comprehensive including the ones related to query/response, ability to find rag relevance, cost related metrics which other tools typically do not offer and comprehensive trace function.

Cons

  • There wasn't much but if I have to highlight something here, self hosting was difficult. And the main dashboard can have drill downs to take it to relevant sections in traces/generations etc

Training

Documentation Supported
Webinars Not Supported
Live Online Not Supported
In Person Not Supported

Training

Documentation Supported
Webinars Not Supported
Live Online Supported
In Person Not Supported

Company Information

Langfuse
Founded: 2023
Germany
langfuse.com

Company Information

Scorable
Finland
scorable.ai/

Alternatives

Alternatives

Categories

AI Observability Supported
LLM Evaluation Supported
Observability Supported
Prompt Management Supported

Categories

Integrations

Claude Supported
Dograh Supported
Flowise Supported
Hugging Face Supported
Lamatic.ai Supported
LangChain Supported
LiteLLM Supported
LlamaIndex Supported
Mirascope Supported
Model Context Protocol (MCP) Not Supported
Netguru Omega Supported
Okta Not Supported
OpenAI Supported
Python Not Supported
Slack Not Supported
TierZero Supported
TypeScript Not Supported
Voker Supported
Workers by Delos Supported

Integrations

Claude Not Supported
Dograh Not Supported
Flowise Not Supported
Hugging Face Not Supported
Lamatic.ai Not Supported
LangChain Not Supported
LiteLLM Not Supported
LlamaIndex Not Supported
Mirascope Not Supported
Model Context Protocol (MCP) Supported
Netguru Omega Not Supported
Okta Supported
OpenAI Not Supported
Python Supported
Slack Supported
TierZero Not Supported
TypeScript Supported
Voker Not Supported
Workers by Delos Not Supported
Claim Langfuse and update features and information
Claim Langfuse and update features and information
Claim Scorable and update features and information
Claim Scorable and update features and information