
Langfuse
by Langfuse · Open-source tracing, evaluation and prompt management for LLM apps
70.2 — BenchRank score out of 100
Screenshots of Langfuse
Homepage Pricing page
Overview
Langfuse is an open-source platform for tracing and evaluating LLM applications and agents. It records hierarchical traces of LLM calls, tool invocations and retrieval steps, and adds prompt management, datasets, experiments, LLM-as-a-judge scoring and human annotation, with cost and latency dashboards. It runs as a hosted cloud service or self-hosted via Docker, Kubernetes or Terraform.
- Best for
- Engineering teams tracing, evaluating and improving LLM apps who want open source and self-hosting.
- Pricing
- Free Hobby plan, then $29/month (Core), $199/month (Pro) and $2,499/month (Enterprise), each including 100k units with additional usage at $8 per 100k units; self-hosting is free under the MIT licence.
- Runs on
- Self-hosted
Strengths and trade-offs
Strengths
- MIT licensed; self-host via Docker, Kubernetes or Terraform
- OTel-native, with SDKs and 100+ framework integrations
- Tracing, prompts, evals and dashboards in one platform
- Free tier: 50k units/month, no credit card required
Trade-offs
- Data access capped: 30 days on Hobby, 90 on Core, 3 years on Pro
- Included usage stays 100k units on all paid plans; extra is $8/100k
- SSO and fine-grained RBAC need the $300/mo Teams add-on or Enterprise
- Hobby is 2 users with GitHub-only support and no response-time SLO
Pricing
Published plans from Langfuse’s own pricing page, in USD. Usage charges and add-ons may apply on top.
| Plan | Monthly | Includes |
|---|---|---|
| Hobby | Freeper month |
|
| Core | $29per month |
|
| Pro | $199per month |
|
| Enterprise | Contact sales |
|
How this score is made up
Each dimension is scored out of 100 and combined into the headline score using fixed weights.
- MCP support
- 100 out of 100
- API quality
- 0 out of 100
- Documentation
- 100 out of 100
- Agent friendliness
- 73 out of 100
- Pricing transparency
- 85 out of 100
- Changelog
- 100 out of 100
- Marketing site structure
- 55 out of 100
- Operational trust
- 60 out of 100
Measured, but not part of the score
Useful to know, but not a mark for or against the product — so these do not affect the ranking.
- Openness
- 55 out of 100
- Maintenance
- 100 out of 100
This doesn’t look right — report a problem with Langfuse’s score
Alternatives in Model Hosting & Inference
Ranked 1
77.3 — BenchRank score out of 100Phoenix
Arize Phoenix · Open-source tracing, evaluation and experimentation for AI agents
Best for: AI engineers who need to trace, evaluate and iterate on LLM agents on their own infrastructure
Ranked 2
72.5 — BenchRank score out of 100Helicone
Helicone · AI gateway and LLM observability for routing, debugging and analysing apps
Best for: AI engineering teams routing, debugging and monitoring LLM calls across many providers
Ranked 3
72.2 — BenchRank score out of 100Replicate
Replicate · Run, fine-tune and deploy AI models through a cloud API
Best for: Developers who want to run, fine-tune or deploy AI models via an API without managing GPUs

