← Articles

Comparison · Updated October 2026

Langfuse Alternatives (2026): 7 Tools Compared

Why teams leave Langfuse, what each alternative is best at, verified pricing and licensing as of October 2026, and when Langfuse is still the right call.

Macy Mody
· 8 minute read
Many thin grey trace lines gathered into a few tidy blue bundles that each end in a ranked marker

The best Langfuse alternative depends on why you are leaving. If you are on LangChain or LangGraph and want failures found and fixed for you, look at LangSmith. If your work is mostly evaluation and experiments, look at Braintrust. If you want an OpenTelemetry-native tool you can run yourself, look at Arize Phoenix or Opik. If you are already on Datadog, its Agent Observability keeps everything in one place. And if the real problem is that nobody has time to read the traces at all, you may not need a different tracing tool. You need something on top of it that tells you what to fix.

Below is an honest comparison of the main options as of October 2026, including when Langfuse is still the right choice. We checked pricing, licensing, and features against each vendor’s own pricing pages and documentation on October 1, 2026. Pricing changes often, so confirm before you buy.

Why teams look for a Langfuse alternative

Langfuse is a good product, and for many teams it is the default for open-source LLM tracing. The core is MIT-licensed, self-hosting is free, and the cloud plans are inexpensive. People who leave usually do so for one of four reasons.

Self-hosting is more work than it looks. A production Langfuse deployment runs a web service and a worker alongside Postgres, ClickHouse, Redis, and object storage. That is fine for a platform team and a real burden for a three-person startup. Some enterprise features, such as audit logs, data retention policies, and SCIM, also require a license key when self-hosted.

Nobody reads the traces. Langfuse gives you an excellent view of each trace, plus dashboards, evaluators, and datasets. What it does not do, as of this writing, is go through your production traffic on its own and tell you which failures are recurring. Several competitors shipped exactly that in 2026: LangSmith Engine, Arize Signal, Braintrust Topics, and similar features from Laminar and Latitude.

Evaluation workflows. Teams whose day-to-day work is running experiments, comparing prompts, and gating releases on eval scores sometimes prefer a tool built around that loop.

Ownership. ClickHouse acquired Langfuse in January 2026. Langfuse has said it plans no licensing changes and that self-hosting remains first-class, so this is a reason to watch rather than a reason to move. Some teams still prefer to re-evaluate after an acquisition.

Langfuse alternatives at a glance

ToolLicenseFree tierFirst paid tierFinds recurring failures automatically
LangfuseMIT core50k units/month$29/month, 100k unitsNo built-in equivalent found
LangSmithProprietary1 seat, 5k traces/month$39 per seat/month, 10k tracesYes: Engine, with proposed fixes
BraintrustProprietary1 GB, 10k scores$249/monthYes: Topics; Patterns in preview
Arize Phoenix / AXPhoenix: Elastic License 2.0AX: 25k spans/monthAX Pro: $50/monthAX: Signal (scheduled)
Opik (Comet)Apache-2.025k spans/month$19/month, 100k spansOllie (interactive, on request)
LaminarApache-2.01 GB data$30/monthYes: Signals you define
Datadog Agent ObservabilityProprietary—$160/month (annual) for 100k LLM spansTopic clustering and Insights
PapayaProprietary1 agent, 1,000 traces/month$500/month, 3 agents, 50k tracesYes, per business use case, ranked by impact

A note on units: Langfuse counts traces, observations, and scores together as “units”; LangSmith counts traces; Braintrust meters gigabytes and scores; Arize, Opik, and Datadog meter spans; Laminar meters gigabytes plus analysis credits. A multi-step agent produces one trace but many spans, so a span-metered plan fills up much faster than the headline number suggests.

The alternatives, by what they are best for

Papaya: best if you want ranked fixes instead of more traces to read

Papaya is the alternative to consider when the real problem is not your tracing tool but the hours it takes to turn traces into decisions. Instead of a trace viewer, you get a short list of what to fix. Papaya classifies every run by the business use case it served, compares succeeding and failing runs with 200+ research-backed analyses, and ranks each finding by its impact on quality, latency, and cost. It reads Langfuse-format traces directly, so you can move without re-instrumenting or keep Langfuse alongside it. The free plan covers one agent and 1,000 traces a month; Business is $500 a month, never per seat. More on how it differs.

LangSmith: best if you are on LangChain and want fixes proposed for you

LangSmith is LangChain’s agent engineering platform: tracing, evaluation, monitoring, annotation queues, and deployment. Its most important 2026 addition is Engine, which scans your tracing projects on a schedule, clusters failures into prioritized issues, and, with a connected repository, proposes fixes as pull requests along with evaluators and dataset examples. The trade-offs are per-seat pricing, short default retention on base traces, and self-hosting only as an Enterprise add-on. It works with other frameworks, but it is at its best inside the LangChain ecosystem.

Braintrust: best for evaluation-heavy teams

Braintrust is built around one data model shared by production logs and experiments, so a failing production trace can become a test case in a few clicks. It has online scoring, alerts, an OpenTelemetry endpoint, and Loop, an assistant for investigating logs in natural language. Topics clusters traces by task, sentiment, and issues, and Patterns, in preview, looks for recurring problems with suggested fixes. It is closed source, and the jump from the free tier to the $249 Pro plan has no middle step.

Arize Phoenix and Arize AX: best for OpenTelemetry-first teams

Phoenix is Arize’s free, self-hostable platform built on OpenTelemetry and OpenInference, with tracing, evals, datasets, experiments, and prompt management. Its license is Elastic License 2.0, which is source-available rather than OSI open source and prohibits offering it as a hosted service to third parties. Arize AX is the commercial product, and it includes Signal, which scans traces on a schedule and groups recurring failures into ranked issues with suggested fixes. Dynatrace completed its acquisition of Arize on October 1, 2026, and says Phoenix will continue as an open-source project.

Opik: best open-source option on a budget

Opik, from Comet, is Apache-2.0 licensed and self-hostable, with trace trees, datasets, experiments, LLM-as-a-judge metrics, online evaluation rules that can trigger alerts, and an agent optimizer. Its cloud Pro plan is the cheapest paid tier here at $19 a month. Its assistant, Ollie, compares failing and successful runs and can propose code changes, but it works interactively, when you ask, rather than scanning production on its own.

Laminar: best for long-running and browser agents

Laminar is an Apache-2.0, OpenTelemetry-native platform aimed at agents, with SQL access to traces, evals, and Signals: plain-language detectors that run over traces, return structured events, and cluster similar events automatically. Signals are powerful, but you write them, and they are billed as credits on top of the plan.

Datadog Agent Observability: best if you already live in Datadog

Datadog renamed its LLM Observability product to Agent Observability in 2026. Its main advantage is correlation: agent traces sit next to your APM, logs, and infrastructure metrics. It includes evaluations, datasets, topic clustering of production traffic, and automatic Insights on cost and reliability issues. It is proprietary and priced per LLM span, which can add up at agent scale.

Two tools that often appear on older lists need a caveat. Helicone was acquired by Mintlify in March 2026 and is in maintenance mode, so it is hard to recommend for a new project. Galileo has been acquired by Cisco and is becoming Splunk Agent Observability, so check its standalone plans before committing.

When Langfuse is still the right call

Stay on Langfuse if you want genuinely open-source tracing you control, your team is comfortable running ClickHouse, you value prompt management alongside tracing, and you have people who actually review traces and maintain evaluators. Its pricing is among the most predictable here, and nothing on this list is clearly better at being an open, self-hosted system of record for LLM traces.

Moving tracing tools is real work. Before you migrate, be specific about what is missing. If the answer is “we have the data but nobody knows what to fix,” a new tracing tool will give you the same problem in a different UI.

When you need fixes, not another trace viewer

Every tool above starts from the same idea: give engineers visibility, then help them investigate. Papaya starts from the other end. Instead of showing you more detail, it tries to show you only what you need to fix.

It works with the traces you already have. Papaya reads traces in Langfuse, LangSmith, Arize Phoenix, and Braintrust formats, collects them through its own SDK, and maps other formats automatically, with no data preparation on your side. It classifies every run into the business use case it was serving, because the changes that improve a booking workflow are not the changes that improve a refund workflow. Within each use case it splits runs into succeeding and failing cohorts, clusters them, and runs 200+ research-backed analyses to find what separates them: failing tool calls, hallucination-prone steps, sub-agents that flood the context, workflows that are not working at all. Each finding is quantified, so you fix the most valuable ones first, and with code access Papaya can open the pull request.

You can keep Langfuse as your system of record and put Papaya on top. You do not have to choose between them.

Try it on your own traces

Papaya’s free plan covers one agent and 1,000 traces a month, enough to see what it finds in a real workflow. See pricing or talk to us.

Frequently asked questions

What is the best Langfuse alternative?

It depends on why you are switching. LangSmith suits LangChain users who want failures clustered and fixes proposed. Braintrust suits evaluation-heavy teams. Arize Phoenix and Opik suit teams that want OpenTelemetry-based tools they can self-host. Datadog suits teams already on Datadog. If the gap is knowing what to fix rather than tracing, Papaya works on top of your existing traces.

Is there a free, open-source alternative to Langfuse?

Yes. Opik is Apache-2.0 licensed and self-hostable, as is Laminar. Arize Phoenix is free to self-host but uses the Elastic License 2.0, which is source-available rather than OSI open source. Langfuse itself is MIT-licensed at its core.

Does Langfuse detect failures automatically?

Langfuse provides tracing, dashboards, alerts on metric thresholds, LLM-as-a-judge and code evaluators, and datasets. As of October 2026 we did not find a built-in feature that scans production traces and clusters recurring failures on its own, comparable to LangSmith Engine or Arize Signal.

Who owns Langfuse now?

ClickHouse acquired Langfuse in January 2026. Langfuse has said it plans no licensing changes, that self-hosting remains first-class, and that its cloud product continues unchanged.

Can I use Papaya with Langfuse?

Yes. Papaya reads traces in Langfuse format, as well as LangSmith, Arize Phoenix, and Braintrust formats, or collects them through its own SDK. You can keep Langfuse as your tracing system and use Papaya to find, rank, and fix recurring failures across each business use case.

Sources used in this article

Related