Back to Langfuse

Alternatives to Langfuse

30 tools that compete with or replace Langfuse. Ranked by direct product-type match — not generic category overlap.

Last updated
Cross-checked through our multi-step verification ·
Arize Phoenix

Arize Phoenix

Open-source LLM agent observability with tracing, evals, and experiments

FreemiumTry
Dash0

Dash0

OpenTelemetry-native observability with autonomous AI SRE Agent0, plus AI Coding Insights to monitor coding agents in production.

FreemiumTry
Phoenix

Phoenix

Open-source observability and evaluation for AI agents.

FreemiumTry
Lilypad

Lilypad

Open-source OpenTelemetry LLM observability for Python devs

FreeTry
WhyLabs

WhyLabs

Open-source AI observability with privacy-preserving logging and LLM security toolkit

FreeTry
Evidently AI

Evidently AI

Open-source AI evaluation and observability for LLMs, RAG, agents, and ML models.

FreemiumTry
Opik (Comet)

Opik (Comet)

Open-source AI observability and evals for the agentic era

FreemiumTry
Langfuse Prompt Experiments

Langfuse Prompt Experiments

Open-source LLM engineering platform for observability, prompt management, and evaluation—self-hostable or cloud.

FreemiumTry
Galileo AI Evals

Galileo AI Evals

AI observability and eval engineering platform that turns offline evals into production guardrails.

FreemiumTry
TruLens

TruLens

Open-source OpenTelemetry-native framework for evaluating and tracing AI agents

FreeTry
MLflow

MLflow

Open source AI engineering platform for building, debugging, and monitoring agents, LLMs, and ML models.

FreeTry
Compare Langfuse vs MLflow
Agenta

Agenta

Open-source workspace to build, evaluate, and deploy AI agents

FreemiumTry
PostHog

PostHog

Open-source product OS unifying analytics, session replay, feature flags, and a data warehouse.

FreemiumTry
RAGAS

RAGAS

Open-source framework for systematic LLM evaluation replacing vibe checks

FreeTry
Opencompass

Opencompass

Open-source LLM and VLM evaluation platform for standardized benchmarking

FreeTry
Token Monitor

Token Monitor

Free open-source widget tracking tokens, cost, and limits across 28+ AI coding tools.

FreeTry
OpenLIT

OpenLIT

Free, self-hosted OpenTelemetry LLM observability and AI engineering platform for teams

FreemiumTry
Tokentelemetry

Tokentelemetry

Free local observability for AI coding agents — track tokens, cost & traces, no cloud.

FreeTry
Burntop

Burntop

Free open-source AI usage tracking & analytics for developers

FreeTry
LangSmith

LangSmith

LangSmith: AI agent observability, tracing, and evaluation platform

FreemiumTry
Comet

Comet

AI observability and evals that auto-fix agent code via git

FreemiumTry
Confident AI

Confident AI

Enterprise AI quality platform unifying LLM evaluation, observability, red teaming, and governance.

FreemiumTry
Fiddler AI

Fiddler AI

Enterprise AI control plane uniting agentic observability, guardrails, and governance

FreemiumTry
Honeycomb Query Assistant

Honeycomb Query Assistant

Turn plain English into production-ready Honeycomb queries for faster debugging.

FreemiumTry
Galileo

Galileo

Turn offline evals into real-time AI guardrails with observability and eval engineering.

FreemiumTry
Neptune.ai

Neptune.ai

Real-time experiment tracking for frontier AI model training, now owned by OpenAI

Contact SalesTry
AgentOps

AgentOps

Trace, debug, and deploy reliable AI agents with full observability

FreemiumTry
Braintrust

Braintrust

Active observability for AI agents: trace, evaluate, and discover patterns at scale.

FreemiumTry
Helicone

Helicone

AI gateway & LLM observability for routing, debugging, and analyzing AI apps

FreemiumTry
Metoro

Metoro

Autonomous AI SRE agent for Kubernetes with eBPF observability and automated fix PRs.

FreemiumTry

Frequently asked questions

What are the best alternatives to Langfuse?

We currently list 30 alternatives to Langfuse: Arize Phoenix, Dash0, Phoenix, Lilypad, WhyLabs. Each is ranked by direct product-type match rather than generic category overlap.

How do you choose which Langfuse alternatives to show?

Alternatives are ranked by direct product-type match — tools that do the same job — not by shared category tags. Every listed tool is independently re-verified on a continuous cycle.