Arize Phoenix vs Braintrust vs Langfuse vs Opik

A side-by-side look at scores, pricing and features — with RECATOOLS' ASEAN-aware verdict for each.

Arize Phoenix Open-source LLM evaluation and observability Visit Braintrust LLM evaluation and prompt management for production Visit Langfuse Open-source LLM observability and analytics Visit Opik Open-source LLM tracing, evals and guardrails you can self-host Visit
RECATOOLS Score 7.8 / 10 7.5 / 10 8.4 / 10 7.8 / 10
Capability 8 8 8 8
Value for money 9 7 9 8.5
Ease of use 6 7 7 7.5
ASEAN readiness 8 6 7 6.5
API quality 8 8 9 8
Pricing Open Source Freemium Open Source Open Source
Free tier
Paid from
Has API
Open source
Free to use
Users
Founded 2023 2023 2023
Maker
Verdict

Arize Phoenix is an open-source observability and evaluation toolkit for LLM and AI applications, offering tracing, prompt and RAG debugging, evals, and OpenTelemetry-based instrumentation that runs locally or self-hoste...

Braintrust is an evaluation, observability and prompt-iteration platform for teams building LLM-powered products, offering systematic evals, logging, datasets, scoring and a playground that bring engineering rigour to ot...

Langfuse is a leading open-source LLM observability platform covering tracing, prompt management, evaluations, cost tracking and analytics, with framework-agnostic SDKs that work well alongside LangChain, LlamaIndex or r...

Opik is the rare open-source LLM observability tool where the free self-hosted version isn't a stripped-down teaser — Comet's own docs describe it as the same codebase as the hosted product. Tracing, LLM-as-a-judge evalu...

Full review → Full review → Full review → Full review →
← Back to AI Directory

Comparisons cover up to 4 tools. Scores are RECATOOLS editorial assessments; verify current pricing on each vendor's site.