Arize Phoenix vs Braintrust vs Langfuse vs LangSmith

A side-by-side look at scores, pricing and features — with RECATOOLS' ASEAN-aware verdict for each.

Arize Phoenix Open-source LLM evaluation and observability Visit Braintrust LLM evaluation and prompt management for production Visit Langfuse Open-source LLM observability and analytics Visit LangSmith LLM observability and evaluation from the LangChain team Visit
RECATOOLS Score 7.8 / 10 7.5 / 10 8.4 / 10 7.9 / 10
Capability 8 8 8 8
Value for money 9 7 9 7
Ease of use 6 7 7 8
ASEAN readiness 8 6 7 6
API quality 8 8 9 8
Pricing Open Source Freemium Open Source Freemium
Free tier
Paid from
Has API
Open source
Free to use
Users
Founded 2023 2023 2023 2023
Maker
Verdict

Arize Phoenix is an open-source observability and evaluation toolkit for LLM and AI applications, offering tracing, prompt and RAG debugging, evals, and OpenTelemetry-based instrumentation that runs locally or self-hoste...

Braintrust is an evaluation, observability and prompt-iteration platform for teams building LLM-powered products, offering systematic evals, logging, datasets, scoring and a playground that bring engineering rigour to ot...

Langfuse is a leading open-source LLM observability platform covering tracing, prompt management, evaluations, cost tracking and analytics, with framework-agnostic SDKs that work well alongside LangChain, LlamaIndex or r...

LangSmith is LangChain's first-party platform for tracing, debugging, evaluating and monitoring LLM applications, with the tightest integration into LangChain and LangGraph of any observability tool. It suits teams alrea...

Full review → Full review → Full review → Full review →
← Back to AI Directory

Comparisons cover up to 4 tools. Scores are RECATOOLS editorial assessments; verify current pricing on each vendor's site.