Best free LangSmith alternatives

Looking for a free alternative to LangSmith? Here are 9 llm eval & observability worth trying — 9 with a free tier. LangSmith itself is free tier + paid.

ToolLangSmith (you searched)Arize PhoenixGoogle StaxDeepEval
Pricingfreemiumfreefreefree
Best forTracing, evaluation datasets and prompt playground for LLM apps and agentsFully open-source, self-hostable with zero setup via `uvx` (pip/conda also supported)Build evals with human raters or LLM autoraters on your own dataOpen-source LLM evaluation framework with pytest-style tests and 50+ metrics
VisitOpen ↗Open ↗Open ↗Open ↗

LangSmith alternatives — straight answers

What is the best free alternative to LangSmith?

Arize Phoenix is the closest free alternative to LangSmith: same LLM Eval & Observability sub-category, and it is free. Fully open-source, self-hostable with zero setup via `uvx` (pip/conda also supported)

Are there free LangSmith alternatives?

Yes — 9 of the 9 alternatives listed here are free or freemium: Arize Phoenix, Google Stax, DeepEval, EleutherAI LM Evaluation Harness, OpenLLMetry (Traceloop). None of them are paid-only.

Is LangSmith free?

LangSmith is free tier + paid. There is a free tier you can work in, with paid plans above it.

How were these LangSmith alternatives chosen?

They are the other tools in LLM Eval & Observability, then the rest of AI Models & Local Execution, ordered free tiers first and capped at 9. There is no sponsorship and no paid placement in that ordering — only the pricing tier decides.

All LangSmith alternatives

Browse LLM Eval & Observability →

Arize Phoenix

LLM Eval & Observability
free

✦ WhyA LangSmith alternative — Fully open-source, self-hostable with zero setup via `uvx` (pip/conda also supported).

  • Fully open-source, self-hostable with zero setup via `uvx` (pip/conda also supported)
  • Combines tracing, evals, datasets, experiments and prompt playground in one local-first tool
  • AI engineering agent (PXI) built in for automated troubleshooting of traces

Google Stax

LLM Eval & Observability
freeNew

✦ WhyA LangSmith alternative — Build evals with human raters or LLM autoraters on your own data.

  • Build evals with human raters or LLM autoraters on your own data
  • Compare models and prompts on quality, latency and token cost
  • Prebuilt and custom evaluators

DeepEval

LLM Eval & Observability
freeNew

✦ WhyA LangSmith alternative — Open-source LLM evaluation framework with pytest-style tests and 50+ metrics.

  • Open-source LLM evaluation framework with pytest-style tests and 50+ metrics
  • RAG, agent and conversational metrics; red-teaming module
  • Apache-2.0; optional Confident AI cloud has a free tier

EleutherAI LM Evaluation Harness

LLM Eval & Observability
freeNew

✦ WhyA LangSmith alternative — Standard open framework for running hundreds of benchmarks (MMLU, GSM8K and more) against any LLM.

  • Standard open framework for running hundreds of benchmarks (MMLU, GSM8K and more) against any LLM
  • Backs the Hugging Face Open LLM evaluations
  • MIT licensed CLI and Python library

OpenLLMetry (Traceloop)

LLM Eval & Observability
freeNew

✦ WhyA LangSmith alternative — OpenTelemetry-based instrumentation for LLM calls, vector DBs and frameworks.

  • OpenTelemetry-based instrumentation for LLM calls, vector DBs and frameworks
  • Ships traces to Datadog, Grafana, Honeycomb or any OTel backend
  • Apache-2.0

Laminar

LLM Eval & Observability
freeNew

✦ WhyA LangSmith alternative — Open-source observability and evals platform for AI agents (Apache-2.0).

  • Open-source observability and evals platform for AI agents (Apache-2.0)
  • Traces, session replay for browser agents and evaluations
  • Self-hostable

Langfuse

LLM Eval & Observability
freemium

✦ WhyA LangSmith alternative — Fully open-source and self-hostable for free via Docker Compose/Kubernetes, not just a hosted SaaS.

  • Fully open-source and self-hostable for free via Docker Compose/Kubernetes, not just a hosted SaaS
  • Hobby cloud plan is free with no credit card, 50k observability units/month
  • Combines tracing, prompt management, evals and datasets in one open platform

Braintrust

LLM Eval & Observability
freemium

✦ WhyA LangSmith alternative — Free Starter plan ships model credits and scored evals with no card required.

  • Free Starter plan ships model credits and scored evals with no card required
  • Unlimited users/projects/datasets/playgrounds even on the free tier
  • Built-in playground for side-by-side prompt/model comparison against production traces

Promptfoo

LLM Eval & Observability
freemium

✦ WhyA LangSmith alternative — Open-source CLI/library for evaluating prompts, models and RAG pipelines side by side, runs in CI.

  • Open-source CLI/library for evaluating prompts, models and RAG pipelines side by side, runs in CI
  • Automated red-teaming to surface prompt injection, jailbreak and data-leak vulnerabilities
  • 300,000+ user community with an enterprise tier for teams needing hosted guardrails

LangSmith head-to-head