LangSmith alternatives
6 evaluation tools in the same lane as LangSmith, ranked by editorial score.
LA
Langfuse
Evaluation · Model-agnostic
7.3
Open-source LLM observability, prompt management, and evaluation in one platform.
Freemium· Free self-host & Hobby tier; Core $29/mo, Pro $199/mo, Enterprise $2,499/mollm-observabilityprompt-management
LA
LangWatch
Evaluation · Model-agnostic; supports OpenAI, Anthropic, AWS Bedrock, Azure OpenAI, Vertex AI, and any OpenTelemetry-instrumented LLM
Simulation-based testing, evaluation, and observability for LLM agents
Freemium· Developer: €0 · Growth: €29/ core-seat / month · Enterprise: CustomLLM agent regression testing in CIRAG answer-quality evaluation
HE
Helicone
Evaluation · Platform (any LLM)
8.3
Open-source LLM observability — one-line proxy install.
Freemium· Free 100k req/mo; Pro from $25/moobservabilitycost tracking
OE
OpenAI Evals
Evaluation · OpenAI GPT models (extensible)
8.1
OpenAI's open-source framework for benchmarking LLMs against a shared registry of evaluations.
Free· Free (MIT); you pay OpenAI API costs for eval runsllm-benchmarkingregression-testing
TR
TruLens
Evaluation · Multi-model (LLM-as-judge)
8.1
Open-source evaluation and tracing framework for LLM apps and agents, built on OpenTelemetry.
Free· Free, open source (Apache-licensed Python package)llm-evaluationrag-evaluation
AG
Agenta
Evaluation · Multi-model
6.9
Open-source LLMOps platform for prompt engineering, evaluation, and observability in one workspace.
Freemium· Hobby: $0 forever · Pro: $29 /month · Business: $299 /month · Enterprise: Custom · Open source: Free foreverprompt-engineeringllm-evaluation