Skip to main content

AI for agent observability

Trace, evaluate, and debug LLM apps.

#ToolCategoryPricingVisit
1Langfuse

Open-source LLM observability — tracing, evaluation, and prompt management

LLM FrameworksFreeVisit
2LangSmith

Debug, test, and monitor LLM applications built with LangChain or any framework

LLM FrameworksFreemiumVisit
3Helicone

One-line LLM observability proxy — log, monitor, and debug LLM API calls

LLM FrameworksFreemiumVisit
4Braintrust

Enterprise AI evaluation platform — experiments, datasets, and production monitoring

LLM FrameworksFreemiumVisit
5Promptfoo

Open-source LLM testing and red-teaming — test prompts, evaluate outputs, catch regressions

LLM FrameworksFreeVisit
6TruLens

LLM app evaluation and monitoring — measure quality and detect hallucinations in production

LLM FrameworksFreeVisit
7DeepEval

Open-source LLM evaluation framework — 14+ metrics for RAG, agents, and chatbots

LLM FrameworksFreeVisit
8Humanloop

LLM evaluation and prompt management — improve and deploy AI features in production

LLM FrameworksPaidVisit

See also