AI for agent observability
Trace, evaluate, and debug LLM apps.
| # | Tool | Category | Pricing | Visit |
|---|---|---|---|---|
| 1 | Langfuse Open-source LLM observability — tracing, evaluation, and prompt management | LLM Frameworks | Free | Visit |
| 2 | LangSmith Debug, test, and monitor LLM applications built with LangChain or any framework | LLM Frameworks | Freemium | Visit |
| 3 | Helicone One-line LLM observability proxy — log, monitor, and debug LLM API calls | LLM Frameworks | Freemium | Visit |
| 4 | Braintrust Enterprise AI evaluation platform — experiments, datasets, and production monitoring | LLM Frameworks | Freemium | Visit |
| 5 | Promptfoo Open-source LLM testing and red-teaming — test prompts, evaluate outputs, catch regressions | LLM Frameworks | Free | Visit |
| 6 | TruLens LLM app evaluation and monitoring — measure quality and detect hallucinations in production | LLM Frameworks | Free | Visit |
| 7 | DeepEval Open-source LLM evaluation framework — 14+ metrics for RAG, agents, and chatbots | LLM Frameworks | Free | Visit |
| 8 | Humanloop LLM evaluation and prompt management — improve and deploy AI features in production | LLM Frameworks | Paid | Visit |