Galileo is a comprehensive software solution for project management and collaboration. It offers features such as task tracking, team communication, and file...
Starting from $10 per user per month
β Update queued β the AI is re-ranking this list. The page will refresh shortly.
This page is already up to date.
DeepEval is an open-source Python framework for testing and evaluating LLM applications with built-in and custom metrics. It is designed for developers who want unit-test-style checks for RAG pipelines, agents, and conversational systems.
Galileo is a comprehensive software solution for project management and collaboration. It offers features such as task tracking, team communication, and file...
Starting from $10 per user per month
Weave is a comprehensive communication platform designed for small business owners to manage customer interactions.
Starting at $59/month
Humanloop is an enterprise platform for managing, testing, evaluating, and deploying prompts and AI applications. It targets product and engineering teams that...
Contact sales
Evalflow is an evaluation platform for testing and monitoring large language model applications. It helps AI teams run repeatable evaluations against datasets,...
LangSmith is an observability and evaluation platform for applications built with language models and agents. It provides tracing, dataset management, prompt testing,...
Free, paid from $39/seat/mo
BenchLLM is an open-source toolkit for evaluating and benchmarking large language model applications. It is aimed at developers and ML teams that...
Your feedback helps us improve the AI rankings.
β Thanks for your feedback!
Suggest a product and our AI will verify it's a real alternative to DeepEval before adding it to the list.