Galileo
An AI observability and evaluation platform that helps developers monitor, evaluate, debug, and improve the quality, reliability, and performance of large language model (LLM) applications and AI systems.
Explore more about Observability Tools
Related terms
The practice of monitoring, tracing, and analyzing AI systems to understand their behavior, performance, reliability, and operational health in production.
LangSmithA platform for tracing, evaluating, monitoring, and debugging LLM and agent applications.
LangfuseAn observability platform for tracing, monitoring, evaluating, and debugging LLM applications and agentic systems.
BraintrustAn AI evaluation and observability platform for testing, tracing, benchmarking, and improving the quality of LLM applications.
TruLensAn open-source software library for evaluating and tracking large language model applications and RAG pipelines.