Skip to main content
Splunk Agent Observability provides a comprehensive suite of pre-built evaluators designed to evaluate various aspects of AI system performance without requiring custom implementation. These evaluators span across eight categories including: Each evaluator addresses specific evaluation needs, from measuring factual correctness to detecting potential biases or tracking tool usage effectiveness. These evaluators apply to different node types (such as session, trace, or different span types), depending on the evaluator. Use the sortable, filterable table below to explore all available native evaluators and find the right measurements for your AI applications. Hover over an icon in the Modalities column to see what it represents.

Deprecated evaluators

The following evaluators have been deprecated. Use the evaluators documented in the previous table instead.