0
2.2.2
Australia Patch 4, Australia Patch 3, Australia Patch 2, Zurich Patch 11, Zurich Patch 10, Zurich Patch 9, Zurich Patch 8, Zurich Patch 5
Standalone Application
AI Control Tower – Evaluations provides insight into the runtime executions of AI agents by generating scores and detailed reasoning through an LLM-as-a-judge. It produces average quality and safety scores for each agent and tracks performance trends over time.
- Generates scores and detailed reasoning using an LLM-as-a-judge.
- Produces average quality and safety scores for each agent.
- Tracks performance trends over time.
New:
- Support sample rate configuration for each metric for external AI evaluations. Default sample rates are provided for all metrics automatically to ensure continuity
- Defect fixes
Plugin Dependencies:
- com.glide.hub.etl_consumer.kafka
- com.glide.tokenbased_auth
App Dependencies:
- sn_ai_governance (7.0.1)
- sn_ai_metric_ui (1.3.1)
- sn_telemetry_data (1.2.1)
Other app dependencies to support evaluations for third-party agents:
- sn_ai_disc (2.0.6)