eval

Name: eval
Author: mikeyobrien

EvalKit is a conversational evaluation framework for AI agents that guides you through creating robust evaluations using the Strands Evals SDK. Through natural conversation, you can plan evaluations, generate test data, execute evaluations, and analyze results.

檢視原始碼 productivity-tools

maintainer

mikeyobrien

更新於 1/20/2026

星標

961

分支

115

quick start

Installation and usage

安裝

$ install --globalskills.sh

使用

安裝後，您可以透過在終端機執行以下指令來使用此技能：

skills use eval