home/categories/productivity-tools/mikeyobrien-ralph-orchestrator-claude-skills-eval-skill-md
productivity-toolstools

eval

EvalKit is a conversational evaluation framework for AI agents that guides you through creating robust evaluations using the Strands Evals SDK. Through natural conversation, you can plan evaluations, generate test data, execute evaluations, and analyze results.

mikeyobrien
maintainer
mikeyobrien
更新于 1/20/2026
星标
961
分支
115
quick start

Installation and usage

EvalKit is a conversational evaluation framework for AI agents that guides you through creating robust evaluations using the Strands Evals SDK. Through natural conversation, you can plan evaluations, generate test data, execute evaluations, and analyze results.

安装
$ install --globalskills.sh
使用

安装后,您可以通过在终端运行以下命令来使用此技能:

skills use eval