home/categories/productivity-tools/mikeyobrien-ralph-orchestrator-claude-skills-eval-skill-md
productivity-toolstools

eval

EvalKit is a conversational evaluation framework for AI agents that guides you through creating robust evaluations using the Strands Evals SDK. Through natural conversation, you can plan evaluations, generate test data, execute evaluations, and analyze results.

mikeyobrien
maintainer
mikeyobrien
اپ ڈیٹ ہوا 1/20/2026
اسٹارز
961
فورکس
115
quick start

Installation and usage

EvalKit is a conversational evaluation framework for AI agents that guides you through creating robust evaluations using the Strands Evals SDK. Through natural conversation, you can plan evaluations, generate test data, execute evaluations, and analyze results.

انسٹالیشن
$ install --globalskills.sh
استعمال

انسٹال کرنے کے بعد، آپ یہ اسکل ٹرمینل میں درج ذیل کمانڈ چلا کر استعمال کر سکتے ہیں:

skills use eval