home/categories/testing/dwmkerr-claude-toolkit-plugins-toolkit-skills-anthropic-evaluations-skill-md

testingtesting-security

anthropic-evaluations

Name: anthropic-evaluations
Author: dwmkerr

This skill should be used when the user asks to "create evals", "evaluate an agent", "build evaluation suite", or mentions agent testing, graders, or benchmarks. Also suggest when building coding agents, conversational agents, or research agents that need quality assurance.

عرض المصدر testing

maintainer

dwmkerr

آخر تحديث 1/19/2026

النجوم

التفرعات

quick start

Installation and usage

التثبيت

$ install --globalskills.sh

الاستخدام

بعد التثبيت، يمكنك استخدام هذه المهارة بتشغيل الأمر التالي في الطرفية:

skills use anthropic-evaluations