home/categories/machine-learning/wshobson-agents-plugins-llm-application-dev-skills-llm-evaluation-skill-md
machine-learningdata-ai

llm-evaluation

Implement comprehensive evaluation strategies for LLM applications using automated metrics, human feedback, and benchmarking. Use when testing LLM performance, measuring AI application quality, or establishing evaluation frameworks.

wshobson
maintainer
wshobson
اپ ڈیٹ ہوا 3/7/2026
اسٹارز
33377
فورکس
3622
quick start

Installation and usage

Implement comprehensive evaluation strategies for LLM applications using automated metrics, human feedback, and benchmarking. Use when testing LLM performance, measuring AI application quality, or establishing evaluation frameworks.

انسٹالیشن
$ install --globalskills.sh
استعمال

انسٹال کرنے کے بعد، آپ یہ اسکل ٹرمینل میں درج ذیل کمانڈ چلا کر استعمال کر سکتے ہیں:

skills use llm-evaluation