evaluation-metrics

Automatically applies when evaluating LLM performance. Ensures proper eval datasets, metrics computation, A/B testing, LLM-as-judge patterns, and experiment tracking.

সোর্স দেখুন machine-learning

maintainer

ricardoroche

আপডেট হয়েছে 11/18/2025

স্টার

ফর্ক

quick start

Installation and usage

Automatically applies when evaluating LLM performance. Ensures proper eval datasets, metrics computation, A/B testing, LLM-as-judge patterns, and experiment tracking.

ইনস্টলেশন

$ install --globalskills.sh

ব্যবহার

ইনস্টল করার পর, টার্মিনালে নিচের কমান্ড চালিয়ে আপনি এই স্কিল ব্যবহার করতে পারবেন:

skills use evaluation-metrics