home/categories/academic/applied-artificial-intelligence-claude-code-toolkit-skills-llm-evaluation-skill-md
academicresearch
llm-evaluation
LLM evaluation and testing patterns including prompt testing, hallucination detection, benchmark creation, and quality metrics. Use when testing LLM applications, validating prompt quality, implementing systematic evaluation, or measuring LLM performance.
maintainer
applied-artificial-intelligence
업데이트됨 1/14/2026
스타
26
포크
7
quick start
Installation and usage
LLM evaluation and testing patterns including prompt testing, hallucination detection, benchmark creation, and quality metrics. Use when testing LLM applications, validating prompt quality, implementing systematic evaluation, or measuring LLM performance.
설치
$ install --globalskills.sh
사용법
설치 후 터미널에서 다음 명령을 실행하여 이 스킬을 사용할 수 있습니다:
skills use llm-evaluation