home/categories/llm-ai/tkhongsap-llama-index-rag-pipeline-claude-skills-evaluating-rag-skill-md
llm-aidata-ai

evaluating-rag

Evaluate RAG systems with hit rate, MRR, faithfulness metrics and compare retrieval strategies. Use when testing retrieval quality, generating evaluation datasets, comparing embeddings or retrievers, A/B testing, or measuring production RAG performance.

tkhongsap
maintainer
tkhongsap
更新日 10/30/2025
スター
1
フォーク
0
quick start

Installation and usage

Evaluate RAG systems with hit rate, MRR, faithfulness metrics and compare retrieval strategies. Use when testing retrieval quality, generating evaluation datasets, comparing embeddings or retrievers, A/B testing, or measuring production RAG performance.

インストール
$ install --globalskills.sh
使い方

インストール後、ターミナルで以下のコマンドを実行してこのスキルを使用できます:

skills use evaluating-rag