home/categories/llm-ai/tkhongsap-llama-index-rag-pipeline-claude-skills-evaluating-rag-skill-md
llm-aidata-ai

evaluating-rag

Evaluate RAG systems with hit rate, MRR, faithfulness metrics and compare retrieval strategies. Use when testing retrieval quality, generating evaluation datasets, comparing embeddings or retrievers, A/B testing, or measuring production RAG performance.

tkhongsap
maintainer
tkhongsap
Mis à jour 10/30/2025
Étoiles
1
Forks
0
quick start

Installation and usage

Evaluate RAG systems with hit rate, MRR, faithfulness metrics and compare retrieval strategies. Use when testing retrieval quality, generating evaluation datasets, comparing embeddings or retrievers, A/B testing, or measuring production RAG performance.

Installation
$ install --globalskills.sh
Utilisation

Après l'installation, vous pouvez utiliser ce skill en exécutant la commande suivante dans votre terminal :

skills use evaluating-rag