home/categories/llm-ai/tkhongsap-llama-index-rag-pipeline-claude-skills-evaluating-rag-skill-md
llm-aidata-ai

evaluating-rag

Evaluate RAG systems with hit rate, MRR, faithfulness metrics and compare retrieval strategies. Use when testing retrieval quality, generating evaluation datasets, comparing embeddings or retrievers, A/B testing, or measuring production RAG performance.

tkhongsap
maintainer
tkhongsap
Обновлено 10/30/2025
Звёзды
1
Форки
0
quick start

Installation and usage

Evaluate RAG systems with hit rate, MRR, faithfulness metrics and compare retrieval strategies. Use when testing retrieval quality, generating evaluation datasets, comparing embeddings or retrievers, A/B testing, or measuring production RAG performance.

Установка
$ install --globalskills.sh
Использование

После установки вы можете использовать этот skill, выполнив следующую команду в терминале:

skills use evaluating-rag