Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

evaluating-rag-retrieval

النجوم٢
التفرعات٠
آخر تحديث١٧ يونيو ٢٠٢٦ في ٠٠:٥٥

Produces a complete evaluation report for a Retrieval-Augmented Generation (RAG) system from a golden Question-Answer set, separating retrieval-stage metrics (recall@k, MRR, nDCG@k) from generation-stage metrics (faithfulness, answer relevance, context utilization). Triggers whenever a RAG pipeline needs scoring, whenever the user reports the RAG "feels worse than direct LLM" but cannot localize where, whenever retrieval-vs-generation failure attribution is needed, or whenever a chunking or embedding-model change is being evaluated. Refuses to report a single aggregate score that conflates retrieval and generation failures.

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

مستكشف الملفات
8 ملفات
SKILL.md
readonly