Skip to main content

ai-eval-failure-analysis

Analyze Business Central AI Test Toolkit (AIT) evaluation results to explain WHY an eval run failed. Use when the user has an AIT run (a version/suite, or a saved aitTestLogEntries JSON export) and wants failure analysis, a confusion matrix, precision/recall/F1, specificity, MatchRate, accuracy, sibling-confusion, model comparison, or a breakdown of failing rows (misses, spurious matches, wrong accounts, crashes). Triggers on phrases like "analyze my eval", "why did the AIT fail", "compare these model runs", "precision/recall for this run", "AI test failure analysis", "analyze the latest <suite> run".

الانتقال إلى التثبيت

معلومات المصدر

المستودع
microsoft/BCTech
آخر نشاط في المصدر
١١ يونيو ٢٠٢٦ في ١١:٢٧
لغة SKILL.md المكتشفة
الإنجليزية
النجوم
٧٣٧
التفرعات
٣٨٣

خيارات التثبيت

يُحدَّد Prompt الذي يراجع المصدر أولًا بشكل افتراضي. يمكنك التبديل إلى أمر مباشر أو تنزيل نسخة محلية.

مراجعة ملفات المصدر

اقرأ SKILL.md وأي ملفات مرافقة يعرضها SkillsMP قبل أن تقرر التثبيت.