Skip to main content

ai-eval-failure-analysis

Analyze Business Central AI Test Toolkit (AIT) evaluation results to explain WHY an eval run failed. Use when the user has an AIT run (a version/suite, or a saved aitTestLogEntries JSON export) and wants failure analysis, a confusion matrix, precision/recall/F1, specificity, MatchRate, accuracy, sibling-confusion, model comparison, or a breakdown of failing rows (misses, spurious matches, wrong accounts, crashes). Triggers on phrases like "analyze my eval", "why did the AIT fail", "compare these model runs", "precision/recall for this run", "AI test failure analysis", "analyze the latest <suite> run".

インストールへ移動

ソース情報

リポジトリ
microsoft/BCTech
ソースの最終更新活動
2026年6月11日 11:27
検出された SKILL.md の言語
英語
スター
737
フォーク
383

インストール方法

デフォルトでは、最初にソースを確認する Prompt が選択されています。直接コマンドに切り替えるか、ローカルコピーをダウンロードすることもできます。

ソースファイルを確認

インストールを決める前に、SKILL.md と SkillsMP に表示されている付属ファイルをお読みください。