Skip to main content

ai-eval-failure-analysis

Analyze Business Central AI Test Toolkit (AIT) evaluation results to explain WHY an eval run failed. Use when the user has an AIT run (a version/suite, or a saved aitTestLogEntries JSON export) and wants failure analysis, a confusion matrix, precision/recall/F1, specificity, MatchRate, accuracy, sibling-confusion, model comparison, or a breakdown of failing rows (misses, spurious matches, wrong accounts, crashes). Triggers on phrases like "analyze my eval", "why did the AIT fail", "compare these model runs", "precision/recall for this run", "AI test failure analysis", "analyze the latest <suite> run".

Jump to install

Source facts

Repository
microsoft/BCTech
Last source activity
June 11, 2026 at 11:27
Detected SKILL.md language
English
Stars
737
Forks
383

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.