Use for complex software tasks that need interactive main-session planning followed by structured checklist execution.
Use to validate agent skill structure, frontmatter, path/name consistency, required sections, and broken references before skill changes are accepted.
Use to audit the skill ecosystem for capability coverage, overlapping responsibilities, fragmented guidance, and consolidation opportunities.
Use to evaluate agent skills with should-trigger and should-not-trigger cases, executable validation hooks, and pass/fail reports for review or CI.
Use when the user asks to create, improve, copy, port, evaluate, or brainstorm an agent skill. Probes the desired workflow, researches local and upstream skill examples, recommends patterns, then invokes plan-and-execute for implementation.
Trigger keywords: "grill me", "grill-me". Use only when the user explicitly requests grill-me style relentless, one-question-at-a-time interviewing.
Use to pressure-test implementation plans against codebase docs, refine language, capture ADR decisions, and emit machine-readable handoff JSON for downstream workflows.
Use to draft ADR records for strict-threshold decisions discovered during grilling and produce blocker-ready issue metadata.