Autonomous experimentation skill -- agent interviews the user, sets up a lab, then explores freely (think, test, reflect) until stopped or a target is hit. Works for any domain where you can measure or evaluate a result: code performance, prompt quality, build times, test coverage, document parsing accuracy, API latency, bundle size, or any custom metric. MANDATORY TRIGGERS: optimize, experiment, research, benchmark, improve metric, reduce latency, speed up, make faster, increase accuracy, A/B test, hypothesis, lab, autonomous optimization, iterative improvement, performance tuning. Use this skill whenever the user wants to systematically improve something measurable through repeated experimentation, even if they don't use the word 'experiment' or 'research'.
2026-04-13