Skip to main content
Exécutez n'importe quel Skill dans Manus
en un clic

coding-agent-robustness

Étoiles1
Forks0
Mis à jour20 avril 2026 à 18:09

Systematic stress-testing and robustness measurement of coding agents (AI coding assistants, LLM-based code generators, or agentic coding systems). Use this skill whenever the user wants to benchmark a coding agent, evaluate its reliability, stress-test it on adversarial inputs, measure how it degrades under hard conditions, audit its security awareness, or produce a structured robustness report. Trigger on phrases like: "evaluate my coding agent", "how robust is X", "stress-test this agent", "benchmark coding assistant", "does it handle edge cases", "measure agent reliability", "what are the failure modes", "adversarial coding eval", "coding agent audit", "how does it perform under pressure", or any request to systematically assess an LLM's coding capability beyond basic correctness. Even if the user only describes a rough goal like "I want to know if my agent is production-ready", use this skill.

Installation

Installer avec Codex ou Claude Copiez ce prompt, collez-le dans Codex, Claude ou un autre assistant, puis laissez-le vérifier la page du skill et l'installer pour vous.

Explorateur de fichiers
6 fichiers
SKILL.md
readonly