Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

coding-agent-robustness

النجوم١
التفرعات٠
آخر تحديث٢٠ أبريل ٢٠٢٦ في ١٨:٠٩

Systematic stress-testing and robustness measurement of coding agents (AI coding assistants, LLM-based code generators, or agentic coding systems). Use this skill whenever the user wants to benchmark a coding agent, evaluate its reliability, stress-test it on adversarial inputs, measure how it degrades under hard conditions, audit its security awareness, or produce a structured robustness report. Trigger on phrases like: "evaluate my coding agent", "how robust is X", "stress-test this agent", "benchmark coding assistant", "does it handle edge cases", "measure agent reliability", "what are the failure modes", "adversarial coding eval", "coding agent audit", "how does it perform under pressure", or any request to systematically assess an LLM's coding capability beyond basic correctness. Even if the user only describes a rough goal like "I want to know if my agent is production-ready", use this skill.

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

مستكشف الملفات
6 ملفات
SKILL.md
readonly