adversarially-robust-tool-orchestration
Use this skill when the user wants to test an agent's ability to resist manipulation of its internal reasoning, task plans, or session memory. It targets scenarios where an attacker might inject fictitious plans into the conversation history, 'poison' the stored context with malicious follow-ups, or use logical bridges to redirect the agent from a legitimate goal to a harmful one. Trigger it for requests like 'test if the agent follows a fake plan,' 'see if it gets hijacked by corrupted memory,' 'make the task instructions change midway using a trap,' and 'check robustness against logical bridges or plan injections.' It applies whenever the orchestrator must verify that its current execution track still aligns with the original user intent despite adversarial context manipulation.
معلومات المصدر
- المستودع
- Dingxingdi/paper_fast_search_backup
- آخر نشاط في المصدر
- ١٠ أبريل ٢٠٢٦ في ٠١:٢٧
- لغة SKILL.md المكتشفة
- الإنجليزية
- النجوم
- ٠
- التفرعات
- ٠
خيارات التثبيت
يُحدَّد Prompt الذي يراجع المصدر أولًا بشكل افتراضي. يمكنك التبديل إلى أمر مباشر أو تنزيل نسخة محلية.
مراجعة ملفات المصدر
اقرأ SKILL.md وأي ملفات مرافقة يعرضها SkillsMP قبل أن تقرر التثبيت.