Skip to main content

adversarially-robust-tool-orchestration

Use this skill when the user wants to test an agent's ability to resist manipulation of its internal reasoning, task plans, or session memory. It targets scenarios where an attacker might inject fictitious plans into the conversation history, 'poison' the stored context with malicious follow-ups, or use logical bridges to redirect the agent from a legitimate goal to a harmful one. Trigger it for requests like 'test if the agent follows a fake plan,' 'see if it gets hijacked by corrupted memory,' 'make the task instructions change midway using a trap,' and 'check robustness against logical bridges or plan injections.' It applies whenever the orchestrator must verify that its current execution track still aligns with the original user intent despite adversarial context manipulation.

الانتقال إلى التثبيت

معلومات المصدر

المستودع
Dingxingdi/paper_fast_search_backup
آخر نشاط في المصدر
١٠ أبريل ٢٠٢٦ في ٠١:٢٧
لغة SKILL.md المكتشفة
الإنجليزية
النجوم
٠
التفرعات
٠

خيارات التثبيت

يُحدَّد Prompt الذي يراجع المصدر أولًا بشكل افتراضي. يمكنك التبديل إلى أمر مباشر أو تنزيل نسخة محلية.

مراجعة ملفات المصدر

اقرأ SKILL.md وأي ملفات مرافقة يعرضها SkillsMP قبل أن تقرر التثبيت.