Skip to main content

mempot-defending-against-memory

Defend LLM agent memory systems against extraction attacks using optimized honeypot injection and sequential detection. Implements the MemPot framework: generates trap documents that lure attackers while staying invisible to legitimate users, then detects extraction attempts via Wald's Sequential Probability Ratio Test (SPRT). Trigger phrases: - "protect agent memory from extraction" - "add honeypots to my RAG memory" - "detect memory extraction attacks on my LLM agent" - "defend my vector database against adversarial queries" - "implement SPRT-based attack detection for my agent" - "harden my LLM agent's retrieval system"

الانتقال إلى التثبيت

معلومات المصدر

المستودع
ndpvt-web/arxiv-claude-skills
آخر نشاط في المصدر
١٣ فبراير ٢٠٢٦ في ٠٨:٣٧
لغة SKILL.md المكتشفة
الإنجليزية
النجوم
١٤
التفرعات
٣

خيارات التثبيت

يُحدَّد Prompt الذي يراجع المصدر أولًا بشكل افتراضي. يمكنك التبديل إلى أمر مباشر أو تنزيل نسخة محلية.

مراجعة ملفات المصدر

اقرأ SKILL.md وأي ملفات مرافقة يعرضها SkillsMP قبل أن تقرر التثبيت.