Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

eval-memory

النجوم٠
التفرعات٠
آخر تحديث١٨ يونيو ٢٠٢٦ في ٠٨:٢٩

Autonomously check the health of the personal-knowledge memory system AND validate whether the autolearn improvements are actually working (not just active). Runs a blind LLM-as-judge efficacy eval against a frozen probe set, scores re-rank vs raw retrieval, tracks the trend in a ledger, and judges how well the dashboard surfaces that progress. Invoke when the user wants a memory-system health + efficacy report, asks "is autolearn working / improving", or on a schedule (cron-ready).

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

SKILL.md
readonly