Skip to main content
Run any Skill in Manus
with one click

eval-memory

Stars0
Forks0
UpdatedJune 18, 2026 at 08:29

Autonomously check the health of the personal-knowledge memory system AND validate whether the autolearn improvements are actually working (not just active). Runs a blind LLM-as-judge efficacy eval against a frozen probe set, scores re-rank vs raw retrieval, tracks the trend in a ledger, and judges how well the dashboard surfaces that progress. Invoke when the user wants a memory-system health + efficacy report, asks "is autolearn working / improving", or on a schedule (cron-ready).

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

SKILL.md
readonly