Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

longmemeval-iterate

النجوم٥٦
التفرعات٨
آخر تحديث١ يونيو ٢٠٢٦ في ٠٥:٢٨

Autonomous LongMemEval benchmark iteration toward ≥95% strict J-Score on full N=500. Use when the user asks to run, iterate, improve, or continue the LongMemEval campaign on branch `longmemeval-iter`. Loops baseline → cluster-analyze → propose fix → re-test → net-positive decision → commit → repeat. Terminates only when a clean-rerun confirmation also clears ≥475/500. SKIP for other benchmarks (LoCoMo, MuSiQue, NarrativeQA, etc.) — those have their own runners.

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

مستكشف الملفات
8 ملفات
SKILL.md
readonly