Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

eval-corpus-forge

النجوم١
التفرعات٠
آخر تحديث٧ يوليو ٢٠٢٦ في ٢٢:٠٥

Build, validate, and package reusable evaluation datasets for agent and LLM systems from prompts, traces, tool-call logs, metadata, and expected outcomes. Use this when the user asks to gather eval data, create benchmark datasets, normalize traces, define reusable ground truth, bootstrap eval suites, or produce retrieval, tool-invocation, response, or end-to-end evaluation artifacts.

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

مستكشف الملفات
29 ملفات
SKILL.md
readonly