Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

eval-design

النجوم٢
التفرعات٠
آخر تحديث٢٦ يونيو ٢٠٢٦ في ٠٤:٤٤

Designs LLM evaluation frameworks with metric selection by task type, test set sizing, pass/fail thresholds, and drift triggers. Use when setting up evals for an LLM feature, before first production deployment, or when asked how to measure AI quality.

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

SKILL.md
readonly