Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

compare-skill-model-performance

النجوم١٣
التفرعات٢
آخر تحديث١٩ يونيو ٢٠٢٦ في ٠٨:٢١

Run task evals across multiple Claude models, compare results side-by-side, and optimise. Use when you want to benchmark a skill across models, compare haiku vs sonnet vs opus performance, run multi-model comparison or benchmark reports, identify model-specific gaps versus universal plugin gaps, evaluate whether a skill works for all model tiers, or validate a skill before publishing it to the registry.

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

SKILL.md
readonly