Skip to main content

capopt-capability-evaluation-benchmarking-agent

Capability/optimization role: **Capability evaluation & benchmarking agent** (AI agent) — measures capability, robustness, and regression across methods and model tiers and finds the efficient frontier. Part of the layer that decides *how* robot and machine capabilities are built — across model tiers (LLM, SLM, tiny LM, deterministic) and many training methods (imitation, model-based/offline RL, RLHF/RLAIF, sim-to-real, distillation, classical control, formal methods). Use this skill when choosing or building how a capability is trained, optimized, or run on-device, even if the user only describes the underlying need. Works under a evaluation lead.

الانتقال إلى التثبيت

معلومات المصدر

المستودع
TuringWorks/civstack
آخر نشاط في المصدر
١٦ يونيو ٢٠٢٦ في ٠٧:٠٣
لغة SKILL.md المكتشفة
الإنجليزية
النجوم
١
التفرعات
٠

خيارات التثبيت

يُحدَّد Prompt الذي يراجع المصدر أولًا بشكل افتراضي. يمكنك التبديل إلى أمر مباشر أو تنزيل نسخة محلية.

مراجعة ملفات المصدر

اقرأ SKILL.md وأي ملفات مرافقة يعرضها SkillsMP قبل أن تقرر التثبيت.