Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

inspect-ai-python

النجوم١٧
التفرعات٢٣
آخر تحديث١٩ مايو ٢٠٢٦ في ٠٦:٥٠

Build LLM evaluations with Inspect AI (inspect-ai Python package by UK AISI). Use this skill whenever the user mentions Inspect AI, inspect-ai, LLM evaluation frameworks, eval tasks, eval datasets, solvers, scorers, agent evals, model grading, agentic benchmarks, SWE-bench evals, or any task involving evaluating language models systematically. Also trigger for questions about sandboxing model code execution, tool-use evals, multi-agent evaluation, eval log analysis, or running benchmarks like MMLU, HumanEval, GSM8K, HellaSwag, ARC with Inspect. Even if the user just says "write an eval" or "benchmark this model", consider this skill.

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

مستكشف الملفات
11 ملفات
SKILL.md
readonly