Skip to main content
Manusで任意のスキルを実行
ワンクリックで

skill-eval-model-benchmark

スター0
フォーク1
更新日2026年3月10日 15:19

Run task evals across multiple Claude models and compare results side-by-side. Use when you want to understand how a skill performs across different models, identify model-specific gaps versus universal tile issues, or validate a skill before publishing it to the registry.

インストール

Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。

SKILL.md
readonly