Skip to main content
Manusで任意のスキルを実行
ワンクリックで

model-benchmark

スター0
フォーク0
更新日2026年7月10日 08:24

Create, inspect, resume, analyze, integrity-verify, and seal reusable Harbor/Codex model benchmark campaigns. Use when an agent in Codex CLI, Claude Code, Gajae Code, or Hermes Agent needs to compare exact model IDs on a repository-editing dataset; reuse an existing model-benchmark workspace; run model, container-auth, and oracle gates; prevent network or oracle leakage; aggregate pass@1, latency, token, and cost results; or audit preserved benchmark artifacts.

インストール

Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。

ファイルエクスプローラー
9 ファイル
SKILL.md
readonly