Skip to main content
在 Manus 中运行任何 Skill
一键导入

model-benchmark

星标0
分支0
更新时间2026年7月10日 08:24

Create, inspect, resume, analyze, integrity-verify, and seal reusable Harbor/Codex model benchmark campaigns. Use when an agent in Codex CLI, Claude Code, Gajae Code, or Hermes Agent needs to compare exact model IDs on a repository-editing dataset; reuse an existing model-benchmark workspace; run model, container-auth, and oracle gates; prevent network or oracle leakage; aggregate pass@1, latency, token, and cost results; or audit preserved benchmark artifacts.

安装

用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。

文件资源管理器
9 个文件
SKILL.md
readonly