Skip to main content
Manusで任意のスキルを実行
ワンクリックで

evaluation-reality-check

スター11
フォーク4
更新日2026年5月30日 07:45

Design evaluation strategies that match task type, deployment setting, time, groups, and decision risk. Use when an agent needs a judgment-heavy data science workflow for design trustworthy model evaluation, including evidence review, local artifact inspection, risk classification, stakeholder-ready decisions, reproducibility, governance, or agent-to-agent handoff. Trigger for Codex, Claude, Gemini, Copilot, Cursor, Windsurf, Gravity, LangGraph, CrewAI, AutoGen, or local agents when this exact workflow is needed.

インストール

Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。

ファイルエクスプローラー
26 ファイル
SKILL.md
readonly