Skip to main content
Exécutez n'importe quel Skill dans Manus
en un clic

evaluation-reality-check

Étoiles11
Forks4
Mis à jour30 mai 2026 à 07:45

Design evaluation strategies that match task type, deployment setting, time, groups, and decision risk. Use when an agent needs a judgment-heavy data science workflow for design trustworthy model evaluation, including evidence review, local artifact inspection, risk classification, stakeholder-ready decisions, reproducibility, governance, or agent-to-agent handoff. Trigger for Codex, Claude, Gemini, Copilot, Cursor, Windsurf, Gravity, LangGraph, CrewAI, AutoGen, or local agents when this exact workflow is needed.

Installation

Installer avec Codex ou Claude Copiez ce prompt, collez-le dans Codex, Claude ou un autre assistant, puis laissez-le vérifier la page du skill et l'installer pour vous.

Explorateur de fichiers
26 fichiers
SKILL.md
readonly