Skip to main content
Run any Skill in Manus
with one click

run-assert-eval

Stars197
Forks27
UpdatedJuly 6, 2026 at 22:50

Run an ASSERT evaluation from a plain-language behavior requirement. Use when the user wants to evaluate, test, or check an AI agent, LLM app, or model against requirements/policies (e.g. "evaluate my agent for budget violations", "test that the support bot never gives legal advice"). Generates or reuses an eval_config.yaml, runs the pipeline, and reports pass/violation rates with trace-cited failure examples.

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

SKILL.md
readonly