Skip to main content

llm-evaluation

Design and run an evaluation for an AI or agent feature using representative tasks, explicit success criteria, failure taxonomy, baselines, and reproducible evidence.

Jump to install

Source facts

Repository
Dadmin88/hermes-agency
Last source activity
August 16, 2026 at 06:25
Detected SKILL.md language
English
Stars
8
Forks
2

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.