Skip to main content

model-evaluation

Generates python code that evaluates SageMaker models. Supports two evaluation types: LLM-as-Judge and Custom Scorer. Use when the user says "evaluate my model", "run a benchmark", "test model performance", "how did my model perform", "compare models", or other similar requests.

Jump to install

Source facts

Repository
awslabs/agent-plugins
Last source activity
June 10, 2026 at 16:46
Detected SKILL.md language
English
Stars
881
Forks
152

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.