Skip to main content
Run any Skill in Manus
with one click

hugging-face-evaluation

Stars3
Forks0
UpdatedMarch 22, 2026 at 13:49

Add, import, and manage evaluation results in Hugging Face model cards. Supports extracting eval tables from README content, importing benchmark scores from Artificial Analysis API, and running custom model evaluations with vLLM/lighteval/inspect-ai on HF Jobs or locally. Works with the model-index metadata format for leaderboard and Papers with Code integration. Always activate when the user mentions HF model card evaluation, benchmark scores, model-index YAML, evaluation results, leaderboard submission, Artificial Analysis benchmarks, lighteval, inspect-ai, running evals on HF Jobs, or adding eval metrics to a model card — even if they don't say "skill".

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

SKILL.md
readonly