Skip to main content
Run any Skill in Manus
with one click

run-aeon-benchmark

Stars21
Forks4
UpdatedJuly 19, 2026 at 18:27

Use when asked to run, execute, benchmark, evaluate, or score an LLM with AEON Bench. You deploy the AEON Bench **Pod** locally and drive the whole verified flow — pull or hash-verify a model → serve it → run the comprehensive benchmark (quality + speed + agentic + vision/audio/video) → submit the signed, attested result to the public leaderboard. An AI agent can do the ENTIRE job through the pod's MCP tools, no clicking. The mothership (aeon-bench.com) is a READ-ONLY leaderboard + submission endpoint; it never runs a job — all benchmarking happens on the pod.

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

SKILL.md
readonly