| name | tokenspeed-benchmark |
| description | TokenSpeed benchmark planning skill for LLM inference experiments. |
| license | MIT |
tokenspeed-benchmark
Define model, hardware, prompts, batch/concurrency, latency, throughput, correctness checks, and baseline comparison before installing engines.
OctoAgent usage
- Confirm the user goal and constraints.
- Load any matching plugin command with
get_plugin_command.
- Run
integrated_workflow_run when a workflow ID is available.
- Produce artifacts and review them against the quality gates before side effects.