원클릭으로
LiteBench
LiteBench에는 ahostbr에서 수집한 skills 4개가 있으며, 저장소 수준 직업 범위와 사이트 내 skill 상세 페이지를 제공합니다.
이 저장소의 skills
Orchestrate LiteBench model testing — scan LM Studio for models, run the agent harness against each, collect scores, swap models, and produce a leaderboard. Triggers on "test models", "run benchmark", "run harness", "test all models", "benchmark models", "start testing".
Autonomous agent and skill trainer — transforms the current Claude session into a self-improving loop inspired by Karpathy's autoresearch. Activates on: "train", "evolve agent", "improve agent", "run training loop", "agent trainer", "evolve this agent", "train this agent", "self-improving agent", "auto-improve", "training loop", "agent evolution", "autotrainer", "improve skill", "evolve skill", "optimize skill", "skill training", "improve my skills". Targets agents (.claude/agents/) OR skills (.claude/skills/) — auto-detects which. For agents: spawns A/B variants, evaluates output quality, mutates config. For skills: uses Claude Code's built-in eval system (run_eval.py) to measure trigger accuracy, then mutates the skill description. Use when you want an agent or skill to improve itself over N cycles without manual intervention.
Tune the agent harness system prompts to improve model tool-calling scores. Triggers on "tune harness", "improve scores", "train harness", "optimize prompts".
Download GGUF models from HuggingFace directly into LM Studio's model directory. Triggers on "download model", "get model", "grab model", "install model".