원클릭으로
solo-founder-agent-builder
solo-founder-agent-builder에는 HomenShum에서 수집한 skills 3개가 있으며, 저장소 수준 직업 범위와 사이트 내 skill 상세 페이지를 제공합니다.
이 저장소의 skills
Benchmark-driven development for AI agents — turns "I have an idea / prototype / half-built app and an agent that demos but does not hold up" into "an agent that completes real benchmark tasks IN the live app, browser-verified, without cheating." Runs the loop discover → benchmark → setup → build → adapter → verify → iterate under four non-negotiables (held-out · no answer-keys · in-app transfer · honest provenance), with the honest-lane clean-probe rule, a local-first memory substrate, and a Design Bridge for UI. Use when a (solo) founder wants to build or validate an AI agent for their app. Triggers: "build the agent layer for my app", "benchmark my agent", "prove my agent works in production", "which benchmark fits my agent", "make my agent pass SpreadsheetBench/BankerToolBench/SWE-bench in my app", "my agent demos but fails on real tasks", "eval my agent honestly". The user's coding agent (Claude Code, Codex, OpenClaw, Hermes, Trae) drives; the user steers by comment. This single skill IS the suite — it r
A reusable ProofLoop skill for making coding agents prove completion with external receipts, anti-gaming gates, held-out tasks, and live or stratified browser evidence before they claim done.
A reusable skill for taking one founder product goal through domain discovery, benchmark selection, local setup, app build, live workflow verification, and repair until the demo has receipts instead of vibes.