con un clic
NodeBenchAI
NodeBenchAI contiene 11 skills recopiladas de HomenShum, con cobertura ocupacional por repositorio y páginas de detalle dentro del sitio.
Skills en este repositorio
Turn "my app + an agent that demos" into "benchmarked, browser-verified, evidence-backed, prod-proven, and looping." Use when a (solo) founder wants to prove an AI agent works IN their real app — across all its UI surfaces — without cheating. Triggers: "set up proofloop", "proof-loop my app", "benchmark my agent's UI", "prove my agent works in prod", "run proofloop". The user's coding agent (Claude Code / Codex / Cursor) drives; the user steers by comment.
Use this skill to generate well-branded interfaces and assets for NodeBench AI, either for production or throwaway prototypes/mocks/etc. Contains essential design guidelines, colors, type, fonts, assets, and UI kit components for prototyping.
Migration runbook for @convex-dev/agent. Currently pinned to 0.2.10 because v0.6 requires a coordinated AI SDK v5 → v6 bump across many files. Use when bumping the agent or when CI shows the "args was removed in v0.6.0" error pattern.
Schema and authoring guide for dep-* skills — project-local migration runbooks for high-touch dependencies. Use when a dep major-bump breaks CI and you need to either apply an existing recipe or write a new one.
Pin runbook for @tiptap/pm. Pinned to 3.22.3 because 3.22.4 dropped the ./collab package export, which some transitive consumer (likely @blocknote or @tiptap/extensions) still relies on. Use when bumping tiptap or when build fails with "./collab is not exported".
Coordinated bump runbook for the vega + vega-lite + vega-embed ecosystem. The 5 → 6 jump fixed CVE-2025-59840 (XSS) and required updating the spec schema URL from v5 to v6. Use when bumping any of these three packages or when the vega XSS CVE alert fires.
Remediation runbook for the SheetJS xlsx package. Pinned to the SheetJS CDN release because the npm-published xlsx has 2 unfixed HIGH CVEs (ReDoS + Prototype Pollution) and SheetJS only ships fixes via their own CDN. Use when bumping xlsx, when CVE-2024-22363 or CVE-2023-30533 alerts fire, or when considering a migration to exceljs.
Run the repeatable operational loop on the NodeBench real-time chat pipeline or the report generator pipeline. Every pipeline change flows through: instrument → judge → persist → surface → measure → regress. Triggers: "pipeline", "operational loop", "run the standard", "measure the pipeline", "judge the traces", "chat pipeline change", "report generator change", or any edit to server/pipeline/* or convex/domains/product/diligence*.
Entity intelligence and founder clarity inside Claude Code. Search any company, run banker-grade deep diligence, get actionable remediation steps, and inject company truth into your session. Delegates implementation tasks to Codex when available.
Run NodeBench context sandbox diagnostics. Checks SQLite/FTS5 availability, sandbox table health, indexed content stats, and hook configuration. Trigger: /nodebench-mcp:sandbox-doctor
Show context sandbox savings for the current session. Displays bytes indexed vs returned, savings ratio, estimated tokens saved, and per-tool breakdown. Trigger: /nodebench-mcp:sandbox-stats