Skip to main content
Run any Skill in Manus
with one click
GitHub repository

dotfiles

dotfiles contains 53 collected skills from drewstone, with repository-level occupation coverage and site-owned skill detail pages.

skills collected
53
Stars
3
updated
2026-07-14
Forks
0
Occupation coverage
10 occupation categories · 100% classified
repository explorer

Skills in this repository

breakout
project-management-specialists

Raise the ceiling, don't climb toward it. When a target is nearly hit or a metric has plateaued, question the target itself: name the binding constraint, separate the physics floor from the assumed floor, design the regime change that makes 10x reachable, commit through the valley. Triggers: 'breakout', '10x this', 'raise the ceiling', 'why are we capped', 'think way bigger', 'this metric is a cage'.

2026-07-14
deploy-proof
software-developers

Prove that a merged change is actually live and production behavior is measured against the deployed artifact. Use after merge, before claiming deployed/live/validated, for Cloudflare Pages/Workers deploys, cache claims, perf claims, and release closeout.

2026-07-14
eval-harness-diagnose
software-quality-assurance-analysts-and-testers

Diagnose and improve eval harnesses built on @tangle-network/agent-eval or similar trace-first systems. Use when success/failure rates, scorecard deltas, judge outputs, promotion gates, benchmark rows, or improvement campaigns may be contaminated by harness bugs, evaluator drift, routing/auth failures, missing traces, or invalid baselines.

2026-07-14
evolve
software-developers

Goal-pursuit loop: measure, diagnose, experiment, verify, compare, and iterate against a measurable target. Triggers: evolve, optimize, make this better, push to target.

2026-07-14
finalize
software-developers

Split one messy experiment branch into N clean, independent branches — one logical change each, built from the merge-base, disjoint change-sets, each an isolated reviewable PR. Triggers: 'finalize', 'split this branch', 'atomize this', 'make reviewable PRs', 'decompose this mess'.

2026-07-14
governor
computer-occupations-all-other

Read repo/session state and choose the single next skill: exploit (evolve/polish), explore (pursue/meta-harness/breakout), bootstrap (eval-agent), diagnose (diagnose/eval-harness-diagnose), or step back (reflect). One dispatch, then exit.

2026-07-14
hub-sdk-adoption
software-developers

Adopt @tangle-network/hub-sdk in any product that needs to talk to the Tangle hub (OAuth connections, tools/search/describe/invoke, capability tokens, policies). Use whenever a product imports @tangle-network/agent-integrations, maintains a local hub-client.ts, or hand-rolls fetch() to /v1/hub/*. Forces kill-and-replace, not additive code.

2026-07-14
hypothesize
software-developers

The THINK phase before you spend compute. Research the solution space (prior art, competitors, the real ceiling), generate a diverse field of candidate mechanisms, rank them by expected value, sequence by information gain, and hand a ranked portfolio to /evolve or /pursue. Turns greedy poke-and-measure into science. Triggers: 'hypothesize', 'what should we try', 'research the space first', 'generate options', 'before we optimize', 'what are the bets'.

2026-07-14
meta-harness
software-developers

Automated architecture evolution after metric progress plateaus. Discover the code loop, create missing evals, run parallel proposers, benchmark variants, and keep the best patches.

2026-07-14
orchestrate
computer-occupations-all-other

Dynamic multi-agent workflow composition: decompose a goal no single skill covers, pick a dependency structure, layer coordination policies, wire skills as stages, compile to a Workflow script, run, synthesize. Triggers: 'orchestrate', 'compose a workflow', 'too big for one skill', 'fan this out', 'pipeline of agents'.

2026-07-14
pursue
software-developers

Design and build a generational improvement when the current approach is wrong or plateaued. Audit, choose a coherent architecture, implement, test, and hand off.

2026-07-14
research
software-developers

Merged into /evolve's structured-hypothesis mode. Use /evolve for hypothesis-driven experimentation, competitive landscape, and the bootstrap-CI promotion gate.

2026-07-14
ground-truth
software-developers

Before optimizing, debugging, or speeding up any LIVE system, stand up the FULL measured harness of the real production path FIRST — instrument every hop, benchmark the real (not local) path, build a reversible test loop, trace your own run, baseline + decompose — in one parallel fan-out. Skip it and you burn days optimizing a system you can't see, acting on a number true only in a narrower context than you present it.

2026-07-03
docs-slop-audit
software-developers

Audit technical docs for weak claims, AI slop, unclear product boundaries, false promises, and prose that misleads builders or buyers.

2026-07-01
product-design
web-and-digital-interface-designers

Design or revise visible product UI with reference-first judgment, real mode-specific controls, low-copy surfaces, and screenshot-based verification.

2026-07-01
simplify
software-developers

Capability-preserving simplification for active code work. Use when the user asks to simplify, modularize, remove god objects, reduce duplication, make code reusable, clean up abstractions, deep clean without losing capability, or asks "anything else to simplify?" after a feature/refactor/PR.

2026-07-01
arena-experiment
software-developers

Design + run a rigorous, equal-compute, executable-graded comparison of agent ARCHITECTURES (topologies, coordination policies, profiles) across a controlled difficulty axis — to find WHERE one approach beats another, not just whether it solves a task. Use before any "does smart multi-agent / topology X beat dumb loop / baseline Y" experiment. Reuse the substrate; never rebuild the harness.

2026-06-28
build-agent-app
software-developers

Adopt @tangle-network/agent-app — the shared application-shell framework for agent products — either greenfield (new product) or by migrating an existing app (from ANY stack). Starts with a discovery interview (product surface, agent surface, eval surface, features, sandbox-or-not, billing, integrations), then routes to the right module set + path. Covers the engine/shell/domain layering rule, per-module seams, sandbox AND non-sandbox (browser/edge copilot) wiring, the migration lift-loop, and anti-patterns. Use when standing up a new agent product, deciding what belongs in the app vs the framework, or porting an existing app onto agent-app.

2026-06-28
calibrate-before-measure
software-developers

Before running any eval, A/B, benchmark, or experiment, prove the metric can actually see what you claim to measure — and prove the task is hard enough to need the capability. Skip this and you measure the wrong thing for three experiments straight.

2026-06-28
dont-collapse-the-architecture
software-developers

When tempted to collapse an ambitious-but-unproven architecture (multi-agent topology, context-lifecycle management, a recursive loop system) into a dumb/old pattern because an early A/B looked marginal — don't. Marginal-early almost always means the regime that makes the architecture pay off wasn't active. Find that regime, build the missing competency, then judge.

2026-06-28
push-past-easy
software-developers

When you catch yourself doing the safe/easy version of a task or experiment, force the harder one that could actually fail. Timidity disguises itself as "let's start simple" and as the flattering result you didn't try to kill.

2026-06-28
agent-behavior-audit
computer-occupations-all-other

Audit whether an autonomous agent actually observes state, uses tools, improves from outcomes, and stays aligned to user intent using traces, logs, and artifacts.

2026-06-28
agent-eval
software-quality-assurance-analysts-and-testers

Extend @tangle-network/agent-eval internals: campaigns, scorecards, trace capture, backend integrity, held-out gates, analysts, auto-PR loops, RL bridge, and eval release checks.

2026-06-28
autopsy
software-quality-assurance-analysts-and-testers

Root-cause one null, surprising, or suspicious run result. Verify raw data, classify the cause, and decide whether to fix code, metric, design, or belief.

2026-06-28
bad
software-developers

Browser Agent Driver CLI operator for browser automation, UI/design audits, auth state, showcases, and benchmark runs. Triggers: "run bad", "browser agent", "design audit", "webbench", "automate this site".

2026-06-28
converge
software-developers

Drive failing CI to green by reading remote failures, reproducing locally when possible, fixing root causes, pushing, waiting, and repeating without shortcuts.

2026-06-28
critical-audit
software-quality-assurance-analysts-and-testers

Staff-engineer review of diffs, branches, docs, APIs, SDKs, and customer-facing surfaces. Findings first, severity ranked, file:line grounded, with a concrete fix plan.

2026-06-28
deep-clean
software-developers

Measured codebase cleanup: dead code, dependency cycles, weak types, duplicate logic, deprecated paths, test debt, and complexity, using real tools and before/after proof.

2026-06-28
diagnose
software-quality-assurance-analysts-and-testers

Analyze test, CI, benchmark, or eval failures; cluster by root cause; rank by impact and fix effort; produce concrete fix hypotheses.

2026-06-28
eval-agent
software-developers

Build LLM-as-judge components from real references: gather examples, generate rubrics, score outputs, return findings, and wire improvement loops.

2026-06-28
handoff
software-developers

Produce a session-to-session brief with current state, git/PR status, decisions, blockers, verification, and exact next actions for a fresh agent.

2026-06-28
harden
information-security-analysts

Security adversarial validation: derive invariants, attack surface, fuzz targets, credential risks, race conditions, and coverage gaps; extend existing tests to prove fixes.

2026-06-28
multi-pursue
software-developers

Run multiple independent /pursue-grade architecture tracks in parallel, each with its own brief, build, verification, and central synthesis.

2026-06-28
polish
software-developers

Apply a fixed quality rubric to existing work and fix every gap: correctness, design, robustness, tests, and public interface. Triggers: polish, tighten, production-grade.

2026-06-28
product-design-audit
web-and-digital-interface-designers

Product UI audit/redesign for apps, dashboards, workflows, marketing pages, fake components, visual slop, information architecture, copy hierarchy, and rendered proof.

2026-06-28
product-innovation-audit
project-management-specialists

First-principles product innovation audit for product bets, workflows, AI agents, marketplaces, developer tools, enterprise apps, differentiation, kill/ship decisions, and 10/10 marketability.

2026-06-28
reflect
software-developers

Analyze sessions or projects for patterns, misses, product signals, process improvements, automation opportunities, and skill effectiveness. Modes: session, project, portfolio.

2026-06-28
release-conductor
network-and-computer-systems-administrators

Run opaque or custom releases to production with a ledger, artifact decision, deploy proof, smoke checks, rollback path, ETA updates, and handoff.

2026-06-28
sandbox-sdk-integration
software-developers

Integrate @tangle-network/sandbox SDK without rebuilding stream durability, session replay, browser-safe clients, or idempotent dispatch already provided by the platform.

2026-06-28
semgrep
information-security-analysts

Run Semgrep static analysis for security findings. Supports important-only or full scans, Semgrep Pro when available, merged SARIF, triage, and remediation plans.

2026-06-28
Showing top 40 of 53 collected skills in this repository.