Skip to main content
tangle-network
Profil créateur GitHub

tangle-network

Vue par dépôt de 44 skills collectés dans 9 dépôts GitHub.

skills collectés
44
dépôts
9
mis à jour
22 août 2026
Les 8 principaux dépôts sont affichés ici ; la liste complète continue ci-dessous.
explorateur de dépôts

Dépôts et skills représentatifs

Affichage de 8 skills collectés sur 17.
eval-campaign
Développeurs de logiciels

Wire product measurement and improvement to Agent Eval campaigns, official optimizers, and explicit release decisions.

27 juil. 2026
surface-evolution
Développeurs de logiciels

Optimize one production agent surface with Agent Eval methods, separate data, bounded spend, and measured promotion.

27 juil. 2026
improve-conductor
Développeurs de logiciels

The user-facing controller for the Improve button. Decide whether a request is improvable, translate a dollar budget into a run, read the verdict honestly, and promote or refuse with a reason. Never promise a lift you cannot measure.

8 juil. 2026
measurement-validation
Développeurs de logiciels

Prove a measurement is sound BEFORE spending money optimizing against it. The gate that decides whether an Improve run is allowed to start, and whether its result is allowed to be believed. Refuse metrics whose noise exceeds the effect, that have no held-out…

8 juil. 2026
skill-evolution
Autres occupations informatiques

How every skill in the Improve family stays agentic and general instead of rotting into a brittle rulebook. A skill is a measured hypothesis — a few human-owned invariants plus a wide loop-owned judgment surface that improves from outcome data via its own…

8 juil. 2026
eval-architect
Développeurs de logiciels

Build a measurement that scores an agent's REAL deliverable — not a proxy — for a product you've never seen before. Use when scaffolding or repairing the eval an Improve loop optimizes against. Get this wrong and every downstream optimization perfects a…

6 juin 2026
eval-bootstrap
Développeurs de logiciels

When a product has NO improvement infrastructure yet, build it for real — elicit the RIGHT target, ground the measurement in external truth, and construct a validated harness (often via a delegated agent-runtime build loop) BEFORE any optimization spend. The…

6 juin 2026
9 dépôts affichés sur 9
Tous les dépôts sont affichés