Skip to main content

tangle-network/agent-app

SkillsMP 已收集 tangle-network/agent-app 中的 7 个 Skill。打开任一 Skill 可查看来源和详情。

最近记录的来源活动
SkillsMP 收录数据更新
已收集 skills
7
GitHub 星标
0
GitHub Forks
0

这个仓库中的 skills

2 个职业分类 · 已分类 100%

已展示 7 / 7 个已收集 Skill。

职业分类
软件开发工程师
描述

Wire product measurement and improvement to Agent Eval campaigns, official optimizers, and explicit release decisions.

原文语言:英语

更新
职业分类
软件开发工程师
描述

Optimize one production agent surface with Agent Eval methods, separate data, bounded spend, and measured promotion.

原文语言:英语

更新
职业分类
软件开发工程师
描述

The user-facing controller for the Improve button. Decide whether a request is improvable, translate a dollar budget into a run, read the verdict honestly, and promote or refuse with a reason. Never promise a lift you cannot measure.

原文语言:英语

更新
职业分类
软件开发工程师
描述

Prove a measurement is sound BEFORE spending money optimizing against it. The gate that decides whether an Improve run is allowed to start, and whether its result is allowed to be believed. Refuse metrics whose noise exceeds the effect, that have no held-out…

原文语言:英语

更新
职业分类
其他计算机职业
描述

How every skill in the Improve family stays agentic and general instead of rotting into a brittle rulebook. A skill is a measured hypothesis — a few human-owned invariants plus a wide loop-owned judgment surface that improves from outcome data via its own…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Build a measurement that scores an agent's REAL deliverable — not a proxy — for a product you've never seen before. Use when scaffolding or repairing the eval an Improve loop optimizes against. Get this wrong and every downstream optimization perfects a…

原文语言:英语

更新
职业分类
软件开发工程师
描述

When a product has NO improvement infrastructure yet, build it for real — elicit the RIGHT target, ground the measurement in external truth, and construct a validated harness (often via a delegated agent-runtime build loop) BEFORE any optimization spend. The…

原文语言:英语

更新
已展示 7 / 7 个已收集 Skill。