Skip to main content
Run any Skill in Manus
with one click

threadlight-evals

Stars1
Forks4
UpdatedJuly 2, 2026 at 07:09

DISCOVER + GOVERN + IMPROVE evals leg for threadlight pilots on Microsoft Foundry. Runs and verifies offline batch quality evals, Foundry Continuous Evaluation on live threads, and champion-challenger comparison gates, then emits `specs/evals-manifest.json` for pillar 6 production-readiness scoring. USE FOR: continuous evals, offline eval gate, eval schedule, Foundry Continuous Evaluation, create_agent_evaluation, Application Insights eval results, eval threshold alert, eval run freshness, eval dataset shape, tool_calls tool_outputs, champion challenger, A/B eval gate, model swap gate, prompt swap gate, foundry-evals pipeline leg, continuous-evals pillar, EVAL-001..006, EVAL-101..105, evals-manifest. DO NOT USE FOR: token-level content filtering at the model edge — use the model guardrail / Azure AI Content Safety; adversarial scanning — use threadlight-redteam; agent-runtime action governance — use threadlight-govern; deep evaluator or dataset authoring — use foundry-evals.

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

File Explorer
18 files
SKILL.md
readonly