Skip to main content
Exécutez n'importe quel Skill dans Manus
en un clic
Goodeye-Labs
Profil créateur GitHub

Goodeye-Labs

Vue par dépôt de 10 skills collectés dans 2 dépôts GitHub.

skills collectés
10
dépôts
2
mis à jour
2026-05-08
explorateur de dépôts

Dépôts et skills représentatifs

create-evaluation
Analystes en assurance qualité des logiciels et testeurs

Scope what quality should be measured, convert it into one or more actionable binary evaluations, deploy those evaluations through Truesight MCP, and generate a companion skill that applies them correctly. Use when a user wants to create new evals, quality checks, guardrails, or pass/fail criteria for AI outputs.

2026-03-25
generate-synthetic-data
Scientifiques des données

Generate synthetic test data for LLM evaluations using dimension-based tuple expansion. Use when the user needs synthetic traces, test cases, eval datasets, or when create-evaluation needs synthetic fallback data.

2026-03-24
truesight-workflows
Scientifiques des données

Orchestrator for Truesight MCP skills. Use this when the user needs help choosing the right Truesight workflow or when intent is ambiguous across LLM evaluate, error analysis, review, templates, or evaluation creation.

2026-03-24
bootstrap-template-evaluation
Analystes en assurance qualité des logiciels et testeurs

Fastest route to a deployed live evaluation using a pre-built Truesight template. Use when the user wants a quick start without building judgment configs from scratch.

2026-03-10
build-review-interface
Développeurs web

Build a custom web interface for trace annotation and review. Use when users need a bespoke review surface for their workflow.

2026-03-10
error-analysis
Scientifiques des données

Systematically identify and categorize failure modes in evaluated traces using Truesight datasets and error-analysis tools. Use when quality issues are unclear, after major pipeline changes, or when incidents indicate drift.

2026-03-10
eval-audit
Scientifiques des données

Audit an existing evaluation workflow and produce severity-ranked findings with concrete next actions. Use when inheriting an eval setup, diagnosing quality regressions, or checking LLM evaluation process maturity.

2026-03-10
evaluate-trace
Analystes en assurance qualité des logiciels et testeurs

Evaluate one or more traces against an existing Truesight live evaluation. Use when a deployed live evaluation already exists and the user wants run outputs with optional handoff to review and promotion.

2026-03-10
Affichage des 8 principaux skills collectés sur 9 dans ce dépôt.
2 dépôts affichés sur 2
Tous les dépôts sont affichés