Ejecuta cualquier Skill en Manus
con un clic

Ejecuta cualquier Skill en Manus con un clic

Repositorio de GitHub

agentic-usability

agentic-usability contiene 10 skills recopiladas de PSPDFKit-labs, con cobertura ocupacional por repositorio y páginas de detalle dentro del sitio.

Perfil de PSPDFKit-labs Ver en GitHub

skills recopiladas

Stars

actualizado

2026-05-14

Forks

Cobertura ocupacional

Analistas de garantía de calidad de software y probadores Desarrolladores de software Administradores de redes y sistemas informáticos

3 categorías ocupacionales · 100% clasificado

explorador de repositorios

Skills en este repositorio

creador/repositorio/skill

skill

ocupación

descripción

actualizado

init

Desarrolladores de software

Initialize a new agentic-usability benchmark pipeline project. Use when setting up a new SDK benchmark, creating a config.json, or starting a new evaluation project.

2026-05-14

sandbox

Administradores de redes y sistemas informáticos

Launch an interactive shell inside a microsandbox for debugging. Supports bare mode, executor setup, or judge setup with optional test case scaffolding.

2026-05-14

eval

Analistas de garantía de calidad de software y probadores

Run the full evaluation pipeline (execute, judge, report) for an SDK usability benchmark. Use when running a complete benchmark end-to-end, resuming an interrupted pipeline, or checking pipeline status.

2026-04-27

execute

Analistas de garantía de calidad de software y probadores

Execute benchmark test cases in sandboxed environments with AI agents. Spins up microsandbox containers for each test case and extracts solutions.

2026-04-27

export

Desarrolladores de software

Export a benchmark pipeline as a zip file for sharing or archiving. Excludes cache and large snapshots.

2026-04-27

generate

Analistas de garantía de calidad de software y probadores

Generate SDK usability test cases by exploring source code. Use when creating benchmark test suites, generating test cases for an SDK, or when the user wants to create evaluation scenarios.

2026-04-27

insights

Desarrolladores de software

Analyze benchmark results and identify SDK improvement areas. Use when reviewing evaluation results, finding failure patterns, identifying documentation gaps, or understanding API design issues.

2026-04-27

inspect

Desarrolladores de software

Open the web UI to visually inspect, edit, and run the benchmark pipeline. Use when the user wants a visual interface for their pipeline.

2026-04-27

judge

Analistas de garantía de calidad de software y probadores

Have an LLM judge compare reference and generated solutions, scoring on API discovery, correctness, completeness, and functional correctness.

2026-04-27

report

Analistas de garantía de calidad de software y probadores

Display a terminal scorecard of benchmark results showing pass rates, scores by difficulty, and per-test breakdowns. Use when the user asks about benchmark results, scores, or wants to see how their SDK performed.

2026-04-27