Create a new Harbor task for evaluating agents. Use when the user wants to scaffold, build, or design a new task, benchmark problem, or eval. Guides through instruction writing, environment setup, verifier design (pytest vs Reward Kit vs custom), and solution…
multimodal-art-projection/TACO
SkillsMP has collected 3 skills from multimodal-art-projection/TACO. Open a skill to review its source and details.
- Latest recorded source activity
- SkillsMP catalog refreshed
- skills collected
- 3
- GitHub stars
- 47
- GitHub forks
- 5
Skills in this repository
2 occupation categories · 100% classified
Showing 3 of 3 collected skills.
skill
occupation
description
updated
occupation
Software Quality Assurance Analysts & Testers
description
updated
occupation
Software Developers
description
Publish a Harbor task or dataset to the registry. Use when the user wants to upload, publish, or share tasks or datasets/benchmarks on the Harbor registry.
updated
occupation
Software Quality Assurance Analysts & Testers
description
Write Harbor task verifiers using Reward Kit. Use when creating or editing a task's tests/ directory, adding grading criteria, setting up LLM/agent judges, or designing verifiers that produce a reward score.
updated
Showing 3 of 3 collected skills.