Skip to main content

allenai/vla-evaluation-harness

SkillsMP 已收集 allenai/vla-evaluation-harness 中的 3 个 Skill。打开任一 Skill 可查看来源和详情。

最近记录的来源活动
SkillsMP 收录数据更新
已收集 skills
3
GitHub 星标
590
GitHub Forks
49

这个仓库中的 skills

2 个职业分类 · 已分类 100%

已展示 3 / 3 个已收集 Skill。

职业分类
数据科学家
描述

Run a VLA model evaluation against a simulation benchmark. Use this skill whenever the user wants to evaluate, benchmark, test, or run a model on a sim environment — even if they say it casually like 'try OpenVLA on LIBERO' or 'get me CALVIN scores'. Covers…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Add a new VLA model server to the evaluation harness. Use this skill whenever the user wants to integrate, create, or add a new model — e.g. 'add OpenVLA server', 'integrate RT-2', 'hook up my model', 'write a model server'. Also use when they ask how model…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Add a new simulation benchmark to the VLA evaluation harness. Use this skill whenever the user wants to integrate, create, or add a new benchmark or simulation environment — e.g. 'add ManiSkill3', 'integrate OmniGibson', 'hook up a new sim'. Also use when…

原文语言:英语

更新
已展示 3 / 3 个已收集 Skill。