Skip to main content

meituan-longcat/WBench

SkillsMP は meituan-longcat/WBench から 5 件の skill を収集しています。skill を開くとソースと詳細を確認できます。

記録された最新のソース活動
SkillsMP カタログ更新
収集済み skills
5
GitHub スター
229
GitHub フォーク
11

このリポジトリの skills

収集済み skill 5 件中 5 件を表示しています。

職業分類
ソフトウェア品質保証アナリスト・テスター
説明

Run the WBench 22-metric evaluation pipeline on a model's videos. Use when the user asks to evaluate / score a model, run metrics, or produce a report (e.g. "evaluate kling3", "跑一下 hyworld1.5 的评测", "只算 video_quality"). Drives main.py (precompute → gpu → vlm →…

原文の言語: 英語

更新
職業分類
ソフトウェア開発者
説明

Generate WBench videos for a model. Use when the user asks to generate / produce videos for a registered model (e.g. "generate kling videos", "用 wan 生成全部 case", "跑 camera_preview 的 navi case"). Drives generate.py over data/cases and writes…

原文の言語: 英語

更新
職業分類
ソフトウェア開発者
説明

Package and submit a model's results to WBench (leaderboard). Use when the user asks to prepare a submission, build the meta.json/turns.json package, or upload videos to the WBench-examples HF dataset (e.g. "package kling3 for submission", "生成 turns.json",…

原文の言語: 英語

更新
職業分類
ソフトウェア品質保証アナリスト・テスター
説明

Automate Project Genie (labs.google) benchmark runs. Use when the user asks to run a case on Genie3 (e.g. "run case_XXX on genie3", "用genie3跑case_XXX"). Uses OS file-dialog upload because Genie3 enforces Private Network Access.

原文の言語: 英語

更新
職業分類
ソフトウェア品質保証アナリスト・テスター
説明

Automate happyoyster.cn/create benchmark runs. Use when the user asks to run a case on Happy Oyster (e.g. "run case_XXX on happy", "用happy跑case_XXX"). Fully JS-driven (fetch + DataTransfer for image upload, no pyautogui coordinates).

原文の言語: 英語

更新
収集済み skill 5 件中 5 件を表示しています。