Skip to main content
在 Manus 中运行任何 Skill
一键导入
wzh4464
GitHub 创作者资料

wzh4464

按仓库查看 3 个 GitHub 仓库中的 34 个已收集 skills。

已收集 skills
34
仓库
3
更新
2026-04-25
仓库浏览

仓库与代表性 skills

diff-eval-local
软件质量保证分析师与测试员

Evaluate agent-generated code against ground truth diff and handwritten file list. Prefer reading GT inputs directly from base_repo via experiment metadata. Trigger on "/diff-eval-local", "evaluate diff", "eval experiment".

2026-04-23
commit-push-pr-workflow
软件开发工程师

Full dev workflow - worktree setup, commit, push, PR, and post-merge cleanup. Use when starting feature work, committing, opening PRs, or cleaning up after merge.

2026-03-30
hf-download
软件开发工程师

Use when downloading models from Hugging Face, especially large models (LLM, diffusion, VLM) that fail with timeouts or connection drops. Trigger on "download model", "hf download", "pull model from Hugging Face", or when hf download crashes with httpx.ReadTimeout / RemoteProtocolError.

2026-03-30
diff-eval-func
软件质量保证分析师与测试员

Evaluate agent-generated code changes against a human-approved PR with deterministic file coverage and function-level coverage analysis. Extracts function/symbol changes from diff hunk headers and compares at granular level. Use when evaluating agent code generation quality, benchmarking AI coding tools, or comparing generated patches against human-approved pull requests. Trigger on "/diff-eval-func", "evaluate diff with function coverage", "function-level diff eval".

2026-03-30
extract-paper-images
软件开发工程师

从论文中提取图片,优先从arXiv源码包获取真正的论文图

2026-03-18
paper-analyze
数据科学家

深度分析单篇论文,生成详细笔记和评估,图文并茂

2026-03-18
paper-search
图书馆文员助理

在已整理的论文笔记中搜索相关内容

2026-03-18
start-my-day
市场调研分析师与营销专员

每日研究工作流启动 - 生成论文推荐 + AI 行业动态笔记

2026-03-18
当前展示该仓库 Top 8 / 18 个已收集 skills。
cheat-audit
软件质量保证分析师与测试员

Audit an experiment run (agent+prompt+task) for suspicious output that may have come from the actual GT patch or upstream PR rather than independent implementation. Run whenever a new agent run is declared PASS, or whenever an eval report looks suspiciously high-coverage.

2026-04-25
run-benchmark
软件开发工程师

对指定 task×type 组合,并行运行 4 个 agent(Claude/Cursor/Codex/OpenCode)的标准跑法。包含 base_repo 清洁验证、eval 保护、CONSTRAINT_DIRECTIVE 注入、实验目录创建、各 agent 启动命令。

2026-04-20
diff-eval-claude
软件开发工程师

使用 Claude Code agent team 批量评测指定目录中的实验。并行评测,报告名称包含 "claude"。Trigger on "/diff-eval-claude", "claude eval", "用 claude 评测".

2026-04-19
diff-eval-codex
数据科学家

使用 Codex CLI (默认 OpenAI 订阅,无需额外配置) 批量评测实验目录。顺序或并行运行,报告名称包含 "codex"。Trigger on "/diff-eval-codex", "codex eval", "用 codex 评测".

2026-04-19
diff-eval-local
软件质量保证分析师与测试员

Evaluate agent-generated code against ground truth diff and handwritten file list. Prefer reading GT inputs directly from base_repo via experiment metadata. Trigger on "/diff-eval-local", "evaluate diff", "eval experiment".

2026-04-19
diff-eval-opencode
软件开发工程师

使用 OpenCode CLI 以指定模型批量评测实验目录。报告名称包含 "opencode-<model>"。Trigger on "/diff-eval-opencode", "opencode eval", "用 opencode 评测".

2026-04-19
run-k-benchmark
软件开发工程师

Launch a multi-agent team to execute benchmark tasks in parallel. Team lead handles all setup (repo copy, hooks, metadata); agents directly code in their sessions without spawning subprocess claude invocations.

2026-04-19
inno-code-survey
软件开发工程师

Acquires missing code repositories for the selected idea (Phase A) and conducts comprehensive code survey mapping academic concepts to implementations (Phase B). Outputs acquired_code_repos, updated_prepare_res, and model_survey for downstream use by inno-implementation-plan.

2026-04-19
当前展示该仓库 Top 8 / 15 个已收集 skills。
已展示 3 / 3 个仓库
已展示全部仓库