Skip to main content
Manus에서 모든 스킬 실행
원클릭으로
wzh4464
GitHub 제작자 프로필

wzh4464

3개 GitHub 저장소에서 수집된 34개 skills를 저장소 단위로 보여줍니다.

수집된 skills
34
저장소
3
업데이트
2026-04-25
저장소 탐색

저장소와 대표 skills

diff-eval-local
소프트웨어 품질 보증 분석가·테스터

Evaluate agent-generated code against ground truth diff and handwritten file list. Prefer reading GT inputs directly from base_repo via experiment metadata. Trigger on "/diff-eval-local", "evaluate diff", "eval experiment".

2026-04-23
commit-push-pr-workflow
소프트웨어 개발자

Full dev workflow - worktree setup, commit, push, PR, and post-merge cleanup. Use when starting feature work, committing, opening PRs, or cleaning up after merge.

2026-03-30
hf-download
소프트웨어 개발자

Use when downloading models from Hugging Face, especially large models (LLM, diffusion, VLM) that fail with timeouts or connection drops. Trigger on "download model", "hf download", "pull model from Hugging Face", or when hf download crashes with httpx.ReadTimeout / RemoteProtocolError.

2026-03-30
diff-eval-func
소프트웨어 품질 보증 분석가·테스터

Evaluate agent-generated code changes against a human-approved PR with deterministic file coverage and function-level coverage analysis. Extracts function/symbol changes from diff hunk headers and compares at granular level. Use when evaluating agent code generation quality, benchmarking AI coding tools, or comparing generated patches against human-approved pull requests. Trigger on "/diff-eval-func", "evaluate diff with function coverage", "function-level diff eval".

2026-03-30
extract-paper-images
소프트웨어 개발자

从论文中提取图片,优先从arXiv源码包获取真正的论文图

2026-03-18
paper-analyze
데이터 과학자

深度分析单篇论文,生成详细笔记和评估,图文并茂

2026-03-18
paper-search
도서관 사무 보조원

在已整理的论文笔记中搜索相关内容

2026-03-18
start-my-day
시장조사 분석가·마케팅 전문가

每日研究工作流启动 - 生成论文推荐 + AI 行业动态笔记

2026-03-18
이 저장소에서 수집된 skills 18개 중 상위 8개를 표시합니다.
cheat-audit
소프트웨어 품질 보증 분석가·테스터

Audit an experiment run (agent+prompt+task) for suspicious output that may have come from the actual GT patch or upstream PR rather than independent implementation. Run whenever a new agent run is declared PASS, or whenever an eval report looks suspiciously high-coverage.

2026-04-25
run-benchmark
소프트웨어 개발자

对指定 task×type 组合,并行运行 4 个 agent(Claude/Cursor/Codex/OpenCode)的标准跑法。包含 base_repo 清洁验证、eval 保护、CONSTRAINT_DIRECTIVE 注入、实验目录创建、各 agent 启动命令。

2026-04-20
diff-eval-claude
소프트웨어 개발자

使用 Claude Code agent team 批量评测指定目录中的实验。并行评测,报告名称包含 "claude"。Trigger on "/diff-eval-claude", "claude eval", "用 claude 评测".

2026-04-19
diff-eval-codex
데이터 과학자

使用 Codex CLI (默认 OpenAI 订阅,无需额外配置) 批量评测实验目录。顺序或并行运行,报告名称包含 "codex"。Trigger on "/diff-eval-codex", "codex eval", "用 codex 评测".

2026-04-19
diff-eval-local
소프트웨어 품질 보증 분석가·테스터

Evaluate agent-generated code against ground truth diff and handwritten file list. Prefer reading GT inputs directly from base_repo via experiment metadata. Trigger on "/diff-eval-local", "evaluate diff", "eval experiment".

2026-04-19
diff-eval-opencode
소프트웨어 개발자

使用 OpenCode CLI 以指定模型批量评测实验目录。报告名称包含 "opencode-<model>"。Trigger on "/diff-eval-opencode", "opencode eval", "用 opencode 评测".

2026-04-19
run-k-benchmark
소프트웨어 개발자

Launch a multi-agent team to execute benchmark tasks in parallel. Team lead handles all setup (repo copy, hooks, metadata); agents directly code in their sessions without spawning subprocess claude invocations.

2026-04-19
inno-code-survey
소프트웨어 개발자

Acquires missing code repositories for the selected idea (Phase A) and conducts comprehensive code survey mapping academic concepts to implementations (Phase B). Outputs acquired_code_repos, updated_prepare_res, and model_survey for downstream use by inno-implementation-plan.

2026-04-19
이 저장소에서 수집된 skills 15개 중 상위 8개를 표시합니다.
저장소 3개 중 3개 표시
모든 저장소를 표시했습니다