Skip to main content

这个仓库中的 skills

autohandai/community-skills - 第 3 页

SkillsMP 已收集 autohandai/community-skills 中的 1,040 个 Skill。打开任一 Skill 可查看来源和详情。

autohandai/community-skills

已展示 40 / 1,040 个已收集 Skill。

职业分类
软件质量保证分析师与测试员
描述

Refresh golden values from a GitHub Actions workflow run (failing-only or all jobs), score the change with average normalized relative differences, and produce a PR-ready summary. Use when the user asks to update goldens for a CI run, refresh golden values…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Query and browse evaluation results stored in MLflow. Use when the user wants to look up runs by invocation ID, compare metrics across models, fetch artifacts (configs, logs, results), or set up the MLflow MCP server. ALWAYS triggers on mentions of MLflow,…

原文语言:英语

更新
职业分类
网络与计算机系统管理员
描述

Run commands inside a remote Docker container via the file-based command relay (tools/debugger). Use when the user says "run in Docker", "run on GPU", "debug remotely", "run test in container", "check nvidia-smi", "run pytest in Docker", or needs to execute…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Serve a quantized or unquantized LLM checkpoint as an OpenAI-compatible API endpoint using vLLM, SGLang, or TRT-LLM. Use when user says "deploy model", "serve model", "start vLLM server", "launch SGLang", "TRT-LLM deploy", "AutoDeploy", "benchmark…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Evaluates accuracy of quantized or unquantized LLMs using NeMo Evaluator Launcher (NEL). Triggers on "evaluate model", "benchmark accuracy", "run MMLU", "evaluate quantized model", "accuracy drop", "run nel". Handles deployment, config generation, and…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Run, monitor, analyze, and debug LLM evaluations via nemo-evaluator-launcher. Covers running evaluations, checking status and live progress, debugging failed runs, exporting artifacts and logs, and analyzing results. ALWAYS triggers on mentions of running…

原文语言:英语

更新
职业分类
网络与计算机系统管理员
描述

Monitor submitted jobs (PTQ, evaluation, deployment) on SLURM clusters. Use when the user asks "check job status", "is my job done", "monitor my evaluation", "what's the status of the PTQ", "check on a SLURM job id", or after any skill submits a long-running…

原文语言:英语

更新
职业分类
软件开发工程师
描述

This skill should be used when the user asks to "quantize a model", "run PTQ", "post-training quantization", "NVFP4 quantization", "FP8 quantization", "INT8 quantization", "INT4 AWQ", "quantize LLM", "quantize MoE", "quantize VLM", or needs to produce a…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Cherry-pick merged PRs labeled for a release branch into that branch, then open a PR and apply the cherry-pick-done label. Use when asked to "cherry-pick PRs for release/X.Y.Z", "pick PRs to release branch", or "cherry-pick labeled PRs".

原文语言:英语

更新
职业分类
软件开发工程师
描述

Create custom LLM evaluation benchmarks using the BYOB decorator framework. Use when the user wants to (1) create a new benchmark from a dataset, (2) pick or write a scorer, (3) compile and run a BYOB benchmark, (4) containerize a benchmark, or (5) use…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Query and browse evaluation results stored in MLflow. Use when the user wants to look up runs by invocation ID, compare metrics across models, fetch artifacts (configs, logs, results), or set up the MLflow MCP server. ALWAYS triggers on mentions of MLflow,…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Run, monitor, analyze, and debug LLM evaluations via nemo-evaluator-launcher. Covers running evaluations, checking status and live progress, debugging failed runs, exporting artifacts and logs, and analyzing results. ALWAYS triggers on mentions of running…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Interactive config wizard for NeMo Evaluator Launcher (NEL). Use when the user wants to create a new evaluation config from scratch, set up an evaluation from existing configs, or modify a NEL config (deployment, tasks, multi-node, interceptors). ALWAYS…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Guide for adding a new benchmark or training environment to NeMo-Gym. Use when the user asks to add, create, or integrate a benchmark, evaluation, training environment, or resources server into NeMo-Gym. Also use when wrapping an existing 3rd-party benchmark…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Use when debugging a Nemo Gym run or reward profiling job. Covers rollout collection failures, empty or partial JSONL outputs, stale materialized inputs, verifier/schema errors, Ray or Slurm issues, vLLM readiness, judge failures, tool/sandbox failures, cache…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Maintain the NeMo Gym Fern docs site — add, update, move, or remove pages under fern/. Use for any documentation change. Triggered by: "edit docs", "add doc page", "update docs", "rename page", "fix broken link", "add redirect", "preview docs", "publish…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Use when creating, validating, or documenting Nemo Gym pivot datasets from rollout, trajectory, chat-completion, Responses API, or tool-call artifacts. Covers Gym Responses-style row conversion, pivot selection, single-step tool-use configs, agent_ref…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Use to help users get started with Nemo Gym reward profiling. Covers the basic ng_run, ng_collect_rollouts, and ng_reward_profile workflow, repeated rollouts, materialized inputs, rollout JSONL artifacts, task and rollout identity, output inspection, partial…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Autonomous NeMo-RL research agent workflow for directed hypothesis testing and open-ended discovery. Guides agents through the full experiment lifecycle: understanding recipes and environments, wiring RL or NeMo-gym runs, launching reproducible baselines and…

原文语言:英语

更新
职业分类
网络与计算机系统管理员
描述

Brev instance operating guidance for NeMo-RL agents working in /home/ubuntu/RL with limited workspace disk, a larger /ephemeral volume, and optional /home/ubuntu/RL/.env secrets. Use when running auto-research campaigns, experiments, training jobs, model or…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Build and dependency management for NeMo-RL. Covers Docker image building and running, uv usage, venv setup, and adding dependencies.

原文语言:英语

更新
职业分类
软件开发工程师
描述

CI/CD reference for NeMo-RL. Covers GitHub Actions pipeline structure, CI triggering via /ok to test, and CI failure investigation.

原文语言:英语

更新
职业分类
软件开发工程师
描述

Configuration conventions for NeMo-RL. YAML is the single source of truth for defaults. Covers TypedDict usage, exemplar YAML updates, and forbidden default patterns.

原文语言:英语

更新
职业分类
软件开发工程师
描述

Contribution conventions for NeMo-RL. Covers PR title format, commit sign-off, and CI triggering.

原文语言:英语

更新
职业分类
软件开发工程师
描述

NVIDIA copyright header requirements for NeMo-RL. Covers which files need headers and the exact header text.

原文语言:英语

更新
职业分类
软件开发工程师
描述

Documentation conventions for NeMo-RL. Covers docs/index.md updates and docstring format.

原文语言:英语

更新
职业分类
软件开发工程师
描述

Error handling guidelines for NeMo-RL. Covers exception specificity, minimal try bodies, and else blocks.

原文语言:英语

更新
职业分类
网络与计算机系统管理员
描述

Playbook for launching, monitoring, stopping, and debugging NeMo-RL recipes on a Kubernetes cluster via the nrl-k8s CLI. Covers ephemeral vs long-lived RayCluster modes, iterating on runs, and debugging hung or failed training jobs.

原文语言:英语

更新
职业分类
软件开发工程师
描述

Code style guidelines for NeMo-RL (Python and shell). Covers naming, indentation, comments, docstrings, reflection avoidance, and uv usage.

原文语言:英语

更新
职业分类
软件质量保证分析师与测试员
描述

Interactive code review for NVIDIA-NeMo/RL pull requests. Checks out PR locally, reads existing comments, applies coding guidelines from skills, previews findings, and posts review comments. Also supports reviewing the current branch locally.

原文语言:英语

更新
职业分类
其他计算机职业
描述

Manage durable working-session memory for coding agents. Use when a user asks to preserve or recover agent context across disconnects, VS Code restarts, long-running work, handoffs, or any session where important state should be written periodically under the…

原文语言:英语

更新
职业分类
软件质量保证分析师与测试员
描述

Testing conventions for NeMo-RL. Covers Ray actor coverage pragmas, nightly test requirements, and recipe naming rules.

原文语言:英语

更新
职业分类
软件开发工程师
描述

Create GitHub pull requests that follow the NemoClaw PR template. Use when the user wants to create a new PR, submit code for review, open a pull request, or push changes for review. Trigger keywords - create PR, pull request, new PR, submit for review, open…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Scan recent git commits for changes that affect user-facing behavior, then draft or update the corresponding documentation pages and refresh generated user skills for release prep. Use when docs have fallen behind code changes, after a batch of features…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Scans other open issues to find ones a given PR may also fix or accidentally break. Outputs adjacent-fix opportunities and contradiction risks with file:line evidence. Use when reviewing a PR to discover bundling opportunities or downstream impact across the…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Cut a new semver release — bump all version strings via bump-version.ts, open a release PR, and after merge tag main and push. Use when cutting a release, tagging a version, shipping a build, or preparing a deployment. Trigger keywords - cut tag, release tag,…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Runs the daytime maintainer loop for NemoClaw, prioritizing items labeled with the current version target. Picks the highest-value item, executes the right workflow (merge gate, salvage, security sweep, test gaps, hotspot cooling, or sequencing), and reports…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Runs the end-of-day maintainer handoff for NemoClaw. Checks version target progress, bumps stragglers to the next patch version, generates a QA handoff summary, and cuts the release tag. Use at the end of the workday. Trigger keywords - evening, end of day,…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Finds open GitHub PRs with security and priority-high labels, links each to its issue, detects duplicates (multiple PRs fixing the same issue), and presents a table of review candidates. Use when looking for the next PR to review. Trigger keywords - find pr,…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Runs the morning maintainer standup for NemoClaw. Triages the backlog, determines the day's target version, labels selected items, surfaces stragglers from previous versions, and outputs the daily plan. Use at the start of the workday. Trigger keywords -…

原文语言:英语

更新
已展示 40 / 1,040 个已收集 Skill。