Skip to main content

Skills neste repositório

autohandai/community-skills - Página 3

O SkillsMP coletou 1.040 skills de autohandai/community-skills. Abra uma skill para revisar a origem e os detalhes.

autohandai/community-skills

Mostrando 40 de 1.040 skills coletadas.

ocupação
Analistas de garantia de qualidade de software e testadores
descrição

Refresh golden values from a GitHub Actions workflow run (failing-only or all jobs), score the change with average normalized relative differences, and produce a PR-ready summary. Use when the user asks to update goldens for a CI run, refresh golden values…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Query and browse evaluation results stored in MLflow. Use when the user wants to look up runs by invocation ID, compare metrics across models, fetch artifacts (configs, logs, results), or set up the MLflow MCP server. ALWAYS triggers on mentions of MLflow,…

Idioma do texto original: inglês

atualizado
ocupação
Administradores de redes e sistemas de computador
descrição

Run commands inside a remote Docker container via the file-based command relay (tools/debugger). Use when the user says "run in Docker", "run on GPU", "debug remotely", "run test in container", "check nvidia-smi", "run pytest in Docker", or needs to execute…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Serve a quantized or unquantized LLM checkpoint as an OpenAI-compatible API endpoint using vLLM, SGLang, or TRT-LLM. Use when user says "deploy model", "serve model", "start vLLM server", "launch SGLang", "TRT-LLM deploy", "AutoDeploy", "benchmark…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Evaluates accuracy of quantized or unquantized LLMs using NeMo Evaluator Launcher (NEL). Triggers on "evaluate model", "benchmark accuracy", "run MMLU", "evaluate quantized model", "accuracy drop", "run nel". Handles deployment, config generation, and…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Run, monitor, analyze, and debug LLM evaluations via nemo-evaluator-launcher. Covers running evaluations, checking status and live progress, debugging failed runs, exporting artifacts and logs, and analyzing results. ALWAYS triggers on mentions of running…

Idioma do texto original: inglês

atualizado
ocupação
Administradores de redes e sistemas de computador
descrição

Monitor submitted jobs (PTQ, evaluation, deployment) on SLURM clusters. Use when the user asks "check job status", "is my job done", "monitor my evaluation", "what's the status of the PTQ", "check on a SLURM job id", or after any skill submits a long-running…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

This skill should be used when the user asks to "quantize a model", "run PTQ", "post-training quantization", "NVFP4 quantization", "FP8 quantization", "INT8 quantization", "INT4 AWQ", "quantize LLM", "quantize MoE", "quantize VLM", or needs to produce a…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Cherry-pick merged PRs labeled for a release branch into that branch, then open a PR and apply the cherry-pick-done label. Use when asked to "cherry-pick PRs for release/X.Y.Z", "pick PRs to release branch", or "cherry-pick labeled PRs".

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Create custom LLM evaluation benchmarks using the BYOB decorator framework. Use when the user wants to (1) create a new benchmark from a dataset, (2) pick or write a scorer, (3) compile and run a BYOB benchmark, (4) containerize a benchmark, or (5) use…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Query and browse evaluation results stored in MLflow. Use when the user wants to look up runs by invocation ID, compare metrics across models, fetch artifacts (configs, logs, results), or set up the MLflow MCP server. ALWAYS triggers on mentions of MLflow,…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Run, monitor, analyze, and debug LLM evaluations via nemo-evaluator-launcher. Covers running evaluations, checking status and live progress, debugging failed runs, exporting artifacts and logs, and analyzing results. ALWAYS triggers on mentions of running…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Interactive config wizard for NeMo Evaluator Launcher (NEL). Use when the user wants to create a new evaluation config from scratch, set up an evaluation from existing configs, or modify a NEL config (deployment, tasks, multi-node, interceptors). ALWAYS…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Guide for adding a new benchmark or training environment to NeMo-Gym. Use when the user asks to add, create, or integrate a benchmark, evaluation, training environment, or resources server into NeMo-Gym. Also use when wrapping an existing 3rd-party benchmark…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Use when debugging a Nemo Gym run or reward profiling job. Covers rollout collection failures, empty or partial JSONL outputs, stale materialized inputs, verifier/schema errors, Ray or Slurm issues, vLLM readiness, judge failures, tool/sandbox failures, cache…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Maintain the NeMo Gym Fern docs site — add, update, move, or remove pages under fern/. Use for any documentation change. Triggered by: "edit docs", "add doc page", "update docs", "rename page", "fix broken link", "add redirect", "preview docs", "publish…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Use when creating, validating, or documenting Nemo Gym pivot datasets from rollout, trajectory, chat-completion, Responses API, or tool-call artifacts. Covers Gym Responses-style row conversion, pivot selection, single-step tool-use configs, agent_ref…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Use to help users get started with Nemo Gym reward profiling. Covers the basic ng_run, ng_collect_rollouts, and ng_reward_profile workflow, repeated rollouts, materialized inputs, rollout JSONL artifacts, task and rollout identity, output inspection, partial…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Autonomous NeMo-RL research agent workflow for directed hypothesis testing and open-ended discovery. Guides agents through the full experiment lifecycle: understanding recipes and environments, wiring RL or NeMo-gym runs, launching reproducible baselines and…

Idioma do texto original: inglês

atualizado
ocupação
Administradores de redes e sistemas de computador
descrição

Brev instance operating guidance for NeMo-RL agents working in /home/ubuntu/RL with limited workspace disk, a larger /ephemeral volume, and optional /home/ubuntu/RL/.env secrets. Use when running auto-research campaigns, experiments, training jobs, model or…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Build and dependency management for NeMo-RL. Covers Docker image building and running, uv usage, venv setup, and adding dependencies.

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

CI/CD reference for NeMo-RL. Covers GitHub Actions pipeline structure, CI triggering via /ok to test, and CI failure investigation.

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Configuration conventions for NeMo-RL. YAML is the single source of truth for defaults. Covers TypedDict usage, exemplar YAML updates, and forbidden default patterns.

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Contribution conventions for NeMo-RL. Covers PR title format, commit sign-off, and CI triggering.

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

NVIDIA copyright header requirements for NeMo-RL. Covers which files need headers and the exact header text.

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Documentation conventions for NeMo-RL. Covers docs/index.md updates and docstring format.

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Error handling guidelines for NeMo-RL. Covers exception specificity, minimal try bodies, and else blocks.

Idioma do texto original: inglês

atualizado
ocupação
Administradores de redes e sistemas de computador
descrição

Playbook for launching, monitoring, stopping, and debugging NeMo-RL recipes on a Kubernetes cluster via the nrl-k8s CLI. Covers ephemeral vs long-lived RayCluster modes, iterating on runs, and debugging hung or failed training jobs.

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Code style guidelines for NeMo-RL (Python and shell). Covers naming, indentation, comments, docstrings, reflection avoidance, and uv usage.

Idioma do texto original: inglês

atualizado
ocupação
Analistas de garantia de qualidade de software e testadores
descrição

Interactive code review for NVIDIA-NeMo/RL pull requests. Checks out PR locally, reads existing comments, applies coding guidelines from skills, previews findings, and posts review comments. Also supports reviewing the current branch locally.

Idioma do texto original: inglês

atualizado
ocupação
Outras ocupações de informática
descrição

Manage durable working-session memory for coding agents. Use when a user asks to preserve or recover agent context across disconnects, VS Code restarts, long-running work, handoffs, or any session where important state should be written periodically under the…

Idioma do texto original: inglês

atualizado
ocupação
Analistas de garantia de qualidade de software e testadores
descrição

Testing conventions for NeMo-RL. Covers Ray actor coverage pragmas, nightly test requirements, and recipe naming rules.

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Create GitHub pull requests that follow the NemoClaw PR template. Use when the user wants to create a new PR, submit code for review, open a pull request, or push changes for review. Trigger keywords - create PR, pull request, new PR, submit for review, open…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Scan recent git commits for changes that affect user-facing behavior, then draft or update the corresponding documentation pages and refresh generated user skills for release prep. Use when docs have fallen behind code changes, after a batch of features…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Scans other open issues to find ones a given PR may also fix or accidentally break. Outputs adjacent-fix opportunities and contradiction risks with file:line evidence. Use when reviewing a PR to discover bundling opportunities or downstream impact across the…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Cut a new semver release — bump all version strings via bump-version.ts, open a release PR, and after merge tag main and push. Use when cutting a release, tagging a version, shipping a build, or preparing a deployment. Trigger keywords - cut tag, release tag,…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Runs the daytime maintainer loop for NemoClaw, prioritizing items labeled with the current version target. Picks the highest-value item, executes the right workflow (merge gate, salvage, security sweep, test gaps, hotspot cooling, or sequencing), and reports…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Runs the end-of-day maintainer handoff for NemoClaw. Checks version target progress, bumps stragglers to the next patch version, generates a QA handoff summary, and cuts the release tag. Use at the end of the workday. Trigger keywords - evening, end of day,…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Finds open GitHub PRs with security and priority-high labels, links each to its issue, detects duplicates (multiple PRs fixing the same issue), and presents a table of review candidates. Use when looking for the next PR to review. Trigger keywords - find pr,…

Idioma do texto original: inglês

atualizado
ocupação
Desenvolvedores de software
descrição

Runs the morning maintainer standup for NemoClaw. Triages the backlog, determines the day's target version, labels selected items, surfaces stragglers from previous versions, and outputs the daily plan. Use at the start of the workday. Trigger keywords -…

Idioma do texto original: inglês

atualizado
Mostrando 40 de 1.040 skills coletadas.