nsys-timeline
Use for stream overlap, launch overhead, and end-to-end CUDA timeline analysis with Nsight Systems.
Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。
メニュー
Use for stream overlap, launch overhead, and end-to-end CUDA timeline analysis with Nsight Systems.
Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。
SOC 職業分類に基づく
Use when creating or revising CUDA benchmark runners and result artifacts.
Use when CUDA availability, compiler setup, or profiler visibility may be broken or inconsistent.
Use when designing or reviewing handwritten CUDA FlashAttention kernels.
Use when designing or reviewing handwritten CUDA GEMM kernels.
Use for kernel-level analysis with Nsight Compute after a benchmark is reproducible.
Use when deciding whether a CUDA kernel is likely memory-bound or compute-bound.
| name | nsys-timeline |
| description | Use for stream overlap, launch overhead, and end-to-end CUDA timeline analysis with Nsight Systems. |
Use this skill for stream overlap, launch overhead, and end-to-end timeline analysis with Nsight Systems.
scripts/perf/profile-nsys.ps1.artifacts/profiles/.