Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة
$pwd:

kernel-tileir-optimization

// Optimize existing Triton kernels for NVIDIA TileIR backend on Blackwell GPUs (sm_100+). Adds TileIR-specific autotune configs: occupancy, num_ctas, TMA descriptors. Covers kernel classification (dot-related, norm-like, elementwise, reduction), type-specific transformations, and PTX-vs-TileIR benchmarking. Triggered by: "optimize for TileIR", "add TileIR configs", "Blackwell optimization", "TMA descriptors", "2CTA mode", "occupancy tuning". Kernels use standard `import triton`; TileIR activates via ENABLE_TILE=1 when nvtriton is installed.

$ git log --oneline --stat
stars:١٣٬٧٠٢
forks:٢٬٤٠٦
updated:٢٠ مايو ٢٠٢٦ في ٠٧:٣٥
مستكشف الملفات
5 ملفات
SKILL.md
readonly