Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

flash-attention-kernel

النجوم٨
التفرعات١
آخر تحديث١٥ مايو ٢٠٢٦ في ٢٣:٣٩

Optimize FlashAttention-style fused attention kernels in Triton for NVIDIA and AMD GPUs. Covers online softmax, tiled QK/AV GEMM, causal masking, and memory-efficient attention. Use when writing or optimizing self-attention, cross-attention, or any QKV attention kernel.

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

مستكشف الملفات
3 ملفات
SKILL.md
readonly