Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

research-flash-attention

النجوم١
التفرعات٠
آخر تحديث١٣ يوليو ٢٠٢٦ في ٠٩:٢٨

Optimizes transformer attention with Flash Attention for 2-4x speedup and 10-20x memory reduction. Use when training/running transformers with long sequences (>512 tokens), encountering GPU memory ...

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

SKILL.md
readonly