المهنة
علماء البيانات
الوصف
This skill should be used when users want to fine-tune language models or perform reinforcement learning (SFT, DPO, GRPO, ORPO, KTO, SimPO) using the highly optimized Unsloth library. Covers environment setup, LoRA patching, VRAM optimization,…
لغة النص الأصلي: الإنجليزية
آخر تحديث