Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

offpolicy-algorithms

النجوم١٨
التفرعات٤
آخر تحديث٢٢ يوليو ٢٠٢٦ في ١٧:١١

Implement, extend, and run off-policy RL algorithms in active-adaptation (SAC, distributional SAC, SimbaV2, RLPD, BAC/BEE). Use when adding or modifying algo files under learning/offpolicy, wiring Hydra algo configs, debugging train_offpolicy.py runs, replay buffers, or critic/actor training loops.

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

مستكشف الملفات
2 ملفات
SKILL.md
readonly