Skip to main content

refute-your-own-answer-blind

When an answer arrives suspiciously clean — especially on a high-stakes claim, a safety judgment, or anything your own confidence is trying to launder into truth — strip authorship and confidence, pay an adversary to destroy it, and require the attacks to fail on execution rather than on counter-prose. The failure it prevents: self-preference and plausibility bias, where the model protects its own first answer because it sounds coherent and every review quietly reuses that same coherence. Triggers on "this should work", policy/safety claims, migration plans, root-cause explanations, and answers whose per-step confidence is uniformly high.

الانتقال إلى التثبيت

معلومات المصدر

المستودع
XyraSinclair/ideonomy
آخر نشاط في المصدر
٩ يوليو ٢٠٢٦ في ١٠:٤٦
لغة SKILL.md المكتشفة
الإنجليزية
النجوم
٠
التفرعات
٠

خيارات التثبيت

يُحدَّد Prompt الذي يراجع المصدر أولًا بشكل افتراضي. يمكنك التبديل إلى أمر مباشر أو تنزيل نسخة محلية.

مراجعة ملفات المصدر

اقرأ SKILL.md وأي ملفات مرافقة يعرضها SkillsMP قبل أن تقرر التثبيت.