Skip to main content
Manus에서 모든 스킬 실행
원클릭으로

flinch-probe

스타1
포크0
업데이트2026년 4월 23일 12:34

Measures the "flinch" of a language model — the gap between the log-probability a charged word deserves on pure fluency grounds and the probability the model actually assigns it. Use this skill whenever the user wants to: measure token suppression in an LLM, compare pretrain corpora for word-level bias, audit "uncensored" models for hidden censorship, reproduce or extend the morgin.ai flinch methodology, benchmark a model on the anti-china/anti-america/anti-europe/slurs/sexual/violence axes, or generate a flinch radar chart. Trigger for phrases like: "measure flinch", "how suppressed is X in model Y", "does this model avoid the word", "token probability audit", "check for hidden censorship", "compare base vs ablated model", "flinch score", "run the flinch probe", or any request to quantify how much a model deflates specific vocabulary.

설치

Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.

파일 탐색기
5 개 파일
SKILL.md
readonly