Skip to main content
Jeden Skill in Manus ausführen
mit einem Klick

flinch-probe

Sterne1
Forks0
Aktualisiert23. April 2026 um 12:34

Measures the "flinch" of a language model — the gap between the log-probability a charged word deserves on pure fluency grounds and the probability the model actually assigns it. Use this skill whenever the user wants to: measure token suppression in an LLM, compare pretrain corpora for word-level bias, audit "uncensored" models for hidden censorship, reproduce or extend the morgin.ai flinch methodology, benchmark a model on the anti-china/anti-america/anti-europe/slurs/sexual/violence axes, or generate a flinch radar chart. Trigger for phrases like: "measure flinch", "how suppressed is X in model Y", "does this model avoid the word", "token probability audit", "check for hidden censorship", "compare base vs ablated model", "flinch score", "run the flinch probe", or any request to quantify how much a model deflates specific vocabulary.

Installation

Mit Codex oder Claude installieren Kopieren Sie diesen Prompt, fügen Sie ihn in Codex, Claude oder einen anderen Assistant ein und lassen Sie die Skill-Seite prüfen und installieren.

Datei-Explorer
5 Dateien
SKILL.md
readonly