Skip to main content

global-workspace-j-space

LLM interpretability methodology from Anthropic's "A global workspace in language models" (Jul 2026). Use when probing what a language model is "thinking but not saying" — its consciously-accessible / broadcast internal representations — via the Jacobian lens (J-lens). Covers finding J-space patterns, reading them as silent words, and using them to catch hidden goals, deception, or tests. Open-source implementation released by Anthropic.

Aller à l'installation

Informations de source

Dépôt
hiyenwong/ai_collection
Dernière activité de la source
14 juillet 2026 à 01:05
Langue détectée de SKILL.md
anglais
Étoiles
2
Forks
0

Options d'installation

Le prompt qui vérifie d'abord la source est sélectionné par défaut. Vous pouvez passer à une commande directe ou télécharger une copie locale.

Vérifiez les fichiers source

Lisez SKILL.md et les fichiers associés affichés par SkillsMP avant de décider de l'installer.