ワンクリックで
judge-evaluate-policy-defense-loop
Algorithms for LLM-as-a-Judge telemetry and rule synthesis.
Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。
メニュー
Algorithms for LLM-as-a-Judge telemetry and rule synthesis.
Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。
SOC 職業分類に基づく
Instructions for triggering external Context7 connections
Lightweight research-plan-implement workflow for small tweaks, isolated bugfixes, copy edits, and low-blast-radius polish in this repo. Use when the change can be understood from 1-3 files and stays inside one subsystem without changing shared data flow.
Implements tree-based attack branching and prompt refinement optimization.
Fetch external library or framework documentation only when local docs are insufficient.
Guidance for JudgeAgent scoring, guardrail synthesis, and anti-hallucination validation work.
Full research-plan-implement workflow for new features, architectural updates, multi-file changes, or any task with non-trivial blast radius in RedThread.
| name | Judge Evaluate & Policy Defense Loop |
| description | Algorithms for LLM-as-a-Judge telemetry and rule synthesis. |
When assessing an AttackRunner's output against a target, building the Prometheus JudgeAgent framework, or configuring the self-healing telemetry pipeline.
Score_final.Auto-CoT process prior to the verdict so the agent generates its own grading rubric step-by-step.K Core-Distance.