원클릭으로
judge-evaluate-policy-defense-loop
Algorithms for LLM-as-a-Judge telemetry and rule synthesis.
Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.
메뉴
Algorithms for LLM-as-a-Judge telemetry and rule synthesis.
Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.
SOC 직업 분류 기준
Instructions for triggering external Context7 connections
Lightweight research-plan-implement workflow for small tweaks, isolated bugfixes, copy edits, and low-blast-radius polish in this repo. Use when the change can be understood from 1-3 files and stays inside one subsystem without changing shared data flow.
Implements tree-based attack branching and prompt refinement optimization.
Fetch external library or framework documentation only when local docs are insufficient.
Guidance for JudgeAgent scoring, guardrail synthesis, and anti-hallucination validation work.
Full research-plan-implement workflow for new features, architectural updates, multi-file changes, or any task with non-trivial blast radius in RedThread.
| name | Judge Evaluate & Policy Defense Loop |
| description | Algorithms for LLM-as-a-Judge telemetry and rule synthesis. |
When assessing an AttackRunner's output against a target, building the Prometheus JudgeAgent framework, or configuring the self-healing telemetry pipeline.
Score_final.Auto-CoT process prior to the verdict so the agent generates its own grading rubric step-by-step.K Core-Distance.