一键导入
judge-evaluate-policy-defense-loop
Algorithms for LLM-as-a-Judge telemetry and rule synthesis.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
菜单
Algorithms for LLM-as-a-Judge telemetry and rule synthesis.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
基于 SOC 职业分类
Instructions for triggering external Context7 connections
Lightweight research-plan-implement workflow for small tweaks, isolated bugfixes, copy edits, and low-blast-radius polish in this repo. Use when the change can be understood from 1-3 files and stays inside one subsystem without changing shared data flow.
Implements tree-based attack branching and prompt refinement optimization.
Fetch external library or framework documentation only when local docs are insufficient.
Guidance for JudgeAgent scoring, guardrail synthesis, and anti-hallucination validation work.
Full research-plan-implement workflow for new features, architectural updates, multi-file changes, or any task with non-trivial blast radius in RedThread.
| name | Judge Evaluate & Policy Defense Loop |
| description | Algorithms for LLM-as-a-Judge telemetry and rule synthesis. |
When assessing an AttackRunner's output against a target, building the Prometheus JudgeAgent framework, or configuring the self-healing telemetry pipeline.
Score_final.Auto-CoT process prior to the verdict so the agent generates its own grading rubric step-by-step.K Core-Distance.