원클릭으로
guardrail-monitor
The system's value alignment and truth verification layer, ensuring output safety and epistemic integrity.
Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.
메뉴
The system's value alignment and truth verification layer, ensuring output safety and epistemic integrity.
Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.
SOC 직업 분류 기준
A project and task management system for tracking work across the Abraxas system.
The Sovereign Vault for verified truth fragments and historical provenance tracking.
Soter is the Risk Sensing & Triggering pillar of the Sovereign Brain. It monitors for linguistic sycophancy and mechanistic attention sinks to trigger an Epistemic Crisis (T=1) and assign a Risk Score (R[0-5]).
The core governance and system integrity layer of the Abraxas sovereign environment.
The ingestion and filtration loop for external data entering the Abraxas system.
The Void Mapper for identifying and resolving epistemic gaps.
| name | guardrail-monitor |
| description | The system's value alignment and truth verification layer, ensuring output safety and epistemic integrity. |
The Guardrail Monitor is a tripartite engine designed to prevent AI drift and ensure that outputs align with both user values and empirical ground truth. It consists of three specialized modules: Pathos, Pheme, and Kratos.
Pathos monitors the "emotional and ethical salience" of the interaction. It identifies what the user cares about and flags when a proposed output might conflict with those priorities.
Pheme provides a mechanism for verifying claims against trusted external sources.
VERIFIED (sufficient agreement), CONTRADICTED (high-reliability contradiction), or UNVERIFIABLE.Kratos resolves contradictions between conflicting sources using a hierarchical authority model.
check_value_saliencyAnalyzes a topic or decision context against tracked user values.
topic (required), decision_context (optional), user_values (optional).verify_ground_truthVerifies a claim using a list of sources.
claim (required), sources (optional), require_min_sources (optional).arbitrate_conflictResolves a conflict between two competing claims.
claimA, claimB, sourceA, sourceB (required), domain (optional).